Чем предстоит заниматься
Build and maintain ETL workflows
Collaborate with data stakeholders
Conduct production issue investigation and root cause analysis
Create technical documentation for data mappings and runbooks
Design ETL data pipelines
Develop batch scheduling with Control-M
Ensure data quality auditability and data lineage
Implement stability and resilience fixes
Monitor and automate workloads
Optimize PySpark data processing
Perform Linux automation and troubleshooting