N

Senior ETL Modernization Developer (IBM DataStage to Databricks/Spark)

Numentica United State
Remote
Apply
AI Summary

Lead the migration of legacy IBM DataStage ETL workloads to modern Databricks/Spark on AWS, designing scalable cloud data pipelines with PySpark, Python, and SQL. Focus on data validation, CI/CD automation, and seamless production cutover while decommissioning outdated systems. Requires deep expertise in ETL modernization, Delta Lake, and AWS data services.

Key Highlights
Migrate legacy IBM DataStage ETL/ELT jobs to Databricks/Spark on AWS
Build scalable cloud data pipelines using PySpark, Python, and SQL
Implement CI/CD pipelines and automated ETL testing with PyTest
Key Responsibilities
Assess and migrate legacy DataStage ETL/ELT jobs to Databricks/Spark on AWS
Develop scalable cloud data pipelines using PySpark, Python, and SQL
Perform data validation, parity checks, UAT, regression, and performance testing
Implement CI/CD pipelines and automated ETL testing frameworks
Support production cutover, deployment, stabilization, and hypercare
Decommission legacy DataStage processes post-migration
Prepare technical documentation and conduct knowledge transfer
Technical Skills Required
IBM DataStage PySpark AWS
Benefits & Perks
W2 payroll (no contractor-to-contractor)
Remote work eligibility

Job Description


ETL Modernization Developer


Location: Remote

Duration: 12 Months Contract

Payrate on: W2 (No C2C)

Work Authorization: GC / USC Only


Job Summary


Seeking an experienced ETL Modernization Developer to migrate and modernize legacy IBM DataStage workloads to Databricks/Spark on AWS. The ideal candidate should have strong hands-on experience in ETL modernization, PySpark/Python, SQL, AWS data services, Delta Lake, and CI/CD.


Must-Have Skills


  • IBM DataStage – strong hands-on development
  • ETL/ELT Migration & Modernization – legacy-to-cloud migration
  • Databricks + Apache Spark on AWS
  • PySpark, Python, SQL & Shell scripting
  • Delta Lake, Unity Catalog & Photon
  • AWS Glue, Lambda & Redshift
  • CI/CD + Automated ETL Testing
  • PyTest + XML/JSON parsing


Key Responsibilities


  • Assess and migrate legacy DataStage ETL/ELT jobs to Databricks/Spark.
  • Build scalable cloud data pipelines using PySpark/Python/SQL.
  • Perform data validation, parity checks, UAT, regression and performance testing.
  • Implement CI/CD and automated ETL testing.
  • Support production cutover, deployment, stabilization and hypercare.
  • Decommission legacy DataStage processes after successful migration.
  • Prepare technical documentation and conduct knowledge transfer.



Similar Jobs

Explore other opportunities that match your interests

Senior Golang Engineer

Programming
9h ago
Visa Sponsorship Relocation Remote
Job Type Contract
Experience Level Mid-Senior level

anblicks

United State
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Not Applicable

makai labs

United State

Head of People

Programming
19h ago
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Not Applicable

startx med

United State

Subscribe our newsletter

New Things Will Always Update Regularly