Tech Mahindra

Tech Lead

Tech Mahindra$120K — $150K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 5-7 years of experience with IBM InfoSphere DataStage.
  • Strong understanding of automated ETL testing tools and methodologies.
  • Proficiency in SQL, Python, and XML/JSON parsing.
  • Experience with Databricks and PySpark environments.
  • Ability to analyze complex data lineage and operational metadata.

Responsibilities

  • Analyze DataStage exports and metadata for data lineage mapping.
  • Identify and remove redundant code and 'dead jobs' during migration.
  • Explain complex job functions to the PySpark development team.
  • Implement automated testing frameworks for data reconciliation.
  • Create scripts for validating data across dual-run setups.
  • Conduct regression testing on refactored PySpark code.
  • Document and sign off on data validation KPIs prior to migrations.

Benefits

  • Medical, vision, and dental insurance.
  • Life and disability insurance.
  • Paid time off including holidays and parental leave.
  • Sick leave as required by law.
Full Job Description
Job Summary

Job Title: Senior DataStage Developer

Location: Across USA any Location

Objective
Drive the end-to-end rationalization, reverse-engineering, and automated validation of legacy DataStage environments migrating to modern Databricks architectures. This role is critical to eliminating legacy code redundancies, mapping complex data lineage, and implementing automated testing frameworks to guarantee zero data loss and business disruption during system cutovers.

Key Responsibilities
  • Legacy Rationalization: Analyze DataStage .dsx/.isx exports and metadata to map end-to-end data lineage.
  • Code Elimination: Identify and isolate "dead jobs", redundant code, and duplicate logic to streamline migration waves.
  • Technical Translation: Provide functional logic explanations of complex parallel/server jobs to the PySpark development team.
  • Test Automation: Deploy automated test frameworks to execute large-scale data reconciliation between DataStage and Databricks.
  • Data Validation: Build automated scripts to validate data schemas, row counts, and complex transformations across dual-run environments.
  • Regression Testing: Execute regression testing on newly refactored PySpark code against historical legacy data.
  • Migration Sign-off: Document validation execution KPIs and formally sign off on data accuracy before live migration cutovers.

Technical Skills & Competencies
  • Legacy ETL: IBM InfoSphere DataStage (Parallel/Server jobs, Sequences) and operational metadata analysis.
  • Data Quality & Testing: Automated ETL Testing tools, PyTest, and Great Expectations.
  • Languages & Querying: Advanced SQL, Python, and XML/JSON parsing.
  • Target Platforms: Familiarity with Databricks, PySpark, and modern cloud data warehouses.

The pay range for this role is $120k - $150k per annum including any bonuses or variable pay. Tech Mahindra also offers benefits like medical, vision, dental, life, disability insurance and paid time off (including holidays, parental leave, and sick leave, as required by law). Ask our recruiters for more details on our Benefits package. The exact offer terms will depend on the skill level, educational qualifications, experience, and location of the candidate.

Similar Jobs

More Jobs at Tech Mahindra

More Information Technology Jobs

Find similar Tech Lead jobs: