Tata Consultancy Services

Hadoop,Pyspark,Hive,Kafka - Senior Developer

Tata Consultancy Services$110K — $125K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 10+ years of hands-on experience in Big Data engineering.
  • Expertise in PySpark, Kafka, and Hadoop ecosystem components.
  • Extensive experience with Databricks Lakehouse architecture and Delta Lake optimizations.
  • Strong SQL skills for handling massive datasets (TB/PB scale).
  • Experience in Global Markets, Risk, Treasury, or Regulatory Reporting.

Responsibilities

  • Design, develop, and optimize PySpark-based ETL pipelines for on-prem and cloud environments.
  • Build high-volume Kafka ingestion frameworks for real-time market data.
  • Develop and manage Hadoop ecosystem components like HDFS and YARN.
  • Implement Bronze/Silver/Gold data modeling patterns in Databricks Lakehouse.
  • Apply Delta Lake best practices for optimized file management and data handling.
  • Collaborate with cross-functional teams in agile ceremonies and design reviews.
  • Mentor junior engineers and enhance engineering practices within the team.

Benefits

  • Discretionary Annual Incentive.
  • Comprehensive medical coverage, including dental and vision insurance.
  • Family support through maternal and parental leaves.
  • Convenience benefits for commuting and professional training reimbursements.
  • Generous time off policy, including vacation and sick leave.
  • Legal and financial assistance programs, including a 401K plan and student loan refinancing.
Full Job Description
Must Have Technical/Functional Skills

Primary skills: PySpark, Apache Kafka, Hadoop Ecosystem, Hive, Databricks Lakehouse Architecture, Delta Lake, Bronze/Silver/Gold Data Modeling, Big Data ETL Pipeline Development, SQL, Real-time Data Ingestion Frameworks, Data Governance & Cataloging, CI/CD Tools Git, Jenkins, Bitbucket, Workflow Orchestration, and Cloud & On-Prem Big Data Platforms.

Experience: Minimum 10+ years

Roles & Responsibilities

Seeking a Senior Big Data Engineer with 1013 years of experience specializing in Hadoop, PySpark, Kafka, Hive, and strong experience designing data solutions for large-scale financial systems.

In addition, the candidate must possess advanced expertise in Databricks Lakehouse architecture, particularly around Bronze/Silver/Gold layer data modeling, Delta Lake optimizations, and building reliable, scalable pipelines for regulatory, risk, trading, and analytics workloads.

This role focuses on delivering highly performant, well-governed data platforms that support the banks mission-critical global markets functions.

Key Responsibilities:

Big Data Platform Engineering
• Design, develop, and optimize PySpark-based ETL pipelines running on on-prem Hadoop clusters and cloud environments.
• Build high-volume ingestion frameworks using Kafka for real-time and near-real-time trading and market data.
• Develop, tune, and manage Hadoop ecosystem componentsHDFS, YARN, MapReduce, Tez, Oozie/Airflow.
• Build high-performance, optimized Hive data models for regulatory reporting, trade lifecycle, and market risk processing.

Databricks Lakehouse & Delta Framework
• Architect and implement Bronze/Silver/Gold layer modeling patterns within the Databricks Lakehouse.
• Apply Delta Lake best practices including:

o optimized file management

o Z-Ordering

o Delta Change Data Feed (CDF) o schema evolution & enforcement o ACID transaction handling
• Build reusable frameworks for ingestion, cleansing, transformation, and consumption of data across Lakehouse layers.
• Enable governance, lineage, and auditability using Unity Catalog or equivalent cataloging tools.

Collaboration, Leadership & Delivery
• Collaborate closely with quants, product owners, architects, risk tech, and business users.
• Participate in agile ceremonies sprint planning, refinement, design reviews.
• Mentor junior engineers and contribute to building strong engineering practices across tech teams.

Required Skills & Experience
• 1013 years of hands-on experience in Big Data engineering.
• Expert skills in:

o PySpark dataframe optimizations, partitioning, broadcast strategies, distributed computing.

o Kafka producer/consumer design, schema registry, streaming ETLs.

o Hadoop ecosystem HDFS, YARN, MapReduce/Tez, Oozie/Airflow.

o Hive advanced query tuning, TEZ optimization, partition/bucket management.
• Extensive hands-on experience with Databricks Lakehouse, including:

o Bronze/Silver/Gold layer modeling

o Delta Lake optimizations
o Data quality frameworks on Lakehouse

o Structured & unstructured data handling
• Experience in Global Markets, Risk, Treasury, Trade Surveillance, or Regulatory Reporting.
• Strong SQL knowledge with experience working on massive datasets (TB/PB scale).

Experience with CI/CD practices Git, Jenkins, Bitbucket, build pipelines.

TCS Employee Benefits Summary:

Discretionary Annual Incentive.

Comprehensive Medical Coverage: Medical & Health, Dental & Vision, Disability Planning & Insurance, Pet Insurance Plans.

Family Support: Maternal & Parental Leaves.

Insurance Options: Auto & Home Insurance, Identity Theft Protection.

Convenience & Professional Growth: Commuter Benefits & Certification & Training Reimbursement.

Time Off: Vacation, Time Off, Sick Leave & Holidays.

Legal & Financial Assistance: Legal Assistance, 401K Plan, Performance Bonus, College Fund, Student Loan Refinancing.

Salary Range: $110,000- 125,000 a year

About Tata Consultancy Services

Tata Consultancy Services (TCS) is an Indian multinational information technology (IT) services and consulting company, headquartered in Mumbai, Maharashtra, India. It is a subsidiary of Tata Group and operates in 149 locations across 46 countries. TCS is the largest Indian company by market capitalization and is ranked 11th on the Forbes Global 2000 list of the world's biggest public companies. TCS is also the second-largest IT services company in the world by revenue and the largest employer of women in India. The company provides services in areas including IT, consulting, and business solutions.
Learn more about Tata Consultancy Services
Size
469,261 employees
Industry

Similar Jobs

More Jobs at Tata Consultancy Services

  • Tata Consultancy Services
    Control System Engineer
    $100K — $125K *
    Whitakers, NC 27891 (Nash County)
    Manufacturing & Automotive
    In-Person
  • Tata Consultancy Services
    Azure Databricks Architect
    $130K — $135K *
    Frisco, TX 75034 (Denton County)
    Information Technology
    In-Person
  • Tata Consultancy Services
    SME Network Autonomy
    $150K — $180K *
    Edison, NJ 08817 (Middlesex County)
    Telecommunications & Hardware
    In-Person
  • Tata Consultancy Services
    Software Developer
    $100K — $125K *
    Tampa, FL 33647 (Hillsborough County)
    Information Technology
    In-Person
  • Tata Consultancy Services
    Engineer
    $110K — $120K *
    Charlotte, NC 28269 (Mecklenburg County)
    Information Technology
    In-Person

More Information Technology Jobs

Find similar Hadoop,Pyspark,Hive,Kafka - Senior Developer jobs: