Tata Consultancy Services

Hadoop Hive Python Developer

Tata Consultancy Services$110K — $125K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • Minimum 9 years of hands-on experience in Big Data engineering
  • Expertise in Hadoop ecosystem, including HDFS, YARN, and MapReduce/Tez
  • Proficiency in programming with Python and PySpark for ETL pipelines
  • Advanced skills in Databricks Lakehouse architecture and Delta Lake optimizations
  • Strong SQL knowledge with experience on large-scale datasets
  • Familiarity with CI/CD tools such as Git, Jenkins, and Bitbucket
  • Experience in financial systems, particularly in Global Markets and Risk.

Responsibilities

  • Design and optimize PySpark-based ETL pipelines for on-prem and cloud environments
  • Build real-time ingestion frameworks using Kafka for market data
  • Tune and manage Hadoop components for performance optimization
  • Architect Bronze/Silver/Gold modeling patterns in Databricks Lakehouse
  • Implement Delta Lake best practices for data management and governance
  • Collaborate with cross-functional teams including quants and product owners
  • Mentor junior engineers and enhance overall engineering practices.

Benefits

  • Discretionary Annual Incentive
  • Comprehensive Medical Coverage, including dental and vision
  • Parental and Maternal leave support
  • Variety of insurance options, including auto and home
  • Professional development benefits for certifications and training
  • Vacation and sick leave, along with holidays
  • Legal assistance and financial benefits, including 401K and student loan refinancing.
Full Job Description
Must Have Technical/Functional Skills

Primary skills: Hadoop, Hive, Python, PySpark, Apache Kafka, Hadoop Ecosystem, Hive, Databricks Lakehouse Architecture, Delta Lake, Bronze/Silver/Gold Data Modeling, Big Data ETL Pipeline Development, SQL, Real-time Data Ingestion Frameworks, Data Governance & Cataloging, CI/CD Tools Git, Jenkins, Bitbucket, Workflow Orchestration, and Cloud & On-Prem Big Data Platforms.

Experience: Minimum 9+ years

Roles & Responsibilities

Seeking a Senior Big Data Engineer with 9-14 years of experience specializing in Hadoop, Python, Hive PySpark, Kafka, and strong experience designing data solutions for large-scale financial systems.

In addition, the candidate must possess advanced expertise in Databricks Lakehouse architecture, particularly around Bronze/Silver/Gold layer data modeling, Delta Lake optimizations, and building reliable, scalable pipelines for regulatory, risk, trading, and analytics workloads.

This role focuses on delivering highly performant, well-governed data platforms that support the banks mission-critical global markets functions.

Key Responsibilities:

Big Data Platform Engineering
• Design, develop, and optimize PySpark-based ETL pipelines running on on-prem Hadoop clusters and cloud environments.
• Build high-volume ingestion frameworks using Kafka for real-time and near-real-time trading and market data.
• Develop, tune, and manage Hadoop ecosystem componentsHDFS, YARN, MapReduce, Tez, Oozie/Airflow.
• Build high-performance, optimized Hive data models for regulatory reporting, trade lifecycle, and market risk processing.

Databricks Lakehouse & Delta Framework
• Architect and implement Bronze/Silver/Gold layer modeling patterns within the Databricks Lakehouse.
• Apply Delta Lake best practices including:

o optimized file management

o Z-Ordering

o Delta Change Data Feed (CDF)

o schema evolution & enforcement

o ACID transaction handling
• Build reusable frameworks for ingestion, cleansing, transformation, and consumption of data across Lakehouse layers.
• Enable governance, lineage, and auditability using Unity Catalog or equivalent cataloging tools.

Collaboration, Leadership & Delivery
• Collaborate closely with quants, product owners, architects, risk tech, and business users.
• Participate in agile ceremonies sprint planning, refinement, design reviews.
• Mentor junior engineers and contribute to building strong engineering practices across tech teams.

Required Skills & Experience
• 9-14 years of hands-on experience in Big Data engineering.
• Expert skills in:

o PySpark dataframe optimizations, partitioning, broadcast strategies, distributed computing.

o Kafka producer/consumer design, schema registry, streaming ETLs.

o Hadoop ecosystem HDFS, YARN, MapReduce/Tez, Oozie/Airflow.

o Hive advanced query tuning, TEZ optimization, partition/bucket management.
• Extensive hands-on experience with Databricks Lakehouse, including:

o Bronze/Silver/Gold layer modeling

o Delta Lake optimizations

o Data quality frameworks on Lakehouse

o Structured & unstructured data handling
• Experience in Global Markets, Risk, Treasury, Trade Surveillance, or Regulatory Reporting.
• Strong SQL knowledge with experience working on massive datasets (TB/PB scale).
• Experience with CI/CD practices Git, Jenkins, Bitbucket, build pipelines.

TCS Employee Benefits Summary:

Discretionary Annual Incentive.

Comprehensive Medical Coverage: Medical & Health, Dental & Vision, Disability Planning & Insurance, Pet Insurance Plans.

Family Support: Maternal & Parental Leaves.

Insurance Options: Auto & Home Insurance, Identity Theft Protection.

Convenience & Professional Growth: Commuter Benefits & Certification & Training Reimbursement.

Time Off: Vacation, Time Off, Sick Leave & Holidays.

Legal & Financial Assistance: Legal Assistance, 401K Plan, Performance Bonus, College Fund, Student Loan Refinancing.

Salary Range: $110,000-$125,000 a year

About Tata Consultancy Services

Tata Consultancy Services (TCS) is an Indian multinational information technology (IT) services and consulting company, headquartered in Mumbai, Maharashtra, India. It is a subsidiary of Tata Group and operates in 149 locations across 46 countries. TCS is the largest Indian company by market capitalization and is ranked 11th on the Forbes Global 2000 list of the world's biggest public companies. TCS is also the second-largest IT services company in the world by revenue and the largest employer of women in India. The company provides services in areas including IT, consulting, and business solutions.
Learn more about Tata Consultancy Services
Size
469,261 employees
Industry

Similar Jobs

More Jobs at Tata Consultancy Services

  • Tata Consultancy Services
    Analyst Testing
    $100K — $115K *
    San Antonio, TX 78228 (Bexar County)
    Finance & Insurance
    In-Person
  • Tata Consultancy Services
    Tosca Architect
    $100K — $120K *
    Houston, TX 77084 (Harris County)
    Information Technology
    In-Person
  • Tata Consultancy Services
    Engineer
    $100K — $120K *
    Plano, TX 75025 (Collin County)
    Information Technology
    In-Person
  • Tata Consultancy Services
    Client Partner
    $191K — $258K *
    Owings Mills, MD 21117 (Baltimore County)
    Enterprise Technology
    In-Person
  • Tata Consultancy Services
    Hadoop Hive Python Developer
    $110K — $125K *
    Charlotte, NC 28269 (Mecklenburg County)
    Information Technology
    In-Person

More Information Technology Jobs

Find similar Hadoop Hive Python Developer jobs: