Full Job Description
Required Skills & Experience
Strong hands-on experience with GCP Dataproc and Apache Spark (PySpark preferred)
Python scripting for custom data transformations and business logic
Proficiency in BigQuery (SQL, table design, partitioning, optimization)
Experience building ETL pipelines in GCP ecosystem
Solid understanding of Salesforce data model and Salesforce APIs (Bulk API, REST API)
Experience with GCS for file handling and cloud storage patterns
Preferred
Exposure to GitLab CI/CD (understanding pipeline YAML, not building from scratch)
Prior Salesforce data migration projects (especially Health Cloud or Life Sciences Cloud)
Knowledge of DLP implementation and PII handling best practices
Healthcare data compliance knowledge (PIPEDA, HIPAA)
Technical Depth
Ability to debug complex Spark jobs and resolve memory/performance issues
Strong SQL skills for data validation and reconciliation
Experience handling large datasets (300GB+ structured, 3TB+ documents)
Comfortable working with 70+ object schemas and complex relationships
Note:
TCS does not use artificial intelligence tools for candidate screening or evaluation.
This posting is for a current vacancy
The hiring process includes an initial screening by the TCS Hiring Team, followed by a technical evaluation and managerial discussion conducted by the Business Team, and concluding wit h the final HR evaluation.