JOB DESCRIPTION
Work Location: Pittsburgh, NJ or NY.
Work Mode : Onsite
Pay Range :$120K-$140K /Yr Base + Annual Bonus
The posted range is the hiring range for this role — a subset of the broader range available to employees over time — and reflects base salary across our national hiring scale. Final offers are based on several factors, including the candidate's skills and experience, internal pay equity, work location, market conditions for the role, and the specific scope and responsibilities of the position. The top of the range is reserved for candidates who notably exceed the requirements; the lower end applies to those with less experience or fewer preferred qualifications. For positions based in higher-cost zones (e.g., California, New York, New Jersey), actual compensation may exceed the posted range; your recruiter will share specifics during the process
For more information on benefits and what we offer please visit us at https://www.exlservice.com/us-careers-and-benefits
Job Overview:
We are seeking an experienced Technology Lead to own the design and delivery of an enterprise batch feature store. The ideal candidate is a hands-on data engineering leader who can make sound architecture decisions, guide developers, and take complex data pipelines through to production.
JOB RESPONSIBILITIES
- Lead the architecture, development, and production rollout of a batch feature store supporting analytics and AI use cases.
- Design scalable PySpark and Impala SQL pipelines that transform data from Hive and Oracle into consistent, reusable features.
- Define technical standards for feature logic, data quality, reconciliation, performance, monitoring, and failure recovery.
- Own batch scheduling and dependencies using CA7 or an equivalent platform.
- Provide hands-on technical direction through design reviews, code reviews, troubleshooting, and mentoring.
- Establish effective Git/Bitbucket, CI/CD, testing, and AI-assisted development practices.
- Coordinate production releases and change management in line with SDLC requirements.
- Work closely with business, analytics, and technology stakeholders to resolve requirements, communicate trade-offs, and deliver against milestones.
- Plan and track work in an Agile environment using Jira.
JOB QUALIFICATIONS
Experience requirements
- 8+ years of experience in data engineering, data platforms, or enterprise application development.
- 5+ years of hands-on experience with PySpark, large-scale ETL/ELT development, and advanced SQL.
- 3+ years leading technical design and delivery for data engineering teams.
- Experience in banking or financial services is a plus.
Required skills
- Advanced PySpark and Impala SQL, including query and pipeline performance tuning.
- Proven experience designing enterprise-scale ETL/ELT pipelines.
- Strong expertise with Hive and Oracle data platforms.
- Experience with CA7 or equivalent batch scheduling.
- Proficiency in Linux scripting, Git/Bitbucket, and CI/CD.
- Working knowledge of Agile delivery, Jira, SDLC, production release, and change management.
- Practical experience with AI-assisted development tools.
- Strong communication skills and a track record of leading delivery across multiple stakeholder groups.