Job Location : Toronto, CA (Onsite/Hybrid from Day 1)Job Description We are seeking a highly skilled
Ab Initio Developer (Big Data & Hadoop) to join our Agile Scrum team and support the development of enterprise-scale Big Data and Data Lake solutions. This role will be responsible for designing, developing, and supporting data integration applications leveraging
Ab Initio, Hadoop, and modern Big Data technologies . The ideal candidate will combine strong technical expertise with the ability to collaborate across business and technology teams to deliver high-quality data solutions.
Key Responsibilities - Participate in Agile/Scrum ceremonies, sprint planning, design discussions, and development activities.
- Analyze business requirements and translate them into functional and technical specifications.
- Design, develop, test, and support scalable data integration and ETL solutions using Ab Initio.
- Build and maintain enterprise data pipelines supporting the organization's Data Lake and analytics platforms.
- Ensure application quality, performance, scalability, and adherence to development standards.
- Collaborate with architects, data engineers, business analysts, and stakeholders to deliver data-driven solutions.
- Create project estimates and contribute to project planning and delivery discussions.
- Participate in code reviews and promote development best practices.
- Develop reusable and modular components to improve development efficiency and maintainability.
- Troubleshoot production issues and perform performance tuning of ETL and Big Data applications.
- Mentor junior developers and provide technical guidance to team members.
- Support continuous improvement initiatives and adoption of emerging technologies.
Mandatory Skills - Bachelor's degree in Computer Science, Engineering, Statistics, Econometrics, Information Systems, or a related quantitative field.
- 10+ years of overall IT experience with at least 5+ years of hands-on experience in Big Data and Data Integration solutions.
- Expert knowledge of Ab Initio technologies, including Graphical Development Environment (GDE), Co-Operating System, Control Center, Metadata Hub, Enterprise Meta Environment (EME), EME Portal, Acquire>It, Express>It, Conduct>It, Data Quality Environment, and Query>It.
- Strong experience in designing, developing, and supporting complex ETL and enterprise data integration solutions.
- Hands-on experience with shell scripting, complex SQL development, Hadoop commands, and Git version control.
- Strong programming skills in at least two of the following languages: Python, Java, and Scala.
- Proven experience developing reusable, scalable, and maintainable code components.
- Experience working with Hadoop ecosystem technologies such as Hive, Spark, Kafka, HBase, and Sqoop.
- Strong understanding of Big Data architectures, Data Lake technologies, and large-scale data processing frameworks.
- Experience in performance tuning, optimization, and troubleshooting of ETL and Big Data applications.
- Familiarity with Hortonworks, ZooKeeper, Oozie, and NoSQL technologies is preferred.
Good to Have - Hands-on experience with Big Data technologies including:
- Hadoop, Hive, Spark, Kafka, HBase and Sqoop etc
- Experience working with Data Lake architectures and large-scale analytics platforms.
- Knowledge of Hortonworks, ZooKeeper, Oozie, and NoSQL databases.
- Experience with data modeling, data quality, and performance optimization.
- Familiarity with cloud-based data platforms and modern data engineering practices.
- Understanding of CI/CD processes and DevOps methodologies.
The base compensation range for this role in the posted location is: 88786- 95298
Capgemini provides compensation range information in accordance with applicable national, state, provincial, and local pay transparency laws. The base compensation range listed for this position reflects the minimum and maximum target compensation Capgemini, in good faith, believes it may pay for the role at the time of this posting. This range may be subject to change as permitted by law.
The actual compensation offered to any candidate may fall outside of the posted range and will be determined based on multiple factors legally permitted in the applicable jurisdiction.
These may include, but are not limited to: Geographic location, Education and qualifications, Certifications and licenses, Relevant experience and skills, Seniority and performance, Market and business consideration, Internal pay equity.
It is not typical for candidates to be hired at or near the top of the posted compensation range.
In addition to base salary, this role may be eligible for additional compensation such as variable incentives, bonuses, or commissions, depending on the position and applicable laws.
Capgemini offers a comprehensive, non-negotiable benefits package to all regular, full-time employees. In the U.S. and Canada, available benefits are determined by local policy and eligibility and may include:
- Paid time off based on employee grade (A-F), defined by policy: Vacation: 12-25 days, depending on grade, Company paid holidays, Personal Days, Sick Leave
- Medical, dental, and vision coverage (or provincial healthcare coordination in Canada)
- Retirement savings plans (e.g., 401(k) in the U.S., RRSP in Canada)
- Life and disability insurance
- Employee assistance programs
- Other benefits as provided by local policy and eligibility