Python Developer in Data Engineering

Compunnel

$90K — $110K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Advanced proficiency in Python, pandas, and polars for data manipulation.
  • Experience optimizing workflows for large datasets.
  • Hands-on with Docker for containerization.
  • Knowledge of Kubernetes for orchestrating applications.
  • Familiarity with ClickHouse for analytical databases.
  • Proficient in pytest for testing frameworks.
  • Experience using Dask for distributed computing.

Responsibilities

  • Develop and optimize data manipulation workflows using pandas and polars.
  • Design and implement containerized applications with Docker and Kubernetes.
  • Build and maintain data pipelines for ClickHouse databases.
  • Develop event-driven architectures using NATS messaging systems.
  • Implement distributed computing solutions with Dask.
  • Write unit and integration tests with pytest.
  • Manage source code using Git and follow collaborative practices.

Benefits

  • Work in a collaborative environment with a focus on innovative data solutions.
  • Gain hands-on experience with cutting-edge technologies like Docker and Kubernetes.
  • Opportunity to work with large-scale datasets and advanced analytics.
  • Flexible onsite work schedule with 4 days onsite.
Full Job Description
Job Summary

We are seeking a Python Developer with strong data engineering expertise to design, develop, and maintain high-performance data processing pipelines using modern Python frameworks and tools. The role will involve working with large-scale datasets, containerized applications, distributed computing platforms, analytical databases, and event-driven architectures to deliver scalable and reliable data solutions. The position is based in Toronto, ON, with 4 days onsite.

Key Responsibilities
• Develop and optimize data manipulation workflows using pandas and polars for efficient processing of large datasets.
• Design and implement containerized applications using Docker and Kubernetes to support scalable and reliable deployments.
• Build and maintain data pipelines integrating with ClickHouse columnar databases for analytical workloads.
• Develop event-driven architectures using NATS messaging systems for asynchronous data processing.
• Implement distributed computing solutions using Dask for processing datasets beyond single-machine memory constraints.
• Write comprehensive unit and integration tests using pytest and maintain appropriate code coverage standards.
• Manage source code using Git and follow established branching and collaborative development practices.
• Optimize data processing code and workflows for performance and scalability.

Required Qualifications
• Advanced proficiency in Python, pandas, and polars for data manipulation, transformation, and analysis.
• Experience optimizing Python code and data processing workflows for large datasets.
• Hands-on experience with Docker, including building container images and composing multi-container applications.
• Knowledge of Kubernetes for container orchestration and deployment management.
• Working knowledge of ClickHouse or similar columnar databases for OLAP workloads and analytical queries.
• Familiarity with NATS.io for message-driven systems and asynchronous workflows.
• Proficiency with pytest for unit testing, integration testing, and maintaining code coverage.
• Experience with Dask for parallel processing and out-of-core computations.
• Strong command of Git workflows, branching strategies, and collaborative development practices.

Preferred Qualifications
• Experience with additional Python libraries for data science and machine learning.
• Familiarity with CI/CD pipelines and DevOps practices.
• Background in financial services or capital markets data systems.

Similar Jobs

More Jobs at Compunnel

More Information Technology Jobs

Find similar Python Developer in Data Engineering jobs: