Affinity

4531 Data Engineer

Affinity$110K — $130K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 2 years of experience in a Data Engineer role or similar.
  • Strong proficiency in Python or relevant programming languages.
  • Hands-on experience with data orchestration tools like Apache Airflow, dbt, or Prefect.
  • Solid understanding of AWS cloud platform; GCP or Azure experience is a plus.
  • Expertise in SQL and familiarity with databases such as PostgreSQL and Snowflake.
  • Knowledge of big data processing frameworks like Apache Spark or Apache Kafka.
  • Familiarity with MLOps tools for machine learning model deployment and support.
  • Strong troubleshooting skills in monitoring data pipelines.

Responsibilities

  • Write clean, efficient code for data solutions in Python.
  • Design and build robust data workflows using Apache Airflow or similar tools.
  • Work comfortably in AWS and ideally in cloud environments.
  • Extract and integrate data from multiple sources, ensuring quality.
  • Leverage big data frameworks to manage large-scale workflows.
  • Support implementation and scaling of machine learning models in production.
  • Monitor data pipeline health and troubleshoot issues effectively.
  • Collaborate with team members to document processes and communicate solutions.

Benefits

  • Collaborative work environment with cross-functional teams.
  • Opportunity to work with advanced data and machine learning technologies.
  • Chance to impact data infrastructure and influence scalability.
  • Support for professional development and continuous learning opportunities.
  • Work in a location with government-related projects requiring TS/SCI clearance.
Full Job Description
4531 Data Engineer
4531 | TS/SCI

Job Description:

OVERVIEW:

We are seeking a skilled and passionate Data Engineer to join our team. You will play a critical role in designing, building, and maintaining our data infrastructure to ensure seamless data flow, scalability, and reliability. You will work closely with data scientists, analysts, and other stakeholders to develop efficient data pipelines, manage large datasets, and integrate machine learning models into production environments.

GENERAL DUTIES:

  • Programming Fundamentals: Write clean, efficient, and scalable code to build and optimize data solutions using programming languages like Python.
  • Data Pipeline Development: Design, build, and orchestrate robust and reliable data workflows using tools such as Apache Airflow, dbt, Prefect, or Dagster.
  • Cloud Platform Familiarity: Work comfortably in cloud environments, with a strong preference for experience in AWS. Experience in GCP or Azure is also highly valued.
  • Database & Querying Skills: Extract, integrate, and ensure the quality of data from various sources using tools and technologies such as SQL, PostgreSQL, Snowflake, Amazon Redshift, or BigQuery.
  • Big Data Processing: Leverage frameworks like Apache Spark, Databricks, or Apache Kafka to process and manage large-scale data workflows with reliability and efficiency.
  • ML Integration / MLOps: Support the implementation, deployment, and scaling of machine learning models in production environments using tools like Amazon SageMaker, MLflow, or Kubeflow.
  • Monitoring & Troubleshooting: Monitor data pipeline health, troubleshoot issues, and ensure data consistency using tools such as Amazon CloudWatch, Datadog, or Great Expectations.
  • Collaboration & Documentation: Work closely with data scientists, analysts, and other stakeholders to understand data requirements, communicate solutions, and document processes using tools like Git, Jira, and Confluence.


REQUIRED QUALIFICATIONS:

  • 2 years of experience as a Data Engineer or similar role.
  • Strong proficiency in Python or other programming languages relevant to data engineering.
  • Hands-on experience with data pipeline orchestration tools (e.g., Apache Airflow, dbt, Prefect, Dagster).
  • Solid understanding of cloud platforms (AWS strongly preferred; GCP or Azure experience also considered).
  • Expertise in SQL and familiarity with relational and columnar databases (e.g., PostgreSQL, Snowflake, BigQuery).
  • Knowledge of big data processing frameworks (e.g., Apache Spark, Databricks, or Apache Kafka).
  • Familiarity with machine learning workflows and experience implementing MLOps tools (e.g., Amazon SageMaker, MLflow, or Kubeflow) in production environments.
  • Strong troubleshooting skills and experience monitoring data pipelines and system health using tools like Amazon CloudWatch, Datadog, or Great Expectations.
  • Excellent communication skills and a collaborative mindset, with a focus on documentation and best practices.


DESIRED QUALIFICATIONS:

  • Experience working with large-scale distributed systems.
  • Knowledge of data governance and security best practices.
  • Proven ability to work in cross-functional teams and contribute to problem-solving and innovation.


CLEARANCE:

  • Active TS/SCI clearance minimum required


Job Details

City : Suitland

State : Maryland

About Affinity

Affinity’s patented technology structures and analyzes millions of data points across emails, calendars, and third-party sources to offer users the tools they need to automatically manage their most valuable relationships, prioritize important connections, and discover untapped opportunities. Affinity uses artificial intelligence to analyze relationship strength and illuminate the best paths to warm introductions. The platform also offers a holistic view of users’ networks in a centralized, automatically updated database without any manual upkeep. Founded in 2014, Affinity is headquartered in San Francisco, California. Affinity has raised $120M to date and is backed by leading investors including Menlo Ventures, Advance Venture Partners, 8VC and MassMutual Ventures. It has over 2,700 customers in 70 countries, including venture capital firms such as Bain Capital Ventures and Kleiner Perkins, private equity firms such as SoftBank Group, investment bankers such as Woodside Capital Partners, financial services firms such as Fidelity Investments, real estate companies such as Tishman Speyer, insurers such as American Family Insurance and enterprises such as Nike, Qualcomm and Twilio. Affinity has been named in Fortune Magazine's Best Workplaces, Inc. Magazine's Best Workplaces and editor's number one pick, the Data Breakthrough Award, BIG Innovation Award and others.
Learn more about Affinity
Size
1,000 employees
Industry
Founded
2014

Similar Jobs

More Jobs at Affinity

More Information Technology Jobs

Find similar 4531 Data Engineer jobs: