Data Engineer

Marathon TS

$110K — $130K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree or equivalent practical experience.
  • Expertise in distributed computing frameworks for large-scale data processing.
  • Familiarity with Amazon Web Managed Services (AWS) or other cloud environments.
  • Working experience with datastores like PostgreSQL, S3, Redshift, MongoDB/DynamoDB, etc.
  • Proficient in Python with key libraries like pandas and PySpark.
  • Working knowledge of software platforms like Docker, Kubernetes, and JMS/SQS.
  • Experience with Agile development methodology.

Responsibilities

  • Identify and implement internal process improvements for scalability and automation.
  • Develop and design data pipelines for end-to-end solutions.
  • Maintain artifacts related to ETL processes like schemas and data dictionaries.
  • Integrate data pipelines with AWS cloud services for insights extraction.
  • Support stakeholder data infrastructure needs and resolve data-related issues.
  • Design functional dataflows to accommodate both raw and expected data.
  • Provide Tier 3 technical support for applications and dataflows.

Benefits

  • Collaborative team environment with key stakeholders and government partners.
  • Opportunity to work on critical national security data systems.
  • Direct involvement in innovative projects with new technologies.
  • Remote/Hybrid work flexibility with a high on-site presence.
Full Job Description
Marathon TS
Data Engineer
TS/SCI Clearance Required
Hybrid (Washington, DC)


Role Description

As a Data Engineer, you will be required to interpret business needs and select appropriate technologies and have experience in implementing data governance of shared and/or master sets of data. You will work with key business stakeholders, IT experts, and subject-matter experts to plan and deliver optimal data solutions. You will create, maintain, and optimize data pipelines as workloads move from development to production for specific use cases to ensure seamless data flow for the use case. You will perform technical and non-technical analyses on project issues and help to ensure technical implementations follow quality assurance metrics. You will analyze data and systems architecture, create designs, and implement information systems solutions.

Duties:
  • Identify, design and implement internal process improvements including re-designing infrastructure for greater scalability, optimizing data delivery, and automating manual processes.
  • Develop and design data pipelines to support an end-to-end solution.
  • Develop and maintain artifacts (e.g., schemas, data dictionaries, and transforms related to ETL processes).
  • Integrate data pipelines with AWS cloud services to extract meaningful insights.
  • Work with stakeholders to support their data infrastructure needs and assist with data-related technical issues.
  • Design and develop robust and functional dataflows to support raw data and expected data.
  • Provide Tier 3 technical support for deployed applications and dataflows.
  • Define and communicate a clear product vision for our client's software products, aligning user needs and business objectives.
  • Create and manage product roadmaps that reflect both innovation and growth strategies.
  • Partner with a government product owner and a product team of 7-8 FTEs.
  • Collaborate with the rest of data engineering team to design and launch new features.
  • Coordinate and document dataflows, capabilities, etc.
  • Occasionally (as needed) support to off-hours deployment such as evening or weekends.

Qualifications:
  • Bachelor's degree or equivalent practical experience.
  • Expertise in distributed computing frameworks to handle large-scale data processing.
  • Familiarity working with Amazon Web Managed Services (AWS) or any other cloud environments.
  • Working experience with datastores like PostgreSQL, S3, Redshift, MongoDB/DynamoDB, Redis, Elasticsearch/OpenSearch and SQL.
  • Proficient utilizing Python with key libraries like pandas and PySark, NiFi, Airflow, AWS Lambda or similar technologies.
  • Working knowledge with software platforms and services, such as Docker, Kubernetes, JMS/SQS, SNS and Kafka.
  • Familiar with Linux/Unix server environments.
  • Experience with Agile development methodology.
  • Publishing and/or presenting design reports.
  • Coordinating with other team members to reach project milestones and deadlines.
  • Working knowledge with Collaboration tools, such as, Jira and Confluence.

Preferred Qualifications:
  • Master's degree or equivalent experience in a related field.
  • Familiarity and experience with the Intelligence Community (IC), and the Client cycle.
  • Familiarity and experience with the Department of Homeland Security (Client).
  • Direct Experience with Client and Intelligence Community (IC) component's data architectures and environments (IC-GovCloud experience preferred).
  • Experience with cloud message APIs and usage of push notifications.
  • Keen interest in learning and using the latest software tools, methods, and technologies to solve real world problem sets vital to national security.
  • Working knowledge with public keys and digital certificates.
  • Experience with DevOps environments.
  • Expertise in various COTS, GOTS, and open source tools which support development of data integration and visualization applications.
  • Experience with cloud message APIs and usage of push notifications.
  • Specialization in Object Oriented Programming languages, scripting, and databases.

Role Requirements:
  • Active TS/SCI
  • Full Time
  • High colocation, 70-80% onsite
#cjjobs

Similar Jobs

More Jobs at Marathon TS

More Information Technology Jobs

Find similar Data Engineer jobs: