Job DescriptionACTIVE SECURITY CLEARANCE AT THE TS/SCI POLYGRAPH LEVEL IS REQUIREDWe are seeking a results-driven
ETL & Data Process Automation Engineer whose primary focus is transforming reporting workflows through cutting-edge process automation. In this role, you will design, engineer, and deploy highly scalable Extract, Transform, Load (ETL) data pipelines that handle both batch and incremental data ingestion with robust validation controls. Utilizing SQL and Python data frameworks (Pandas, PySpark, or Polars), you will preprocess, normalize, aggregate, and reshape high-volume datasets into clean, actionable structures. Furthermore, you will develop and maintain reusable Python automation packages for agency-wide use, driving seamless data processing and automated reporting across critical mission threads.
The annual base salary range for this role is $156,000-$184,000 (USD) , which does not include discretionary bonus compensation or our comprehensive benefits package. Actual compensation offered to the successful candidate may vary from posted hiring range based upon geographic location, work experience, education, and/or skill level, among other things.
Required SkillsAutomated Reporting & Workflow Engineering: Proven expertise designing and automating end-to-end reporting processes, replacing manual churn with programmatic pipelines.
- Scalable ETL Architecture & Ingestion: Deep experience building robust batch and incremental ETL/ELT pipelines with automated data validation, quality checks, and error handling.
- Advanced Data Transformation & Modeling: Mastery of SQL and Python for preprocessing, filtering, normalizing, aggregating, and reshaping complex, multi-source datasets.
- Python Data Processing Toolkits: Hands-on proficiency with modern Python data libraries, specifically Pandas, PySpark, or Polars.
- Python Package Development & Maintenance: Experience authoring, testing, and maintaining reusable Python packages and libraries for deployment and use by agency personnel.
- Education: Bachelor's Degree in Data Science, Computer Science, Computational Linguistics, Mathematics, or a related technical discipline, PLUS 10+ years of professional experience in data science, NLP, or software engineering OR Associate's Degree PLUS 12+ years of specialized technical experience in lieu of a degree.
Desired SkillsAutomated Dashboarding & BI Tools: Hands-on experience constructing and automating interactive visual dashboards using Power BI, Tableau, or similar business intelligence tools.
Distributed Computing Frameworks: Familiarity operating within distributed big-data environments using Apache Spark or Hadoop ecosystems.
CI/CD & MLOps Infrastructure: Exposure to automated testing and deployment pipelines (GitLab CI, Jenkins) for python packages and data scripts.