ECS

Senior Databricks Migration Engineer

ECS • $150K — $160K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or related field.
  • 5+ years in data engineering or related roles.
  • 5+ years with cloud platforms (Azure, AWS, GCP).
  • 1+ year leading complex data projects and teams.
  • Experience with Databricks Lakehouse, Apache Spark, and Delta Lake.

Responsibilities

  • Lead migration from SQL Server to Databricks Lakehouse.
  • Translate relational data warehousing to scalable Lakehouse frameworks.
  • Design reusable ETL/ELT frameworks using PySpark and Databricks Workflows.
  • Optimize Databricks SQL Warehouses for Power BI performance.
  • Implement advanced optimization techniques for data management.
  • Define governance standards for cluster sizing and auto-scaling.
  • Create monitoring dashboards for Databricks Unit consumption.

Benefits

  • Hybrid work environment in Arlington, VA.
  • Opportunity to lead innovative data migration projects.
  • Access to advanced data technologies and tools.
  • Collaborative team culture with knowledge transfer initiatives.
  • Focus on professional development through workshops and code reviews.
Full Job Description
ECS is seeking a Senior Databricks Migration Engineer to work in our Arlington, VA (hybrid) office. Please Note: This position is contingent upon additional funding.

The individual serves as the authoritative resource for the agency who specializes in preparing big data infrastructure for analytical or operational uses. They are responsible for designing and creating systems that collect, manage, and convert raw data into usable information for data scientists and business analysts to interpret and enables the agency to make smarter decisions and optimize operations.

Responsibilities include:
  • Lead the technical migration from legacy SQL Server stored procedures and ADF pipelines to Databricks Lakehouse (Delta Lake), ensuring best practice Lakehouse design.
  • Translate traditional relational data warehousing paradigms into scalable, distributed Lakehouse frameworks (Bronze, Silver, Gold).
  • Design robust, reusable ETL/ELT frameworks using PySpark, Delta Live Tables (DLT), and Databricks Workflows.
  • Architect and refine the Gold Layer (dimensional models, star schemas) specifically to maximize Power BI performance.
  • Optimize Databricks SQL Warehouses to support high-concurrency, low-latency Power BI queries (DirectQuery and Import modes).
  • Implement advanced optimization techniques, including Z-Ordering, data skipping, liquid clustering, and materialized views.
  • Define and enforce governance standards for cluster sizing, auto-scaling policies, and serverless SQL compute to balance performance with cost.
  • Implement proactive monitoring dashboards to track Databricks Unit (DBU) consumption and identify cost-saving opportunities.
  • Establish best practices for partition strategies and file size management within Delta Lake.
  • Design and implement a robust data security model using Unity Catalog for centralized governance.
  • Enforce row-level and column-level security policies to ensure compliant data access for Power BI consumers and internal analysts.
  • Align the Lakehouse security architecture with existing enterprise Azure Active Directory (Microsoft Entra ID) and RBAC standards.
  • Act as the primary technical lead, conducting dedicated pair-programming sessions, workshops, and code reviews to transition the team from SQL-centric to Spark-centric thinking.
  • Create comprehensive technical documentation, including architecture diagrams, design patterns, and optimization playbooks.
  • Build a foundational knowledge transfer framework to ensure the internal team is fully self-sufficient post-migration.
  • Communicate effectively verbally and in written form to both technical and non-technical audience.
  • Work in an organized fashion, completing tasks timely while paying close attention to details.

Salary Range: $150,000-$160,000

General Description of Benefits

  • Bachelor's degree or higher from an accredited college or university in Computer Science, Engineering, or a related technical field.
  • 5+ years' experience in data engineering, data system development or related roles.
  • 5+ years' experience with cloud platforms (e.g. Azure, AWS, GCP).
  • 1+ year leading complex, cross-functional data projects and technical teams.
  • Experience with Databricks Lakehouse, Apache Spark, Delta Lake, cloud-native databases, storage solutions, and distributed compute platforms.
  • Experience with data warehousing, dimensional modeling, enterprise data lakes, incremental data loads, and metadata-driven ingestion and data quality frameworks using PySpark.
  • Mastery of data engineering principles, including data modeling, ETL (Extract, Transform, Load) processes, and data pipelines.
  • Proficiency with Azure Data Lake data storage and processing services.
  • Skilled at designing, building, and optimizing data pipelines for ingesting, transforming, and loading data.
  • Proficiency in languages such as SQL and Python/PySpark for data manipulation and pipeline development.
  • Skilled at identifying and resolving data-related challenges.
  • Skilled at creating efficient data models that meet business requirements.
  • Skilled at optimizing query performance and system scalability.

About ECS

ECS is a leading provider of digital solutions and services to the federal government. The company was founded in 2001 by Roy Kapani and has since grown to become a trusted partner to a wide range of government agencies. ECS offers a broad range of services, including cloud computing, cybersecurity, and artificial intelligence. The company has been recognized for its innovative solutions and has won numerous awards, including the AWS Public Sector Partner of the Year award.
Learn more about ECS
Size
2,000 employees
Industry

Similar Jobs

More Jobs at ECS

More Information Technology Jobs

Find similar Senior Databricks Migration Engineer jobs: