Presidio

Principal AI Data Engineer

Presidio$130K — $160K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, Information Systems, Data Engineering, or related field, or equivalent experience.
  • 10+ years of experience in data engineering, software engineering, or cloud data platforms.
  • 5+ years of hands-on experience with enterprise data platforms using Microsoft Azure and Microsoft Fabric.
  • Deep expertise with Azure data services including lakehouse, Data Factory, and semantic models.
  • Strong coding abilities in Python/PySpark and SQL for high-volume data pipelines.

Responsibilities

  • Establish engineering standards and practices for AI and data platform solutions.
  • Mentor engineers through architecture and code reviews, promoting best practices.
  • Evaluate and recommend emerging technologies for the data platform.
  • Build and manage the enterprise lakehouse on Microsoft Fabric and Azure.
  • Implement automated data quality frameworks to ensure integrity and compliance.

Benefits

  • Mentorship opportunities with experienced engineers and architects.
  • Access to cutting-edge technologies and frameworks in the data field.
  • Collaborative work environment with cross-functional teams.
  • Professional development support for certifications and learning.
  • Opportunity to lead impactful enterprise-level projects.
Full Job Description
The Role

Responsibilities Include:

Technical Leadership

  • Establish engineering standards, development practices, and implementation patterns for enterprise AI and data platform solutions.
  • Mentor engineers through architecture reviews, code reviews, technical coaching, and engineering best practices.
  • Evaluate emerging technologies and recommend improvements to the enterprise AI and data platform.
  • Partner with the AI Data Architect to translate enterprise strategy into scalable, secure, and production-ready technical solutions.
  • Promote engineering excellence across reliability, maintainability, automation, and operational support.
  • Provide technical leadership in evaluating implementation trade-offs and recommend improvements that strengthen the enterprise architecture while maintaining alignment with strategic objectives.


Data Platform Engineering (Microsoft Fabric & Azure)

  • Build and operate the enterprise lakehouse on Microsoft Fabric and Microsoft Azure, implementing the domain-oriented data products, medallion-layer structures, and Fabric-based semantic models defined in the enterprise architecture.
  • Develop, test, and maintain data pipelines for ingestion, transformation, and serving using Fabric-native tooling, Python, Spark, and SQL, with automated data validation to ensure integrity and timeliness.
  • Administer the Fabric and Azure data environments: capacity, workspaces, deployment pipelines, monitoring, and cost management.
  • Own performance tuning and operational excellence for the data platform, including incident response, root-cause analysis, and continuous improvement.
  • Establish and maintain engineering practices for the platform: version control, CI/CD, code review, testing standards, and release management.


Semantic Model & Data Product Implementation

  • Implement enterprise semantic models and certified data products to specification, encoding governed metric definitions, calculation logic, and business context from the metrics registry.
  • Implement row-level and object-level security in Fabric and OneLake that mirrors source-system permissions (e.g., Salesforce roles and visibility rules) to protect sensitive pipeline, customer, and people data.
  • Integrate source systems - CRM (Salesforce), CPQ, PSA, ERP, HRIS, and finance platforms - into the enterprise model so revenue, pipeline, people, cost, and customer data are consistently defined and analytics-ready.
  • Modernize data flows from legacy and server-based applications into the lakehouse, with reconciliation and validation frameworks that prove parity between legacy outputs and modernized models.
  • Connect governed, certified data sources to Data Visualization Platforms (e.g., Power BI, Tableau) and partner with BI developers to migrate duplicated logic into shared enterprise models.


AI Solution Engineering

  • Build the retrieval and grounding infrastructure - semantic model endpoints, metadata services, RAG patterns, certified MCP connectors, and context APIs - that lets AI applications and agents answer business questions with governed data.
  • Engineer the enterprise context layer in partnership with the AI Data Architect and AI Enablement function, making curated business context, policies, and definitions available to AI tools.
  • Implement guardrails, access controls, and quality gates for AI data consumption in accordance with company policies.


Data Quality & Operations

  • Implement automated data quality frameworks: validation rules, anomaly detection, reconciliation checks, and monitoring aligned to established quality standards.
  • Maintain lineage, documentation, and metadata for pipelines, models, and data products to support governance, certification, and auditability.
  • Support current-state assessment and knowledge capture from existing systems, prior development efforts, and third-party contractors, converting institutional knowledge into documented, maintainable code.


Collaboration

  • Partner daily with the AI Data Architect to refine designs based on implementation realities, propose technical alternatives, and deliver iteratively.
  • Work with BI developers, analysts, and domain teams to gather technical requirements and deliver reliable, well-documented data products.
  • Mentor and review the work of internal engineers and contractors, raising the engineering bar across the data function.


Required Skills and Experience:

  • Bachelor's degree in Computer Science, Information Systems, Data Engineering, or a related field, or equivalent practical experience.
  • 10+ years progressive experience in data engineering, software engineering, cloud data platforms, or enterprise analytics engineering. 5+ years of hands-on experience designing, building, and operating enterprise data platforms using Microsoft Azure, Microsoft Fabric, Databricks, Snowflake, or comparable data technologies.
  • Significant hands-on experience building and operating enterprise data platforms in production, including lakehouse and medallion architectures, domain-oriented data products, and semantic models.
  • Deep, hands-on expertise with Microsoft Fabric and Microsoft Azure data services: lakehouse, Data Factory, notebooks, semantic models, deployment pipelines, and identity-based access control (e.g., Entra ID).
  • Proven experience consolidating heterogeneous legacy source systems (e.g., mainframe, Oracle, PostgreSQL, on-premises SQL Server) into modern cloud data platforms, including reconciliation and validation across migrations.
  • Strong programming skills in Python/PySpark and SQL, with experience engineering high-volume production pipelines with automated auditing, validation, and recovery patterns.
  • Hands-on Salesforce experience, including administration and integration of SFDC data and permission models into analytical platforms.
  • Experience implementing row-level security and access controls in analytics platforms that mirror source-system permission models.
  • Demonstrated engineering discipline: version control, structured deployment (e.g., Fabric deployment pipelines), testing, and production support.
  • Strong communication skills and the ability to work effectively with architects, analysts, business stakeholders, and third-party contractors.


Preferred Skills and Professional Experience:

  • Experience building AI-ready data foundations: RAG pipelines, vector/semantic retrieval, MCP or similar connector frameworks, or agent-based data access patterns.
  • Experience with Power BI semantic model development (DAX, M, Tabular Editor) and/or Tableau connectivity and certified data sources.
  • Experience in sales operations, revenue operations, or go-to-market analytics domains, including territory, pipeline, and quota data models.
  • Experience with data quality tooling, observability, and automated reconciliation frameworks.
  • Experience working in contractor-heavy or transition environments, including knowledge capture, code remediation, and acquisition data integration.
  • Relevant certifications (e.g., Microsoft Fabric, Azure Data Engineer, Salesforce).


Technical Skills Snapshot

  • Cloud & Platform: Microsoft Fabric (OneLake, lakehouse, Direct Lake), Microsoft Azure data & analytics services, Data Factory.
  • Engineering: Python/PySpark, SQL, notebook-based ETL/ELT, version control, Fabric deployment pipelines, automated validation.
  • Modeling: Semantic and dimensional modeling (star/snowflake), medallion architecture, data product implementation, DAX/M.
  • Legacy Modernization: Mainframe, Oracle, PostgreSQL, and SQL Server consolidation into cloud lakehouse platforms.
  • Business Systems: Salesforce/CRM administration and integration, CPQ, PSA, ERP, HRIS.
  • Analytics & BI: Power BI and/or Tableau connectivity, certified data sources, row-level security.
  • AI Engineering: RAG, metadata/context services, MCP connectors, AI data guardrails.

About Presidio

Presidio is a leading provider of information technology services and solutions for businesses of all sizes. The company offers a wide range of services, including cloud computing, cybersecurity, data center solutions, and more. Presidio was founded in 2003 and has since grown to become one of the largest IT services companies in the United States. The company has over 3,000 employees and serves clients in a variety of industries, including healthcare, finance, and government.
Learn more about Presidio
Size
2,900 employees
Industry
Founded
2003
NASDAQ

Similar Jobs

More Jobs at Presidio

More Information Technology Jobs

Find similar Principal AI Data Engineer jobs: