Guidehouse

Databricks Data Scientist

Guidehouse$113K — $188K *
US-AnywhereRemote in United States
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor’s degree in computer science, engineering, mathematics, statistics, or relevant field.
  • 3-8 years of experience in data science, machine learning, or advanced analytics.
  • Strong proficiency in Python and SQL for data handling.
  • Experience with Databricks, Spark, Delta Lake, or similar platforms.
  • Hands-on experience with machine learning models from development to deployment.
  • Familiarity with ML lifecycle practices such as monitoring and version control.
  • Ability to analyze data and communicate findings effectively.

Responsibilities

  • Develop, train, and evaluate machine learning models using Databricks.
  • Prepare, clean, and manage datasets for analysis and modeling.
  • Optimize Python and SQL workflows for data exploration and model development.
  • Work with large-scale datasets using Databricks and Spark.
  • Design reusable workflows and pipelines in Databricks notebooks.
  • Monitor deployed models and recommend adjustments based on performance.
  • Translate business requirements into actionable data science solutions.

Benefits

  • Flexible work environment allowing for virtual work within the U.S.
  • Comprehensive total rewards package including competitive compensation.
  • Commitment to creating a diverse and supportive workplace.
Full Job Description

Job Family:

Data Science & Analysis, Data Science Consulting


Travel Required:

None


Clearance Required:

None

Guidehouse is seeking a Databricks Data Scientist to join ourAI & Data team to support client projects involving advanced analytics, machine learning, and data science solutions. This role focuses on working with data to develop models, generate insights, and support data-driven decision-making across teams. The role requires strong Python and SQL skills, analytical thinking, and the ability tocollaboratewith clients and stakeholders to deliver scalable data science solutions.

This position offers virtual work flexibility within the United States. While remote candidates will be considered, preference will be given to candidates located near a Guidehouse office in one of the following markets: Arlington, VA; Washington, DC; New York, NY; Chicago, IL; Austin, TX; Atlanta, GA; Boston, MA; and Boulder, CO.

Please note, this requisition supports hiring across multiple levels to support our Databricks Data Science Teams. The posted salary range represents a range of potential compensation and will vary based on the selected candidate9s experience, qualifications, location, and the level at which the position is filled.

What You Will Do:

  • Develop, train, and evaluate machine learning and statistical models to support business and missionneeds using the Databricks platform.

  • Prepare, clean, andmaintaindatasets for modeling, experimentation, andanalysis.

  • Write,optimize, andmaintainPython and SQL workflows for data exploration, feature engineering, and modeldevelopment.

  • Work with large-scale datasets using Databricks, Spark, and Delta Lakeplatforms.

  • Design reusable feature engineering workflows and model training pipelines using Databricks notebooks, workflows, and MLflow.

  • Register, version, promote, and document models using MLflow Model Registry and Unity Catalog-based model governance practices.

  • Monitor deployed models for performance, drift, data quality, usage patterns, and operational issues; recommend retraining, tuning, or retirement actions as needed.

  • Analyze data toidentifytrends, patterns, and insights to support businessdecisions.

  • Translate business requirements into analytical approaches, models, and data sciencesolutions.

  • Perform data validation, quality checks, and issue resolution to ensure accuracy andconsistency.

  • Collaborate with cross-functional teams including data engineers, analysts, and businessstakeholders.

  • Communicate model outputs, analytical findings, and recommendations to both technical and non-technicalaudiences.

  • Document models, datasets, and methodologies to support reproducibility, transparency, andreuse.

  • Follow data governance, security, and compliance standards within theplatform.

What You Will Need:

  • Bachelor9s degreein computer science, engineering, mathematics, statistics, or another relevant field.

  • 3-8 years of relevant experience in data science, machine learning, or advanced analytics.

  • Strong experience with Python and SQL for data analysis, modeling, and transformation.

  • Experience with Databricks, Spark, Delta Lake, or similar cloud-native data platforms.

  • Hands-on experience designing, building, evaluating, and deploying machine learning models, including experience moving models from prototype to production or production-like environments.

  • Experience with ML lifecycle practices including experiment tracking, model evaluation, model registry, version control, deployment workflows, monitoring, and retraining approaches.

  • Familiarity with model serving patterns, API-based inference, scheduled batch scoring, and integration of model outputs into dashboards, applications, or operational workflows.

  • Experience with data preparation, feature engineering, and model development.

  • Ability to analyze data and communicate insightsclearly.

  • Ability to troubleshoot technical issues, communicate recommendations clearly, and work effectively in team-based delivery environments.

  • Experience supporting AI governance practices, including model documentation, validation, monitoring, version control, and responsible AI considerations.

  • Ability to work across data science, data engineering, cloud, security, and client stakeholder teams to translate analytical prototypes into scalable, maintainable solutions.

What Would Be NicetoHave:

  • 2+ years of hands-on experience with the Databricks platform.

  • Active Databricks Machine Learning Engineer, GenAI Engineer, Data Analyst, or related certification.

  • Experience with Databricks MLflow, Feature Engineering, Feature Store, Model Serving, Workflows, Unity Catalog, Mosaic AI, Vector Search, AI Gateway, or related Databricks AI/ML capabilities.

  • Experience developing GenAI, LLM, RAG, agentic AI, or prompt evaluation workflows using Databricks Mosaic AI, MLflow, open-source frameworks, or cloud-native AI services.

  • Experience with CI/CD, automated testing, code packaging, environment promotion, and source control practices for data science and machine learning workloads.

  • Experience with machine learning frameworks and statistical modeling techniques.

  • Experience with cloud platforms such as Azure, AWS, or GCP.

  • Experience working in project-based or consultingdelivery environments.

  • Familiarity with data modeling, data warehousing, and large-scale data processing concepts.

The annual salary range for this position is $113,000.00-$188,000.00. Compensation decisions depend on a wide range of factors, including but not limited to skill sets, experience and training, security clearances, licensure and certifications, and other business and organizational needs.


What We Offer:

Guidehouse offers a comprehensive, total rewards package that includes competitive compensation and a flexible benefits package that reflects our commitment to creating a diverse and supportive workplace.

About Guidehouse

Guidehouse is a management consulting firm headquartered in Washington, D.C. The firm provides consulting services to clients in the public and commercial sectors, with a focus on energy, financial services, healthcare, national security, and aerospace and defense. Guidehouse was founded in 2018 as a spin-off from PwC. The firm has over 7,000 employees and operates in more than 50 locations worldwide.
Learn more about Guidehouse
Size
8,000 employees
Industry
Founded
2018

Similar Jobs

More Jobs at Guidehouse

More Information Technology Jobs

Find similar Databricks Data Scientist jobs: