ML Engineer - Austin, TX

Baasi

$120K — $150K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Proficient in Python programming
  • Hands-on with PyTorch and Hugging Face tools
  • Knowledge of LoRA/QLoRA fine-tuning techniques
  • Familiarity with mixed precision training methods
  • Experience in building production-ready training scripts
  • Comfortable in Linux GPU environments
  • Strong collaboration skills with non-ML engineers

Responsibilities

  • Implement and maintain LoRA/QLoRA fine-tuning pipelines
  • Develop logic for incremental training and adapter stacking
  • Automate data preprocessing for user datasets
  • Build training scripts that integrate with orchestration systems
  • Implement monitoring hooks for performance metrics
  • Collaborate with DevOps for training environment reproducibility
  • Write tests to ensure correctness of adapter outputs

Benefits

  • Opportunities for hands-on experience with cutting-edge AI technologies
  • Collaborative working environment with cross-functional teams
  • Flexibility with occasional office presence for team discussions
  • Potential for professional growth in a rapidly evolving field
  • Contributions to productization of machine learning techniques
Full Job Description
About the job ML Engineer - Austin, TX

Role

We are seeking a AI ML Engineer (Python) to help design and implement our AI Pipelines. This is not an academic research role - you will be productizing and automating existing fine-tuning techniques (LoRA/QLoRA) so vendors can train and manage their own adapters with minimal effort.

You'll work closely with backend engineers (Node.js) who orchestrate jobs and dashboards, while you focus on the training pipelines and adapter export logic.

Responsibilities
  • Implement and maintain LoRA/QLoRA fine-tuning pipelines using PyTorch + Hugging Face Transformers + PEFT.
  • Develop logic for incremental training and adapter stacking, producing clean, versioned "delta packs."
  • Automate data preprocessing (tokenization, formatting, filtering) for user-supplied datasets.
  • Build training scripts/workflows that integrate with orchestration backends (Node.js, REST/gRPC, or job queues).
  • Implement monitoring hooks (loss curves, checkpoints, eval metrics) to feed into dashboards.
  • Collaborate with DevOps to ensure reproducible, portable training environments.
  • Write tests to guarantee reproducibility and correctness of adapter outputs.
  • Willingness to occasionally be present in the office for discussions and team collaboration.


Requirements
  • Strong programming skills in Python.
  • Hands-on experience with PyTorch and the Hugging Face ecosystem (Transformers, Datasets, PEFT).
  • Familiarity with LoRA/QLoRA or parameter-efficient fine-tuning methods.
  • Understanding of mixed precision training (FP16/BF16) and memory optimization techniques.
  • Experience building training scripts that are production-ready (reproducibility, logging, error handling).
  • Comfortable working in Linux GPU environments (CUDA, ROCm).
  • Ability to collaborate with backend/frontend engineers who are not ML specialists.


Nice to Have
  • Experience with bitsandbytes, xformers, or flash-attention.
  • Familiarity with distributed training (multi-GPU, NCCL, DeepSpeed, or Accelerate).
  • Prior work in MLOps or packaging ML pipelines for deployment.
  • Contributions to open-source ML libraries.


Similar Jobs

More Jobs at Baasi

More Information Technology Jobs

Find similar ML Engineer - Austin, TX jobs: