Google

Software Engineer, On-Device Machine Learning

Google$147K — $211K *
Consumer Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree or equivalent practical experience.
  • 2 years of software development experience in one or more programming languages, or 1 year with an advanced degree.
  • 2 years of experience with ML infrastructure, including model deployment and optimization.
  • Experience with runtimes and performance tuning for software applications.
  • Experience in mobile development, particularly relevant for on-device applications.

Responsibilities

  • Collaborate with peers and stakeholders through design and code reviews to ensure adherence to best practices.
  • Implement solutions in key ML areas while contributing to model optimization and data processing.
  • Develop LiteRT, Google’s on-device AI framework to enable advanced hardware acceleration.
  • Enable on-device deployment of significant ML models across various platforms and accelerators.
  • Enhance performance of on-device model inference through optimizations in representation and runtime implementation.

Benefits

  • Comprehensive health insurance including medical, dental, and vision coverage.
  • Generous retirement plans and savings programs.
  • Flexible work hours and remote work opportunities.
  • Access to wellness programs and mental health support.
  • Paid parental leave and family care assistance.
Full Job Description
Minimum qualifications:
  • Bachelor's degree or equivalent practical experience.
  • 2 years of experience with software development in one or more programming languages, or 1 year of experience with an advanced degree.
  • 2 years of experience with ML infrastructure (e.g., model deployment, model evaluation, optimization, data processing, debugging).
  • Experience with runtimes and performance tuning.
  • Experience in mobile development.

Preferred qualifications:
  • Master's degree or PhD in Computer Science or related technical fields.
  • Experience in leading and delivering successful ML projects focused on on-device deployment (Android, iOS, web browsers, or embedded devices).
  • Experience in ML frameworks (e.g., PyTorch, JAX, TensorFlow).
  • Experience with on-device ML SDKs/tooling (e.g., TensorFlow Lite, ExecuTorch, Core ML, SNPE/QNN).
  • Strong understanding of Generative AI model architectures and their optimization for on-device execution.
  • Passion for innovation and a strong desire to push the boundaries of what's possible with on-device ML.


About the job

LiteRT is Google's next generation on-device AI framework, succeeding TensorFlow Lite (TFLite). It is designed to maximize the performance, efficiency, and portability of ML models on a wide array of edge devices, from mobile phones to embedded systems. LiteRT significantly upgrades GPU acceleration and introduces native NPU acceleration, while maintaining and enhancing the robust CPU performance inherited from TFLite.

LiteRT enables developers and Google products to deploy AI across mobile, web, desktop, and embedded. Our team focuses on building cross-platform infrastructure aligned with Google's business needs, serving top Google products (Android, Chrome, Photos, Meet, YouTube, etc.), third-party developers, and specialized Pixel solutions. Our goal is to provide on-device AI infrastructure with exceptional performance, enabling framework and device flexibility at scale.

Google Cloud accelerates every organization's ability to digitally transform its business and industry. We deliver enterprise-grade solutions that leverage Google's cutting-edge technology, and tools that help developers build more sustainably. Customers in more than 200 countries and territories turn to Google Cloud as their trusted partner to enable growth and solve their most critical business problems.
Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $147000 - $211000 (USD) 15% bonus target equity benefits

Learn more about benefits at Google .

Responsibilities
  • Collaborate with peers and stakeholders through design and code reviews to ensure best practices amongst available technologies (e.g., style guidelines, checking code in, accuracy, testability, and efficiency).
  • Implement solutions in one or more specialized ML areas, utilize ML infrastructure, and contribute to model optimization and data processing.
  • Develop LiteRT, Google's on-device AI framework for first- and third-party, enabling SOTA hardware acceleration and use cases on edge platforms.
  • Enable on-device deployment of key models, such as Gemini Nano and Gemma, across various accelerators (GPU/Pixel TPU/NPUs/CPU) on Android, Chrome, iOS, desktop, and more.
  • Improve performance of on-device model inference via optimizations in the model representation, on-device runtime and kernel implementation.


Information collected and processed as part of your Google Careers profile, and any job applications you choose to submit is subject to Google's Applicant and Candidate Privacy Policy .

About Google

Google is a multinational technology company that specializes in Internet-related services and products. These include online advertising technologies, search engine, cloud computing, software, and hardware. Google was founded in 1998 by Larry Page and Sergey Brin while they were Ph.D. students at Stanford University. The company has grown tremendously since then and has become one of the most valuable companies in the world. Google's mission is to organize the world's information and make it universally accessible and useful.
Learn more about Google
Size
156,500 employees
Market Cap
$1,115.4 billion
Industry
Net Income
$40.2 billion
Founded
1998
5 Year Trend
+23.3%
Revenue
$182.5 billion
NASDAQ

Similar Jobs

More Jobs at Google

More Consumer Technology Jobs

Find similar Software Engineer, On-Device Machine Learning jobs: