Mirantis

Senior Site Reliability Engineer (Golang, Kubernetes)

Mirantis$110K — $130K *
Enterprise Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years of professional experience in DevOps, focusing on cloud and Kubernetes
  • Experience with high-performance data center processing, networking, and storage
  • Familiarity with Golang and experience with languages like Python, JavaScript
  • Strong understanding of distributed systems, microservices architecture, and CI/CD pipelines
  • Exceptional problem-solving and debugging abilities across networking and storage technologies
  • Proven ability to lead technical tasks and collaborate with diverse teams
  • Excellent English communication skills, both written and verbal
  • Willingness to travel up to 25% internationally if necessary

Responsibilities

  • Collaborate with international teams on technical challenges and improvements
  • Develop and maintain cloud and AI infrastructure solutions using open source software
  • Gather and refine technical requirements in collaboration with stakeholders
  • Optimize system performance, reliability, and scalability
  • Troubleshoot and resolve complex technical issues
  • Conduct code reviews to ensure high-quality standards
  • Stay updated with trends in cloud operations and development
  • Design AI-driven automation across the DevOps lifecycle
  • Facilitate knowledge transfer to customers during delivery phases

Benefits

  • Work with an established leader in cloud infrastructure
  • Engage with passionate and talented colleagues across Fortune 500 customers
  • Be part of cutting-edge open-source innovation
  • Thrive in a collaborative and growth-oriented company culture
  • Receive professional development and training opportunities
  • Participate in conferences, hackathons, and tech talks
  • Enjoy company outings and team-building events
Full Job Description
Job Description

We are looking for a highly experienced and driven Senior Site Reliability Engineer to join our forward-thinking cloud development and operations team. In this role, you will contribute to the design, development, and operation of sophisticated cloud-based AI solutions built on the CNCF ecosystem including Kubernetes, running on cutting-edge hardware from leading vendors. This role focuses mainly on deploying AI infrastructure built on NVIDIA-certified hardware, following architecture and implementation designs produced by our engineering team. You will play a pivotal role in ensuring the reliability, security, and performance of container infrastructure, while mentoring team members and Mirantis customers to deliver high-quality software and services. As a senior engineer, you will work closely with stakeholders to define technical strategies, solve complex challenges, and ensure the seamless integration of cloud and software services. This is an excellent opportunity to make a significant impact while driving innovation in a rapidly evolving cloud ecosystem.

Main Responsibilities:
  • Work with geographically distributed international teams on technical challenges and process improvements.
  • Develop, implement, maintain, and troubleshoot cloud and AI infrastructure solutions based on open source software.
  • Collaborate with stakeholders to gather and refine technical requirements.
  • Optimize system performance, reliability, and scalability.
  • Troubleshoot, debug, and resolve complex technical issues.
  • Participate in code reviews to maintain high quality standards.
  • Stay up to date with industry trends and best practices in cloud operations and development.
  • Design and implement AI-driven automation across the DevOps lifecycle, including code development and maintenance.
  • Facilitate knowledge transfer to customers during the delivery phases.


Qualifications
  • 5+ years of professional experience in DevOps, with a strong focus on Cloud, infrastructure technologies and Kubernetes
  • Experience with high-performance data center processing, networking, and storage
  • Exposure to Golang and working knowledge of other programming languages (Python, JavaScript).
  • Strong knowledge of distributed systems, microservices architecture, and CI/CD pipelines.
  • Exceptional problem-solving and debugging skills across networking and storage (hardware and software), Linux, and Kubernetes, with attention to performance optimization and security.
  • Demonstrated ability to lead technical tasks and collaborate effectively with diverse teams.
  • Comfortable making independent judgment calls when working directly with customers, often with limited day-to-day oversight.
  • Excellent written and spoken English.
  • Excellent customer-facing communication skills.
  • A commitment to innovation, continuous learning, and delivering high-quality results.
  • Ability to travel up to 25% if needed, including internationally.

Nice to have
  • Extensive experience in network and/or storage architecture.
  • Experience with high-performance computing or GPU infrastructure: GPU scheduling, MIG/vGPU, RDMA/RoCE or InfiniBand fabrics, NVLink, DCGM health-checking, GPU driver/firmware lifecycle or NVIDIA AI Enterprise.
  • Working experience with Openstack
  • Presence in the open source community including upstream contribution and conference presentations.
  • Prior experience with commercial container and virtual compute infrastructure platforms such as Rancher, Openshift, and VMware.

Education and Experience:
  • Bachelor's degree in Computer Science or a related field, or equivalent experience.
  • At least 5 years of DevOps or Software Development experience or in a similar role.


Additional Information

What does Mirantis offer you?

- Work with an established Silicon Valley leader in the cloud infrastructure industry;
- Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
- Be a part of cutting-edge, open-source innovation;
- Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
- Professional development and training;
- Attend conferences and working groups;
- Company outings, happy hours, hackathons, and tech talks;
- Receive a competitive compensation package with a strong benefits plan.

We are a Leader for Container Management in G2 (#2 after AWS)!

About Mirantis

Mirantis is a software company that provides cloud computing services and solutions. The company was founded in 2011 and is headquartered in Sunnyvale, California. Mirantis offers a range of cloud computing services, including OpenStack, Kubernetes, and Docker. The company's solutions are used by a variety of industries, including telecommunications, finance, and healthcare. Mirantis has over 1,000 employees and offices in the United States, Russia, Ukraine, and the United Kingdom.
Learn more about Mirantis
Size
1,000 employees
Industry
Founded
2011

Similar Jobs

More Jobs at Mirantis

More Enterprise Technology Jobs

Find similar Senior Site Reliability Engineer (Golang, Kubernetes) jobs: