Virtualization & Orchestration EngineerLocation: Hybrid | Bellevue, WA Area
Titles: Intermediate, Senior and Staff (multiple roles available)
The OpportunityThis is a foundational engineering role within the company's largest infrastructure engineering organization. You'll help design and build the systems that make GPU capacity available, scalable, secure, and reliable across a multi-tenant AI cloud platform.
You'll work at the intersection of virtualization, Kubernetes, distributed systems, GPU infrastructure, and high-performance computing-solving complex challenges around workload scheduling, resource allocation, cluster management, and infrastructure automation.
This opportunity is ideal for engineers who enjoy building large-scale platforms from the ground up and owning critical infrastructure systems end-to-end.
What You'll Do- Design and build virtualization infrastructure supporting GPU-intensive AI and HPC workloads.
- Develop and operate Kubernetes-based orchestration systems for GPU cluster provisioning and workload scheduling.
- Build automated provisioning systems that enable GPU capacity to be allocated, scaled, and reclaimed efficiently across multiple tenants.
- Design solutions for workload placement, resource management, and cluster lifecycle operations.
- Partner closely with hardware, networking, infrastructure, and AI platform teams to ensure orchestration systems align with real-world cluster architectures and constraints.
- Improve the reliability, security, scalability, and operational maturity of the orchestration platform.
- Build tooling and automation that simplifies infrastructure management and improves developer and customer experiences.
- Contribute to architectural decisions, engineering standards, and best practices as the platform evolves.
What We're Looking For- Strong hands-on experience with Kubernetes and container orchestration in production environments.
- Experience designing, building, and operating large-scale infrastructure platforms.
- Background with virtualization technologies supporting cloud, HPC, GPU, or distributed computing environments.
- Understanding of GPU cluster provisioning, workload scheduling, and resource management.
- Experience with Linux-based infrastructure and distributed systems concepts.
- Ability to independently own complex systems from design through production operation.
- Comfortable working in a fast-moving environment where architecture and processes are being established.
Preferred Qualifications- Experience with GPU scheduling technologies such as Slurm, Kubernetes device plugins, NVIDIA GPU Operator, or similar frameworks.
- Experience supporting AI infrastructure, machine learning platforms, HPC environments, or GPU cloud providers.
- Background building multi-tenant infrastructure platforms for cloud providers or large-scale compute environments.
- Experience with infrastructure automation, Infrastructure as Code, and platform engineering practices.
- Familiarity with high-performance networking and GPU cluster architectures.
Compensation- Competitive base pay for Bellevue market
- Certain roles are eligible for additional rewards, including merit increases, annual bonus, and stock. These awards are allocated based on individual performance
- U.S. based employees have access to medical, dental, and vision insurance, a 401(k) plan and company match, employees also receive per calendar year, paid holidays.
Location- Hybrid role based in the Bellevue, WA area.
- Approximately three days per week in the office.
- Candidates elsewhere in the U.S. who are open to relocation are encouraged to apply.
- U.S. work authorization is required. Visa sponsorship is not currently available.