Group/Division
Enabling the movement toward advanced chip design, KLA's Measurement, Analytics and Control group (MACH) is looking for the best and brightest research scientists, software engineers, application development engineers and senior product technology process engineers to join our team. The MACH team's mission is to collaborate with our customers to innovate technologies and solutions that detect and control highly complex process variations—at their source—rather than compensate for them at later stages of the manufacturing process. With over 40 years of semiconductor process control experience, chipmakers around the globe rely on KLA to ensure that their fabs ramp next-generation devices to volume production quickly and cost-effectively. Our MACH team develops leading-edge solutions for patterning process analytics and control technologies, thereby providing customers with critical insight at the feature level, field level and cross-wafer analysis. Our teams also develop advanced modeling simulation, data analytics and process control modeling technologies. As a member of the MACH team, you’ll be joining the most sophisticated and successful process-control company in the semiconductor industry--working across functions to solve the most complex technical problems in the digital age.
Job Description/Preferred Qualifications
In this exciting role you will design and architect next-generation High Performance Computing platforms supporting semiconductor manufacturing, AI workloads, computational lithography, metrology, inspection, and large-scale distributed data processing environments. Lead architecture decisions spanning software, hardware, networking, storage, GPU acceleration, virtualization, and cloud-integrated HPC infrastructure. Drive technology strategy, roadmap planning, and cross-functional alignment from concept through field deployment
Responsibilities will include: HPC Platform Architecture- Define architecture for enterprise-scale HPC clusters.
- Develop compute, network, storage, and virtualization strategies.
- Lead platform scalability planning from tens to thousands of cores.
- Define standards for GPU acceleration, distributed computing, and workload orchestration.
- Establish system performance entitlement targets and capacity planning methodologies.
Software Architecture- Design distributed software services supporting HPC infrastructure.
- Define APIs, automation frameworks, configuration management architecture, and observability solutions.
- Lead adoption of software engineering best practices including CI/CD, testing, infrastructure-as-code, and DevOps methodologies.
- Drive modernization initiatives across Linux, container, and cloud-native platforms.
Technical Leadership- Provide technical leadership across software, systems, manufacturing, operations, and field organizations.
- Lead architecture reviews and design reviews.
- Mentor engineers across multiple levels.
- Establish coding, deployment, and operational standards.
- Serve as escalation leader for critical system issues affecting customers and manufacturing operations.
Product Strategy- Define HPC platform roadmap.
- Evaluate emerging technologies in:
- GPU computing
- AI infrastructure
- Storage architectures
- Virtualization
- Cloud integration
- High-speed networking
- Partner with product management and business leadership on long-term HPC strategy.
Required skills and experience:Expert Linux knowledge:- RHEL
- Rocky
- AlmaLinux
- Ubuntu
- SUSE
HPC Technologies- SGE
- SLURM
- Kubernetes
- Distributed computing architectures
- NUMA optimization
- MPI concepts
- GPU workload scheduling
Infrastructure- NVIDIA GPU ecosystems
- High-speed Ethernet
- RDMA
- NVMe storage
- Parallel file systems
- NAS/NFS architectures
Virtualization- VMware
- Proxmox
- KVM
- Container technologies
Software Development- Python
- C++
- Bash
- Git
- CI/CD pipelines
- Ansible
- Infrastructure automation frameworks
System Architecture- Service-oriented design
- Scalability engineering
- Reliability engineering
- High availability
- Disaster recovery
- Capacity modeling
- Performance benchmarking and tuning
Minimum Qualifications
Doctorate (Academic) Degree and related work experience of 5 years; Master's Level Degree and related work experience of 8 years; Bachelor's Level Degree and related work experience of 12-15 years
Base Pay Range: $186,200.00 - $316,500.00 Annually
Primary Location: USA-CA-Milpitas-KLA
KLA’s total rewards package for employees may also include participation in performance incentive programs and eligibility for additional benefits including but not limited to: medical, dental, vision, life, and other voluntary benefits, 401(K) including company matching, employee stock purchase program (ESPP), student debt assistance, tuition reimbursement program, development and career growth opportunities and programs, financial planning benefits, wellness benefits including an employee assistance program (EAP), paid time off and paid company holidays, and family care and bonding leave.
Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level, and location. The range displayed reflects the pay for this position in the primary location identified in this posting. Actual pay depends on several factors, including state minimum pay wage rates, location, job-related skills, experience, and relevant education level or training. We are committed to complying with all applicable federal and state minimum wage requirements where applicable. If applicable, your recruiter can share more about the specific pay range for your preferred location during the hiring process.