Senior Infrastructure Engineer
Posting Start Date: 8/3/26
Job Location (Short): Tucson, Arizona, USA, 85706
Requisition ID: 36532
Onsite or Remote: Onsite Position
Job OverviewThe Senior Infrastructure Engineer - DevOps & Site Reliability Engineering is responsible for the architecture, implementation, and reliability of Komatsu's mission-critical infrastructure supporting the Autonomous Haulage System (AHS) and other mining technology solutions. This role ensures the scalability, observability, and resilience of infrastructure both in development and in globally deployed production environments.
Working cross-functionally with software engineering, security, IT, and field operations, the Senior Infrastructure Engineer designs and automates infrastructure systems to maximize reliability and performance. They will also mentor peers, establish best practices, and drive continuous improvement of infrastructure platforms that support 24/7 mine operations worldwide
Job Responsibilities- Serve as a technical leader and subject matter expert across multiple infrastructure and DevOps domains.
- Develop and optimize automated CI/CD pipelines to support agile development and fast, reliable deployments.
- Architect and maintain on-premise virtual environments, including provisioning, monitoring, and capacity planning.
- Design and implement secure, scalable, and resilient infrastructure solutions using Infrastructure as Code (IaC) principles (e.g., Terraform, Ansible, Packer).
- Implement observability frameworks for end-to-end system visibility, integrating metrics, tracing, and logging (e.g., Prometheus, Grafana, ELK, OpenTelemetry).
- Define and monitor Site Reliability Engineering (SRE) metrics, including Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to guide operational improvements.
- Lead incident response and root cause analysis for production and field systems, ensuring lessons learned are implemented and systems are hardened.
- Partner with field engineers and support teams to ensure robust infrastructure at customer mine sites, helping to maintain 24/7 high availability and reliability.
- Continuously improve infrastructure in field-deployed environments by applying automation, repeatability, and security best practices.
- You will assist in the career development of others, actively mentoring individuals and the team on advanced technical issues and help managers guide the career growth of their team members
- You will balance technical leadership and savvy with strong business judgment to make the right decisions about technology choices.
- You will be able to occasionally travel internationally.
Required Skills- 7+ years of experience in Infrastructure Engineering, DevOps, or SRE roles, with a track record of leading major initiatives.
- Experience building and maintaining CI/CD pipelines using GitHub, Jenkins, GitLab CI, or similar tools.
- Proven experience managing on-premise and hybrid cloud infrastructure, including hypervisors (VMware, KVM).
- Deep understanding of networking, system security, and Linux-based systems.
- Familiarity with SRE disciplines including SLIs, SLOs, incident management, and error budgets. Experience with observability and monitoring platforms (e.g., Prometheus, Grafana, ELK, Datadog).
- Strong background supporting field-deployed infrastructure and maintaining mission-critical systems with high uptime.
- Experience as a mentor, tech lead or leading an engineering team
- Bachelor's or Master's degree in Computer Science, Engineering, or equivalent experience.