Career Area:
Technology, Digital and Data
Job Description:
Role Definition:
The Simulation Infrastructure Engineer will design, build, automate, and support the compute and platform environments used to develop and execute simulation, digital twin, robotics, and autonomy workloads.
The role combines strong software engineering with cloud, container, GPU, and operational expertise. The engineer will create reliable, repeatable, and observable infrastructure that supports local development, shared environments, demonstrations, continuous integration, and large-scale simulation execution.
Responsibilities:
- Design, implement, and support simulation infrastructure across AWS, Azure, on-premises, and hybrid environments.
- Build reusable infrastructure and platform automation using Kubernetes, Docker, Terraform, Helm, scripting, and configuration management practices.
- Create deployment patterns for simulation services, GPU workloads, NVIDIA Omniverse, Isaac Sim, Unreal Engine, and supporting tools.
- Implement simulation workload scheduling, resource allocation, queueing, isolation, retries, and capacity management across CPU and GPU environments.
- Support GPU and high-performance compute infrastructure, including drivers, container runtimes, resource visibility, scheduling, and performance troubleshooting.
- Develop and maintain CI/CD pipelines for simulation builds, tests, containers, infrastructure, packaging, artifacts, and environment promotion.
- Implement monitoring, logging, metrics, dashboards, alerting, auditability, and operational runbooks for simulation environments.
- Improve platform reliability, availability, scalability, security, cost visibility, and recovery through automation and engineering standards.
- Build self-service tooling and reusable environment templates for simulation developers and test teams.
- Partner with runtime, scenario framework, rendering, digital twin, robotics, autonomy, and QA engineers to resolve deployment and integration issues.
- Support simulation packaging, artifact repositories, dependency management, release workflows, and reproducible environment creation.
- Participate in design reviews, code reviews, incident analysis, root-cause analysis, and continuous improvement activities.
- Document infrastructure architecture, operational procedures, interfaces, dependencies, and support expectations.
Qualifications
- Strong experience with AWS, Azure, Kubernetes, Docker, Linux, networking, CI/CD, and infrastructure automation.
- Programming or scripting experience in Python, Go, C++, Bash, PowerShell, or comparable languages.
- Experience building and operating distributed, cloud-native, or developer platform services.
- Experience with Terraform, Helm, GitHub Actions, Azure DevOps, Jenkins, GitLab CI, Argo CD, or comparable automation tools.
- Knowledge of GPU infrastructure, NVIDIA container tooling, GPU scheduling, HPC, Slurm, Ray, or distributed compute is strongly preferred.
- Experience deploying or supporting NVIDIA Omniverse, Isaac Sim, Unreal Engine, robotics, AI, or simulation workloads is preferred.
- Knowledge of observability, logs, metrics, traces, dashboards, alerting, reliability, and performance engineering.
- Understanding of software packaging, artifact management, container registries, security scanning, release management, and environment governance.
- Strong troubleshooting and root-cause analysis skills across application, container, operating system, network, storage, and infrastructure layers.
- Ability to collaborate across platform, simulation, digital twin, QA, security, and enterprise infrastructure teams.
- Bachelor's degree in Engineering, Computer Science, Information Technology, or a related discipline.
- Typically 5+ years of software, cloud, infrastructure, platform, or simulation engineering experience.
Additional Details:
- This position requires the candidate towork full-time at the Irving, TX office (Dallas)
- Domestic relocation assistance is available for this position.
- Visa sponsorship is available with this position.
Summary Pay Range:
$112,710.00 - $183,140.00
Compensation and benefits offered may vary depending on multiple individualized factors, job level, market location,job-related knowledge, skills, individual performance and experience. Please note that salary is only one component of total compensation at Caterpillar.
Benefits:
Subject to plan eligibility, terms, and guidelines. This is a summary list of benefits.
Medical, dental, and vision benefits*
Paid time off plan (Vacation, Holidays, Volunteer, etc.)*
401(k) savings plans*
Health Savings Account (HSA)*
Flexible Spending Accounts (FSAs)*
Health Lifestyle Programs*
Employee Assistance Program*
Voluntary Benefits and Employee Discounts*
Career Development*
Incentive bonus*
Disability benefits
Life Insurance
Parental leave
Adoption benefits
Tuition Reimbursement
* These benefits also apply to part-time employees
This position requires working onsite five days a week.
Relocation is available for this position.
Visa sponsorship is available for eligible applicants.
Posting Dates: