What Your Job Will Be LikeWe are seeking an Information Systems Architect to join a team that administers next-generation high performance computing (HPC) Architectures and develops innovative solutions for operations and improved utilization. Your work will enhance the effectiveness of vitally important experimental research in support of both our national defense and broader scientific discoveries.
On any given day, you may be called on to:
- Engage in all aspects of the HPC system lifecycle including facility integration, standup, acceptance testing, performance bench marking, operational monitoring and support, upgrades, and reclamation.
- Explore, develop, and deploy advanced operation technologies to improve the efficacy of our HPC systems and the administrators who support them.
- Perform all aspects of security, networking, filesystems, system software installation, and user support work.
- Troubleshoot software and hardware and replace defective components.
- Support research and development staff to deliver functional platforms for pre-production systems.
* The position allows a combination of onsite and offsite work. The candidate must live within 50 miles of the assigned work location and be available to work onsite a minimum of three days per week as necessary.
Salary Range$102,400 - $199,700
*Salary range is estimated, and actual salary will be determined after consideration of the selected candidate's experience and qualifications, and application of any approved geographic salary differential.
Qualifications We Require- Bachelor's degree in Computer Science, Computer Engineering, Information Systems Engineering (CIS/MIS), or a related STEM field plus 5 years of experience; or equivalent (AS + 9 years or no degree + 13 years).
- Included in the experience above:
- 5 years of experience with high performance computing (HPC) systems or large scale cloud deployments based on Linux / Unix operating systems.
- Ability to obtain and maintain a DOE Q clearance.
Qualifications We Desire- Experience managing multiple systems or a variety of system architectures.
- Experience with configuration management technologies such as Ansible or Terraform.
- Experience with cloud infrastructure and container technologies such as Docker, Kubernetes, or Singularity.
- Experience with Python, C, and/or FORTRAN programming languages.
- Strong communication skills and experience working in a team environment.
About Our TeamThe HPC Development Department (9328) partners with different scientific and computing disciplines at Sandia, and externally, to advance high performance computing (HPC) and operations. The Observation, Orchestration, & Optimization Team supplies the HPC and Cloud monitoring capability for Sandia National Laboratory to ensure actionability to performance or system anomalies, uses derived knowledge to inform workload management decisions and find opportunities to improve throughput and efficiencies, and innovate the future of HPC and Cloud Computing. The team supports production monitoring infrastructure and tools and implements innovative solutions to support cutting-edge computing technologies.
Posting DurationThis posting will be open for application submissions for a minimum of three (3) calendar days, including the 'posting date'. Sandia reserves the right to extend the posting date at any time.