Oak Ridge National Laboratory

Senior HPC Engineer - Classified Environment

Oak Ridge National Laboratory$120K — $145K *
Aerospace & Defense
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • BS in computer science, engineering, or related field with 8+ years experience or equivalent education and experience.
  • 5+ years experience in HPC engineering including cluster architecture and performance optimization.
  • Advanced Linux systems engineering with automation skills (Python, Bash, Ansible).
  • Familiarity with HPC workload management tools (e.g., SLURM, PBS).
  • Experience with parallel computing technologies (MPI, OpenMP, CUDA) and high-performance interconnects (InfiniBand).
  • Knowledge of HPC performance measurement and optimization techniques.
  • Understanding cybersecurity requirements for classified computing environments.

Responsibilities

  • Lead design, deployment, and modernization of secure HPC environments.
  • Architect scalable compute solutions using CPU/GPU resources and supporting infrastructure.
  • Establish standards for performance, reliability, security, and scalability.
  • Optimize HPC clusters and analyze performance to identify improvement areas.
  • Develop automation scripts using Python, Bash, and Ansible for operational efficiency.
  • Integrate security measures into HPC systems working with cybersecurity teams.
  • Collaborate with scientists and mission stakeholders to translate technical requirements into solutions.

Benefits

  • Join a collaborative team focusing on technical excellence and innovation.
  • Opportunities for continuous professional growth and meaningful work.
  • An environment encouraging knowledge sharing and diverse perspectives.
  • Work on crucial national security and scientific missions with advanced technologies.
Full Job Description
Requisition Id 16983

Overview:

The Field Intelligence Operations Division (FIOD) of the National Security Directorate (NSSD) is seeking a Senior High Performance Computing (HPC) Engineer for Classified Computing, where you will help shape advanced computing capabilities supporting some of the nation's most important national security and scientific missions.

In this role, you will lead the design, implementation, modernization, and operation of HPC environments built to solve complex computational challenges within highly secure classified environments. You will bring expertise in HPC architecture, cluster management, parallel computing, and performance optimization while helping evaluate and introduce emerging technologies that advance mission capabilities.

You will work alongside a multidisciplinary team of HPC and systems engineers, cybersecurity professionals, scientists, researchers, and mission partners to develop computing solutions that deliver exceptional performance while meeting rigorous security and compliance requirements. This is an opportunity to remain deeply technical while influencing architecture, solving challenging problems, and helping define the future of classified high-performance computing.

You will join a collaborative team that values technical excellence, innovation, service, and shared success. We encourage new ideas, knowledge sharing, and different perspectives, and we are committed to creating an environment where talented people can do meaningful work, continue to grow, and make an impact.

The position requires Top-Secret Clearance with access to Q and SCI.

Major Duties/Responsibilities:

HPC Architecture & Engineering
  • Lead the design, deployment, modernization, and evolution of secure HPC environments supporting classified national security and scientific missions.
  • Architect scalable compute solutions integrating CPU/GPU resources, high-speed interconnects, parallel storage, scheduling, and supporting infrastructure.
  • Establish technical standards, architectures, and engineering approaches that improve performance, reliability, security, and scalability.

Cluster Engineering & Performance
  • Engineer and optimize HPC clusters, including Linux operating systems, workload schedulers such as SLURM, resource management, high-speed fabrics, and parallel file systems.
  • Analyze system and application performance, identify bottlenecks, and optimize compute, memory, storage, networking, and GPU utilization.
  • Lead troubleshooting and root-cause analysis of complex HPC infrastructure and performance issues.

Automation, Security & Modernization
  • Develop and advance automation using Python, Bash, Ansible, Git, and infrastructure-as-code practices to improve deployment, configuration, patching, monitoring, and operational efficiency.
  • Engineer HPC systems to meet applicable NIST, RMF, DISA STIG, DOE, DOW, and classified computing requirements, partnering with cybersecurity teams to integrate security throughout the system lifecycle.
  • Evaluate and introduce emerging HPC, AI/ML, GPU computing, containerization, and infrastructure technologies that advance mission capabilities.

Technical Leadership & Mission Partnership
  • Lead complex HPC initiatives from architecture and design through deployment, optimization, and operational transition.
  • Partner directly with scientists, researchers, engineers, cybersecurity professionals, and mission stakeholders to translate computational requirements into effective technical solutions.
  • Serve as a senior technical resource for HPC architecture and engineering decisions and mentor other engineers through knowledge sharing and technical guidance.
  • Help shape technology roadmaps and continuously improve the performance, resiliency, and capabilities of the classified HPC environment.


Basic Qualifications:
  • BS degree in computer science, engineering, or a related technical discipline and a minimum of eight years of relevant experience. An equivalent combination of education, experience, and certifications may be considered.
  • At least five years of hands-on HPC engineering experience involving cluster architecture, administration, performance optimization, or parallel computing.
  • Advanced Linux systems engineering experience with automation using Python, Bash, Ansible, or comparable technologies.
  • Experience with HPC workload management and scheduling technologies such as SLURM, PBS, or similar platforms.
  • Experience with parallel computing technologies such as MPI, OpenMP, and/or CUDA and high-performance interconnects such as InfiniBand.
  • Experience with HPC performance analysis, monitoring, benchmarking, and optimization.
  • Demonstrated understanding of cybersecurity requirements applicable to classified or highly regulated computing environments, including NIST, RMF, and DISA STIGs.
  • Eligibility for access to Sensitive or Special Access Program information.
  • Current Top-Secret clearance with SCI eligibility.


Preferred Qualifications:
  • Experience designing or operating large-scale HPC, GPU, or AI/ML computing environments.
  • Experience with advanced and parallel storage technologies such as Lustre, GPFS/Spectrum Scale, or BeeGFS.
  • Experience with GPU computing, accelerated workloads, containers, or emerging AI/HPC architectures.
  • Experience with infrastructure automation, configuration management, Git-based workflows, or infrastructure-as-code.
  • Demonstrated ability to lead complex technical efforts, influence architecture, and mentor other engineers while remaining hands-on.
  • Strong analytical, troubleshooting, communication, and collaboration skills.
  • Relevant HPC, Linux, cybersecurity, automation, or infrastructure certifications.
  • A strong technical curiosity and commitment to continuous learning, innovation, and engineering excellence in a rapidly evolving HPC landscape.


Special Requirements:
  • Sensitive Access: Eligibility for access to Sensitive or Special Access Program Information
  • Visa sponsorship: Visa sponsorship is not available for this position.
  • Physical requirements: Work may involve various physical requirements and working conditions.
  • Security, Credentialing, and Eligibility Requirements: Q Clearance with SCI: This position requires the ability to obtain and maintain a Secret Compartmented Information (SCI) clearance from the Department of Energy. As such, this position is a Workplace Substance Abuse (WSAP) testing designated position. WSAP positions require passing a pre-placement drug test and participation in an ongoing random drug testing program. In addition, due to the SCI, you may also be subject to random polygraph testing.


This position will remain open for a minimum of 5 days after which it will close when a qualified candidate is identified and/or hired.

We accept Word (.doc, .docx), Adobe (unsecured .pdf), Rich Text Format (.rtf), and HTML (.htm, .html) up to 5MB in size. Resumes from third party vendors will not be accepted; these resumes will be deleted and the candidates submitted will not be considered for employment.



About Oak Ridge National Laboratory

Oak Ridge National Laboratory (ORNL) is a science and technology national laboratory managed for the United States Department of Energy (DOE) by UT-Battelle. ORNL is the largest science and energy national laboratory in the Department of Energy system by size and by annual budget. ORNL conducts research and development activities in a variety of scientific and technical disciplines. ORNL's scientific programs focus on materials, neutron science, energy, high-performance computing, systems biology and national security. ORNL partners with other national laboratories, universities and industry to solve complex problems and transfer knowledge and technology. ORNL is home to several of the world's most powerful supercomputers, including Summit, the world's most powerful supercomputer as of November 2018.
Learn more about Oak Ridge National Laboratory
Size
5,000 employees
Industry
Founded
1943

Similar Jobs

More Jobs at Oak Ridge National Laboratory

More Aerospace & Defense Jobs

Find similar Senior HPC Engineer - Classified Environment jobs: