Research HPC Systems Administrator
Not SpecifiedBookmark Details
Internal or External Search: External – Open to all applicants
Advertising Summary: Join
New Mexico State University as our High-Performance Computing (HPC) Systems Administrator and play a critical role in advancing
groundbreaking research across diverse disciplines. In this dynamic position, you’ll manage and optimize cutting-edge computing
infrastructure, support researchers tackling complex computational challenges, and help shape the future of scientific discovery. If you’re
passionate about automation, large-scale computing environments, and empowering innovation through technology, this is your opportunity to
make a lasting impact.
Position Details
Position Title: Research HPC
Systems Administrator
College/Division: Information Technology
Department: 450280-IT
SYSTEM ADMINISTRATION
Location: Las Cruces
Offsite Location (if applicable):
Target Hourly/Salary Rate: $67,161.67 – To commensurate with experience
Appointment Full-time
Equivalency: 1.0
FLSA Status: Exempt
Bargaining Unit Announcement: This is NOT a
bargaining unit position with American Federation of State, County & Municipal Employees (AFSCME).
Contingent Upon Funding:
Contingent upon funding
Standard Work Schedule: Standard (M-F, 8-5)
If Not a Standard Work
Schedule:
Job Duties and Responsibilities: Administer, maintain, and optimize NMSU’s research
high-performance computing (HPC)
cluster, including compute nodes, login nodes, storage systems, networking interfaces, and
supporting services. Manage and tune the Slurm workload manager – job scheduling,
partitions, QoS settings, node configurations,
troubleshooting, and end-user support. Oversee
and maintain parallel file systems (PanFS), ensuring reliability, performance, and data
integrity.
Monitor system performance and resource utilization, and identify and resolve performance
bottlenecks. Perform software
installation and environment management (modules, conda,
Spack) and apply system and security updates using modern automation tools
(e.g., Ansible,
Puppet, Terraform). Develop documentation, training materials, and workshops to help
researchers adopt the cluster
effectively, and assist with building, deploying, and scaling
containerized workloads (Singularity/Apptainer, Docker). Ensure system
security, compliance,
backups, monitoring, and data-protection best practices, and assist with integrating the HPC
environment into
campus identity, networking, monitoring, and storage systems. Contribute to
long-term research-computing capacity planning and the
procurement of new technology.
Qualifications
Required Education and
Experience:
Associate’s Degree + 9 years of relevant experience or a Bachelor’s degree + 7 years of relevant experience.
Master’s degree or higher preferred.
Equivalent Qualifications:
Equivalent combinations of education and
experience will be considered. Relevant research computing experience, internships, professional training, and industry-recognized computing
or information technology certifications may be substituted for portions of the required experience, as appropriate.
Preferred Qualifications:
Special Certification/Licensure:
Working
Conditions and Physical Effort
Environment: Work is normally performed indoors.
Physical Effort: No or very limited physical effort required.
Lifting Requirements:
Requires handling of average-weight objects up to 10 pounds or some standing or walking.
Risk: Work
environment involves some exposure to hazards or physical risks, which require following basic safety precautions.
Share
Facebook
X
LinkedIn
Telegram
Tumblr
Whatsapp
VK
Bluesky
Threads
Mail