Systems Engineer position at North Dakota State University

We are seeking an experienced Systems Engineer to join our High-Performance Computing (HPC) Systems team. As a Systems Engineer at CCAST, you will be responsible for architecting, deploying, and maintaining advanced research data storage systems. Therefore, the ideal candidate will be proficient in the Linux command line and be familiar with key Linux-based storage technologies, such as RAID, SCSI, LVM, etc. Given the HPC focus of the center, some familiarity with other areas of the HPC systems stack are also desired, such as fast networking, job/task scheduling frameworks, parallel and accelerated computing, etc. Experience with HPC filesystems directly, such as GPFS, Lustre, BeeGFS, Ceph, or others, is also a plus.

NDSU’s Center for Computationally Assisted Science and Technology (CCAST) provides advanced computing infrastructure for research and education at NDSU and beyond. We maintain the largest academic supercomputing facility in the state of North Dakota, with more than 12,000 CPU cores, 70 GPUs, and 10 petabytes of research data storage capacity, and provide rigorous training and internship programs in advanced research computing. We continually work to enhance NDSU’s capabilities and competitive edge in disciplines and research that rely on advanced computing.

Additional information is available here.

Location: Fargo, ND (hybrid options may be available)

Advertised Salary:

  • $70,000+ commensurate with experience for a junior level
  • $90,000+ commensurate with experience for a senior level.

Minimum Qualifications:

  • Bachelor’s degree or equivalent.
  • 5+ years’ experience maintaining and administering Linux systems in a production environment.
  • Proficiency in one or more scripting/programming languages (e.g. Bash, Korn, Python, Perl, C/C++, etc.).

Preferred Qualifications:

  • 10+ years’ experience maintaining and administering Linux systems in a production environment.
  • Experience working specifically with high-performance computing clusters.
  • In-depth knowledge of and experience with fundamental Linux storage technologies, including software (i.e. md) RAID, Storage Area Networking (SAN), SCSI, Fibre Channel, etc.
  • Experience managing large quantities of data (100s TB to multiple PB).
  • Experience working with higher-level data management systems, such as databases (SQL and NoSQL), data lakes, object stores, etc.
  • Experience creating, maintaining, and optimizing various filesystem formats (i.e. ZFS, NFS, GPFS etc.) and storage protocols (SMB, FcoE, etc.) in multi-tiered, high-availability storage architectures.
  • Experience integrating on-prem clusters with cloud computing resources or next-generation computing, storage, or network resources.
  • Experience tuning/optimizing networking fabrics such as InfiniBand, Slingshot, and Ethernet in large, diverse (multi-OS), HPC cluster environment.
  • Demonstrated leadership abilities and project management skill.
  • A US Citizen or US Permanent Resident—persons in this position may be required to handle export controlled or other sensitive information.

APPLY HERE

Closing date: 09/12/23

Happy to answer questions about this position specifically, or what it’s like to work at North Dakota State University generally! Send me a message or reply.