Harvard University
As a High-Performance Computing (HPC) Engineer, you will support the implementation, operation, and lifecycle management of secure and scalable HPC environments that enable computational research across HMS. Working as part of the Research Computing Infrastructure team, you will contribute to the provisioning and administration of compute clusters, workload scheduling systems such as Slurm, user-facing software environments, and secure platforms that meet institutional compliance needs. This role emphasizes hands-on technical execution, operational reliability, and collaboration with colleagues and researchers to support evolving HPC workflows and infrastructure. 
Core Duties:
• Perform provisioning, configuration, and decommissioning of HPC compute clusters.
• Support the administration and tuning of workload schedulers (e.g., Slurm) to ensure efficient job management and cluster utilization.
• Help maintain secure, regulated compute environments (e.g., NIST 800-171).
• Contribute to the integration of user accounts and identity management with institutional systems.
• Maintain and optimize user-facing software environments, including module systems and containerized applications.
• Support development and maintenance of scripts, automation, and tools used in cluster operations.
• Monitor system health, respond to alerts, and assist with compliance reporting and documentation.
• Collaborate with team members and researchers to troubleshoot and improve the computing environment.
• Contribute to operational documentation and support knowledge-sharing across the team.
• Participate in off-hours on-call rotation.
• Perform other duties as assigned.
Real, currently open roles at Harvard University, sourced from their public smartrecruiters careers page.
Industry
Other
Qualdoc
Not disclosed
Today