Job Summary
The Senior HPC Systems Engineer will play a vital role in enhancing the existing on-premises systems of the laboratory by integrating burst capacity solutions. Working in a small engineering firm, this position involves collaborating closely with a dynamic team to meet project priorities and client needs.
Responsibilities
- Optimize and manage the laboratory’s on-premises HPC systems to improve performance and scalability.
- Implement burst capacity solutions to accommodate fluctuating workloads.
- Collaborate with team members to align engineering projects with organizational priorities.
- Utilize DevOps practices to enhance infrastructure management.
- Oversee production systems, ensuring robust operation and maintenance.
- Administer out-of-band management for system reliability.
Qualifications
- A minimum of 10 years of experience in operating and managing production systems.
- Proficient in out-of-band management techniques.
- Solid background in DevOps methodologies with a focus on infrastructure.
- Strong problem-solving skills and ability to work collaboratively in a team setting.
- Experience in high-performance computing environments is preferred.


