3008 HPC Cluster Management
Indexed description
Location: Remote
Experience Level: 5+ years
English: B2–C1
Full-time | Engineering
✈️ Valid visa is mandatory
Responsabilities
- Maintain and Optimize parallel job management
- Maintain system software and firmware revisions, including patches, updates, and OS upgrades
- Solve system hardware and software issues
- Implement solutions, repairs and workarounds
- Manage software issues for both the system and user applications
- Support users when they encounter issues
- BSc in Computer Science, Engineering or related area
- 5+ years of Linux Administration
- 2+ years of HPC Management
- - Job scheduling
- - Software installation
- - User support
- AWS ParallelCluster experience
- 1+ years of Python Scripting experience
- 1+ years of Bash Scripting experience
- Familiarity with CAE software
- Good organization
- Strong communication
- Strong English
- PBS experience
- SLURM experience
- SOCA
- Professional growth and development (Both in Mexico, Canada and Europe as well as the United Kingdom, Spain)
- Indeterminate Contract
- Hybrid scheme
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search