IC Resources is supporting a specialist technology organisation requiring an experienced HPC Infrastructure Engineer for a focused infrastructure improvement project.
The successful engineer will take technical ownership of an existing on-premise HPC/compute environment, assess the current infrastructure and implement improvements across reliability, manageability, monitoring and automation.
This would suit a hands-on Linux/HPC engineer who is comfortable entering an existing environment, identifying weaknesses and leaving the infrastructure in a more robust and supportable state.
Key experience:
-
Linux HPC / compute infrastructure
-
HPC cluster administration
-
Linux server and node management
-
Cluster troubleshooting and remediation
-
Infrastructure monitoring and management
-
Networking and storage within compute environments
-
Infrastructure automation
-
Improving resilience and reliability
-
Technical documentation
-
Knowledge transfer and handover
Experience with workload schedulers, infrastructure-as-code, configuration management, high-performance networking or distributed storage would be beneficial.
We’re particularly interested in engineers who have previously delivered HPC cluster upgrades, infrastructure remediation, migrations or short-term improvement programmes.
If this opportunity could be of interest, please contact Bradley Wilson at IC Resources for more information.