10+ years of hands-on experience operating, crafting or engineering large-scale computing environments, such as HPC, HTC or BC
Drive innovative computational solutions and exploit emerging technologies
Experience of administration of large-scale cluster and server computing and related software (e.g. Slurm, LSF, Grid Engine)
Hands-on experience working in a DevOps team and using agile methodologies
Operating and consuming virtualized private cloud resources (e.g. OpenStack)
Understanding of Linux system administration, the TCP/IP stack, and storage subsystems
Experience in implementing and administering large-scale parallel filesystems (e.g. Weka, GPFS, Lustre)
Proven experience of using configuration management (e.g. ansible, salt, puppet) and technology frameworks in IT operations
Experience of developing and managing relationships with 3rd party suppliers
Scripting and tool development for HPC & DevOps style platform operations using bash and Python
Scientific degree, and/or experience in computationally intensive analysis of scientific data
Previous experience in high performance computing (HPC) environments, especially at large scales (>10,000 cores)
Operation and configuration of public cloud computing infrastructure (e.g. AWS, Azure, GCP) is a plus
Managing a virtualized private cloud environment (e.g. OpenStack) is a plus
Container technology (e.g. LXD, Singularity, Docker, Kubernetes) is a plus
Demonstrated development experience with a variety of programming languages, tools, and technologies (Java/C++, Python/Ruby/Perl, SQL, AWS) is a plus
Experience with Hashicorp tools like terraform, vault, consul and nomad is a plus
Working experience with high-speed networks (e.g. InfiniBand)