Nebius
HPC Specialist Solutions Architect
engineeringfull-timeRemote - United States
SALARY
Not listed
WORK TYPE
remote
JOB TYPE
full-time
INDUSTRY
ai
✦ AutoApply Sick of applying? We apply to roles like this for you, up to 20 a month.
Learn more
About the role
The role
We are seeking a Specialist HPC Infrastructure Solutions Architect to design, build, and optimize next-generation high-performance computing (HPC) platforms for AI, simulation, and large-scale data processing workloads. The ideal candidate combines deep knowledge of cloud-native architecture, Kubernetes orchestration, networking, and HPC system design with hands-on experience implementing NVIDIA GPU-based compute environments and MLOps toolchains. This role sits at the intersection of infrastructure engineering, accelerated computing, and AI systems design, shaping the foundation for high-throughput, low-latency distributed workloads in cloud environment.
You’re welcome to work remotely from the United States or Canada.
Your responsibilities will include:
- Architect and implement scalable HPC clusters optimized for AI, simulation, and distributed training, leveraging container orchestration frameworks and schedulers (e.g., Kubernetes, Slurm).
- Design and integrate GPU-accelerated compute infrastructures featuring NVIDIA Hopper, Blackwell architectures, NVLink/NVSwitch, and InfiniBand/RoCE Interconnects.
✦ Sick of applying to 40 jobs a month?
I rewrite your resume for ATS by hand first. Once you sign off on it, AutoApply applies to up to 20 roles like this a month, cover letter in your own voice each time. From $14.99/mo, cancel anytime.
Get AutoApply