Research Engineer
About the role
What We're Looking For
We are seeking a highly skilled Research Engineer to help optimize training and inference workloads running on Lightning AI infrastructure. This role sits at the intersection of ML systems, AI infrastructure, performance engineering, and practical research. You’ll work across models, inference systems, and platform infrastructure to improve performance, scalability, and reliability for real-world AI workloads.
This is a highly cross-functional role that combines deep technical problem solving with hands-on implementation. Successful candidates are comfortable working broadly across the stack — from model behavior and inference systems to distributed infrastructure and developer tooling — while collaborating closely with customers and internal engineering teams to solve complex AI performance challenges.
Location: This role can be based in one of our hubs (NYC, SF, Seattle, or London) or remote, with a minimum of 2 in-office days per week and occasional team and company offsites.
What You'll Do
- Optimize large-scale training and inference workloads across GPUs, accelerators