Engineering Manager, Deep Learning Inference
USonsitemanager
Posted today · via Workday
About this role
NVIDIA is seeking an exceptional Manager, Deep Learning Inference Software, to lead a world-class engineering team advancing the state of AI model deployment. You will shape the software powering today’s most sophisticated AI systems — from large language models to multimodal generative AI — all accelerated on NVIDIA GPUs. The Deep Learning Inference team develops and optimizes open-source frameworks that make AI deployment scalable, efficient, and accessible — including vLLM / SGLang, and FlashInfer. Our work enables developers worldwide to harness NVIDIA accelerators for real-time inference at every scale, from datacenter clusters to edge devices.…
What we'd score you on
reqspace match rubricFive dimensions, recruiter-grade. Upload your resume and we'll generate a written explanation of where you fit and where the gaps are.
1
Skills match
For this role: python, c++, pytorch, teams
2
Level fit
This role is manager-level. We check your trajectory against it.
3
Domain experience
Your work in the role's domain matters more than your years total. We weight recent and direct experience.
4
Recency
A skill you used last quarter weighs more than one from five years ago. We grade on recency, not lifetime.
5
Location fit
This role is based in US. We weight your proximity and willingness to relocate.
Score yourself on this role.
Free · no card · written explanation included
Skills in this role
Pulled from the job description. These are the keywords we'll weight when scoring your fit.
pythonc++pytorchteams
More at Nvidia
- View →Release ManagerIsrael, Raanana
- View →Software Engineer, StorageIsrael, Raanana
- View →ASIC Verification EngineerIndia, Bengaluru
- View →Hardware Test Engineer - Silicon ValidationUS, CA, Santa Clara
- View →Senior Network Engineer, DGX Cloud - BackboneSingapore, Remote
- View →Advanced Development Engineer, AI NetworkingIsrael, Tel Aviv
