NVIDIA is seeking highly skilled software engineers to build AI inference systems that scale large models efficiently across multi-GPU, multi-node, and multi-cloud environments. You will architect and implement high-performance inference stacks, optimize GPU kernels and compilers, and collaborate across inference, compiler, scheduling, and performance teams.
You’ll drive industry benchmarks, contribute to vLLM, SGLang, and related tooling, and publish research that advances ML systems.
📌 Senior AI Inference Systems Engineer (Italia)
🏢 NVIDIA
📍 Italia