Senior deep learning software engineer, inference
Pubblicato il 01-08-2026 - Experteer Italy in Italia
Experteer Overview In this role you design, build, and optimize GPU-accelerated software powering AI applications. You contribute to high-performance deep learning frameworks like SGLang and v LLM, focusing on efficient model serving and inference across datacenter GPUs to edge accelerators. You will implement algorithms and optimize pipelines using NVIDIA tools, collaborating with open-source communities and cross-functional teams. The work centers on advancing inference performance for state-of-the-art LLMs and Generative AI, delivering scalable, production-ready solutions. Retribuzione / Benefits Performance optimization, analysis, and tuning of DL models across LLM, Multimodal, and Generative AI domains Scale DL model performance across NVIDIA accelerator architectures Contribute features and code to inference libraries (v LLM, SGLang, Flash Infer and related solutions)
Collaborate with cross-functional teams across frameworks, NVIDIA libraries, and optimization projects Responsabilità Masters or Ph D or equivalent experience in Computer Engineering, Computer Science, EECS, or AI 5+ years of relevant software development experience Excellent C/C++ programming and software design skills Python experience is a plus Experience with training, deploying, or optimizing DL model inference in production is a plus Background with performance modeling, profiling, debugging, and GPU/CPU architecture knowledge is a plus Requisiti fondamentali highly competitive salaries extensive benefits package diversity, inclusion, and flexibility #J-18808-Ljbffr
