Human role

Senior Performance Engineer - LLM Inference Frameworks

NVIDIA

Israel, Yokneam workday 4mo ago
Apply now

The role

Job description

NVIDIA is hiring exceptional software engineers to build and optimize the core inference infrastructure for large language models. Join the TensorRT‑LLM team - the group defining how generative AI performs at global scale on NVIDIA GPUs. We’re looking for engineers who love squeezing every drop of throughput, memory efficiency, and scalability out of modern model runtimes. Your work will directly shape the frameworks behind state‑of‑the‑art LLM inference used across NVIDIA and the AI community.

View full posting

Index terms

Skills & keywords

LLMGenerative AI

Get roles like this in your inbox

New agentic AI jobs, curated every Thursday. No spam.

Explore more

All LLM Engineer jobs All jobs at NVIDIA
NVIDIASenior Performance Engineer - LLM Inference Frameworks
Apply