Direct from source · No middlemen
30 open positions · Updated 1 week ago
30 Tensorrt Llm roles across 12 companies, most in AI/ML; 7% fully remote; typical advertised salary $194k
Who is hiring (12 companies)
Role types
Work arrangement: 2 fully remote · 8 hybrid · 8 on-site · 12 not stated
Advertised salaries: p25 $160k · median $194k · p75 $205.3k (from 18 disclosed annual salaries, USD)
Counts are open roles Joblaze currently tracks on company career pages for this exact skill/location; salaries are advertised minimums, annual, converted to USD.
Showing 20 of 30 positions
Search with filters →Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.
Own Sarvam's production serving path for large distributed models, integrating and optimizing performance across a multi-node stack.
Lead and build a new team focused on developing Jockey Core, a reasoning LLM for video understanding at Twelve Labs.
Own the end-to-end lifecycle of production ML serving systems for a top-performing AI Shopping Agent.
Join Baseten as a Software Engineer to build and optimize large-scale LLM inference systems in a collaborative environment.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.
Join Periodic Labs as an ML Systems Engineer to build and optimize large-scale training and reinforcement learning infrastructure.
Own the inference systems that power frontier AI models in production and research at a tech-first startup.
Drive the adoption of AI runtime services at CoreWeave, leveraging your expertise in distributed systems and AI infrastructure.
Join CoreWeave as a Staff Software Engineer to lead the development of a Kubernetes-native inference platform for AI workloads.
Join CoreWeave as an Applied AI Engineer to enhance the performance of our inference platform through benchmarking and optimization.
Join SpaceX as a Software Engineer to develop high-performance AI inference systems for mission-critical applications.
Join Together AI as a Forward Deployed Engineer to optimize inference systems for strategic customers in a hands-on role.
Join Fireworks AI as an AI Field Engineer to build production systems for generative AI with large organizations across EMEA.
Join Fireworks AI as a senior AI Field Engineer to build production systems for generative AI with leading organizations.
Join Fireworks AI as a senior AI Field Engineer to build production systems and engage with enterprise customers on generative AI solutions.
Join Fireworks AI as a senior AI Field Engineer to build production systems for innovative AI-native companies.