"llm inference" Jobs

190 open tech roles matching “llm inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 190 results

Handshake

Join Handshake as a Senior Software Engineer to build scalable ML infrastructure for a fast-growing AI data business.

Handshake San Francisco, CA Published 2 months ago
Flexible on stack
AssemblyAI

Join AssemblyAI as a Senior Design Engineer to shape and build exceptional product experiences in a fast-paced AI environment.

AssemblyAI Denver $180k–$240k/yr Published 3 weeks ago
Flexible on stack
AssemblyAI

Join AssemblyAI as a Senior Software Engineer to design and build exceptional product experiences in a fast-paced, meritocratic environment.

AssemblyAI United States $180k–$240k/yr Published 1 month ago
Flexible on stack
Arena

Explore and analyze large datasets to uncover insights about AI model behavior in a mission-driven team.

Arena Bay Area Published 2 weeks ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 2 months ago
70% coding
Sesame

Join Sesame as a Research Engineer to innovate in NLP, Speech, and Computer Vision with a focus on deep learning.

Sesame San Francisco Published 2 months ago
Flexible on stack
Doctronic

Own full-cycle recruiting for exceptional AI and engineering talent at a startup focused on transforming healthcare.

Doctronic New York City Published 3 months ago
Flexible on stack
Fireworks AI

Lead the development of lifecycle marketing strategies to enhance customer engagement and drive growth at Fireworks AI.

Fireworks AI San Mateo Published 1 week ago
Flexible on stack
AssemblyAI

Join AssemblyAI as a Senior Design Engineer to shape and build exceptional product experiences in a fast-growing Voice AI company.

AssemblyAI Chicago $180k–$240k/yr Published 3 weeks ago
Flexible on stack
Yuno

Join Yuno as a Senior AI Engineer to architect and scale AI-powered systems that redefine payment orchestration.

Yuno Hyderabad Published 5 months ago
Flexible on stack
AssemblyAI

Join AssemblyAI as a Senior Design Engineer to shape and build exceptional product experiences in a fast-growing AI company.

AssemblyAI New York City $180k–$240k/yr Published 3 weeks ago
Flexible on stack
AssemblyAI

Join AssemblyAI as a Senior Design Engineer to shape and build exceptional product experiences in a fast-growing AI company.

AssemblyAI San Francisco $180k–$240k/yr Published 3 weeks ago
Flexible on stack
Sierra

Join Sierra as a Software Engineer on the Site Reliability team to enhance the reliability and scalability of AI-driven infrastructure.

Sierra San Francisco, CA Published 10 months ago
Flexible on stack
Faire

Join Faire as a Senior Applied AI/ML Scientist to drive retailer growth through innovative machine learning solutions.

Faire Kitchener-Waterloo, ON; Toronto, ON CA$180k–CA$247.5k/yr Published 1 month ago
Flexible on stack
AssemblyAI

Join AssemblyAI as a Senior Research Engineer to enhance large-scale distributed training and inference systems in Voice AI.

AssemblyAI Remote - New York $270k–$310k/yr Published 2 weeks ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a senior AI Field Engineer to lead technical partnerships and drive AI transformation in enterprise environments.

Fireworks AI San Mateo $200k–$260k/yr Published 3 months ago
Flexible on stack
AssemblyAI

Join AssemblyAI as a Senior Security Operations Engineer to shape security practices in a high-growth Voice AI company.

AssemblyAI Remote $180k–$220k/yr Published 3 months ago
Flexible on stack
AssemblyAI

Join AssemblyAI as a Senior Design Engineer to shape and build exceptional product experiences in a fast-growing AI company.

AssemblyAI Seattle $180k–$240k/yr Published 3 weeks ago
Flexible on stack
baseten

Join Baseten as a Technical Program Manager to build and optimize the core algorithms for high-performance AI inference.

baseten San Francisco Published 2 weeks ago