"llm inference" Jobs

451 open tech roles matching “llm inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 451 results

Coreweave

Join CoreWeave as a Staff Software Engineer to lead the development of a Kubernetes-native inference platform for AI workloads.

Coreweave Sunnyvale, CA / Bellevue, WA $188k–$275k/yr Published 4 months ago
Flexible on stack
Inflection AI

Lead model training and post-training strategies for emotionally intelligent AI at Inflection AI.

Inflection AI Palo Alto, California, United States $400k–$550k/yr Published 2 months ago
Inflection AI

Lead the development of Inflection's realtime Voice AI stack, shaping emotionally intelligent AI for enterprise voice interactions.

Inflection AI Palo Alto, California, United States $400k–$550k/yr Published 2 months ago
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to build and optimize large-scale AI training and inference clusters.

Perplexity AI London Published 5 months ago
Flexible on stack
Omnifold

Lead a research team at Omnifold to develop advanced forecasting and optimization models in a startup environment.

Omnifold San Francisco HQ Published 2 days ago
Inworld AI

Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and impact AI applications globally.

Inworld AI Germany Published 5 months ago
Databricks
Databricks San Francisco, California $280k–$350k/yr Published 4 months ago
SpaceX

Join SpaceX as a Software Engineer to develop high-performance AI inference systems for mission-critical applications.

SpaceX Palo Alto, CA $135k–$175k/yr Published 3 weeks ago
Flexible on stack
Preference Model

Join Preference Model as a Senior Machine Learning Engineer to design RL environments for advancing ML capabilities.

Preference Model San Francisco, United States Published 1 day ago
Flexible on stack
Coreweave

Join CoreWeave as an Applied AI Engineer to enhance the performance of our inference platform through benchmarking and optimization.

Coreweave Bellevue, WA/ San Francisco, CA/ Sunnyvale, CA $188k–$275k/yr Published 6 months ago
Flexible on stack
Stuut

Join Stuut as a Member of the Technical Staff to design and deploy AI-powered systems for financial operations.

Stuut San Francisco Published 1 month ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Member of Technical Staff to advance generative AI through foundational research and collaboration with top experts.

Fireworks AI San Mateo Published 1 month ago
Flexible on stack
Fundamental

Join Fundamental as a Model Serving Engineer to optimize and scale the NEXUS model for enterprise decision-making.

Fundamental Europe Published 5 months ago
Flexible on stack
Typeface

Lead the design and strategy for large-scale ML systems and generative AI at Typeface, influencing company-level initiatives.

Typeface Palo Alto, CA $230k–$260k/yr Published 4 months ago
Flexible on stack
MaintainX

Lead the technical direction for predictive maintenance and asset intelligence initiatives at MaintainX, leveraging deep ML expertise.

MaintainX San Francisco Published 1 month ago
Flexible on stack
Preference Model

Join Preference Model as a senior ML Engineer to design RL environments for advancing machine learning capabilities.

Preference Model San Francisco Published 2 weeks ago
Flexible on stack
Fireworks AI

Design and optimize infrastructure for large-scale AI model training at a leading generative AI company.

Fireworks AI San Mateo Published 1 month ago
Flexible on stack
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Member of Technical Staff to advance generative AI and multimodal systems through foundational research.

Fireworks AI San Mateo Published 4 days ago
Flexible on stack