"llm inference" Jobs

451 open tech roles matching “llm inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 451 results

Fireworks AI

Join Fireworks AI as a Member of Technical Staff to build innovative AI solutions on a large inference platform.

Fireworks AI London Published 5 days ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $92k–$135k/yr Published 10 months ago
Together AI

Join Together AI as a Forward Deployed Engineer to optimize inference systems for strategic customers in a hands-on role.

Together AI Singapore Published 1 month ago
Flexible on stack 70% coding
Perplexity

Lead the financial strategy for AI products at Perplexity, optimizing model spend and driving pricing decisions.

Perplexity San Francisco Published 1 week ago
Inworld AI

Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models within a fully remote team in Switzerland.

Inworld AI Switzerland Published 7 months ago
Genesis Molecular AI

Join a world-class team to lead transformative research in generative AI for drug discovery at Genesis Molecular AI.

Genesis Molecular AI San Mateo, CA Published 1 year ago
Flexible on stack
Airbnb

Fine-tune state-of-the-art LLMs and develop AI products to enhance the travel experience at Airbnb.

Airbnb Remote - USA $248k–$310k/yr Published 4 months ago
Flexible on stack 60% coding
Inworld AI

Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and impact AI applications globally.

Inworld AI UK £140k–£200k/yr Published 5 months ago
Sarvam AI

Own the model lifecycle for defence and strategic sector deployments as an MLOps Engineer at Sarvam AI.

Sarvam AI Delhi Published 4 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Software Engineer to design and optimize backend services for cloud inference at scale.

Anthropic San Francisco, CA $320k–$485k/yr Published 3 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Databricks
Databricks San Francisco, California $54–$60/hr Published 2 years ago
Perplexity AI

Lead the economics of AI products at Perplexity AI, optimizing model spend and driving pricing and margin decisions.

Perplexity AI San Francisco Published 1 week ago
Reflection AI

Lead the post-training and evaluation capabilities for large language models in a dynamic AI research lab.

Reflection AI New York, NY Published 10 months ago
Instacart

Lead the design and development of core ML models for Instacart’s ads ecosystem in a fully remote role.

Instacart United States - Remote $201k–$253.5k/yr Published 3 months ago
Flexible on stack
Freenome

Join Freenome as a Senior Machine Learning Engineer to develop and optimize deep learning pipelines for cancer detection.

Freenome Remote $173.8k–$246.8k/yr Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic's Inference team to design and maintain distributed systems that serve AI models to millions of users worldwide.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $320k–$485k/yr Published 3 months ago
Flexible on stack