"llm inference systems" Jobs

392 open tech roles matching “llm inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 392 results

Inworld AI

Join Inworld AI as a Staff/Principal Software Engineer to develop cutting-edge backend systems for real-time voice models.

Inworld AI Mountain View, California, USA $280k–$350k/yr Published 1 year ago
Flexible on stack 70% coding
Preference Model

Join Preference Model as a senior ML Engineer to design RL environments for advancing machine learning capabilities.

Preference Model San Francisco Published 2 weeks ago
Flexible on stack
HappyRobot

Join HappyRobot as a Machine Learning Engineer to build AI models for human-like conversations and shape the future of AI infrastructure.

HappyRobot Spain Published 4 months ago
Flexible on stack
Preference Model

Join Preference Model as a new graduate Machine Learning Engineer to design and build reinforcement learning environments.

Preference Model San Francisco Published 4 months ago
Flexible on stack
SpaceX

Join SpaceX as a Software Engineer to develop high-performance AI inference systems for mission-critical applications.

SpaceX Palo Alto, CA $135k–$175k/yr Published 3 weeks ago
Flexible on stack
Inflection AI

Lead model training and post-training strategies for emotionally intelligent AI at Inflection AI.

Inflection AI Palo Alto, California, United States $400k–$550k/yr Published 2 months ago
Databricks
Databricks San Francisco, California $54–$60/hr Published 2 years ago
Reflection AI

Join Reflection AI to build systems that transform pre-trained models into aligned agents in a fast-paced startup environment.

Reflection AI San Francisco, CA Published 1 month ago
Together AI

Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.

Together AI San Francisco $200k–$290k/yr Published 2 months ago
Flexible on stack
DeepL

Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.

DeepL London Published 1 month ago
Flexible on stack
NODA AI

Join NODA AI as an AI/ML Engineer to design intelligent agents for autonomous systems in a hybrid work environment.

NODA AI Austin Published 10 months ago
Flexible on stack
Wonderful

Design and maintain backend services and AI systems while collaborating with research and product teams in a fast-paced startup environment.

Wonderful Tel Aviv, HQ Published 5 months ago
Harvey

Lead the design and development of systems powering AI requests at Harvey, a fast-scaling company in the legal tech space.

Harvey San Francisco $231k–$340k/yr Published 6 days ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $165k–$242k/yr Published 11 months ago
Hilbert

Join Hilbert as an AI Engineer to build production-grade AI systems that drive enterprise outcomes in a fast-paced startup environment.

Hilbert San Francisco Published 6 months ago
Flexible on stack 70% coding
Profound

Lead the Answer Engine Insights team at Profound, focusing on analytics and AI search in a fast-paced environment.

Profound Buenos Aires, Argentina $125k–$165k/yr Published 3 months ago
Flexible on stack
Instacart

Lead the design and development of core ML models for Instacart’s ads ecosystem in a fully remote role.

Instacart United States - Remote $201k–$253.5k/yr Published 3 months ago
Flexible on stack
Perplexity

Join Perplexity as a Machine Learning Engineer to enhance search quality through innovative ranking solutions.

Perplexity Belgrade Published 1 month ago
Fundamental

Join Fundamental as a Model Serving Engineer to optimize and scale the NEXUS model for enterprise decision-making.

Fundamental Europe Published 5 months ago
Flexible on stack