"inference serving" Jobs

350 open tech roles matching “inference serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 350 results

baseten

Join Baseten as a senior software engineer to develop cutting-edge AI training products and enhance user workflows.

baseten San Francisco Published 7 months ago
Flexible on stack
Together AI

Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.

Together AI San Francisco $220k–$280k/yr Published 3 months ago
Flexible on stack 60% coding
Inferact

Join Inferact as a cloud orchestration engineer to build reliable systems for AI model deployment at scale.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Peregrine

Lead the development of AI-powered features for an end-to-end intelligence platform in public safety.

Peregrine San Francisco, CA $225k–$320k/yr Published 7 months ago
Twelve Labs

Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 5 months ago
Flexible on stack
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
Patreon

Join Patreon as a Senior Machine Learning Engineer to architect and maintain high-throughput ML infrastructure for creator discovery.

Patreon New York Published 4 weeks ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA $315k–$560k/yr Published 10 months ago
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
Harvey AI

Lead the design and development of systems powering AI requests at Harvey, collaborating with multiple teams to ensure reliability and efficiency.

Harvey AI San Francisco $236k–$290k/yr Published 1 month ago
Flexible on stack
Hilbert

Join Hilbert as an AI Engineer to build production-grade AI systems that drive enterprise outcomes in a fast-paced startup environment.

Hilbert San Francisco Published 6 months ago
Flexible on stack 70% coding
Databricks

Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI workloads.

Databricks San Francisco, California $190k–$265k/yr Published 2 months ago
Flexible on stack
Inductive Bio

Join Inductive Bio as a software engineer to build AI tools that accelerate drug discovery.

Inductive Bio New York City, San Francisco, or Boston Published 5 months ago
Reflection AI

Lead large-scale AI infrastructure engagements with governments and enterprises, shaping complex partnerships and driving strategic outcomes.

Reflection AI San Francisco, CA Published 2 weeks ago
Twitch

Join Twitch's Monetization team as a Data Scientist to drive product decisions through rigorous analysis and causal inference.

Twitch San Francisco, CA $136k–$212.8k/yr Published 2 months ago
Flexible on stack
Twitch

Join Twitch's Monetization team as a Data Scientist to drive product decisions through causal analysis and experimentation.

Twitch Seattle, WA $136k–$212.8k/yr Published 2 months ago
Flexible on stack
Twitch

Join Twitch's Monetization team as a Data Scientist to apply causal inference methods and optimize revenue for creators.

Twitch New York City $136k–$212.8k/yr Published 2 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack