"inference infrastructure" Jobs

912 open tech roles matching “inference infrastructure”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 912 results

baseten

Join Baseten as a Product Manager to shape the future of AI infrastructure and enhance production inference capabilities.

baseten San Francisco Published 5 months ago
Ambient

Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced AI models.

Ambient Redwood City Published 2 months ago
Flexible on stack 70% coding
Anthropic

Join Anthropic as a Demand Planning expert to optimize AI infrastructure capacity and ensure timely delivery across multiple platforms.

Anthropic San Francisco, CA | New York City, NY $320k–$405k/yr Published 1 month ago
Flexible on stack
ElevenLabs

Join ElevenLabs as a Research Engineer to deploy and optimize AI models for real-time applications in a fully remote environment.

ElevenLabs United Kingdom Published 2 weeks ago
Flexible on stack
Together AI

Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.

Together AI San Francisco $200k–$290k/yr Published 2 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Product Manager to shape the future of AI inference and platform capabilities.

Fireworks AI San Mateo Published 3 weeks ago
AI-first team
Pika

Join Pika as a Senior/Staff ML Engineer to enhance AI-driven products through advanced inference acceleration and GPU optimization.

Pika Palo Alto HQ Published 2 months ago
Flexible on stack
baseten

Join Baseten as an AI Inference Engineer to architect and deploy high-scale production AI applications while collaborating with customers.

baseten San Francisco Published 1 month ago
Flexible on stack 70% coding
Inceptive

Architect and implement secure data infrastructure for AI model training and deployment in a collaborative, antedisciplinary team.

Inceptive Berlin, Germany $200k–$275k/yr Published 2 months ago
Flexible on stack 70% coding
Ambient

Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced machine learning models.

Ambient Redwood City, United States Published 2 days ago
Flexible on stack 70% coding
Anthropic

Join Anthropic as a Staff Software Engineer to build scalable ML infrastructure for AI safety systems.

Anthropic San Francisco, CA $320k–$485k/yr Published 3 days ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to architect and develop scalable infrastructure for ML training platforms.

baseten San Francisco Published 1 year ago
Flexible on stack
Perplexity

Join Perplexity as a technical program manager to drive the core inference platform and coordinate between model providers and engineering teams.

Perplexity San Francisco Published 1 week ago
Omnifold

Lead the infrastructure team at Omnifold, focusing on AI model training and deployment in a fast-paced startup environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
Reflection AI

Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.

Reflection AI San Francisco, CA Published 5 months ago
Flexible on stack
Cloudflare

Join Cloudflare as a Senior Systems Engineer to build core AI Gateway systems for high-volume inference traffic.

Cloudflare In-Office Published 1 month ago
Harvey

Lead the design and development of systems powering AI requests at Harvey, a fast-scaling company in the legal tech space.

Harvey San Francisco $231k–$340k/yr Published 6 days ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Lyft

Join Lyft as a Data Scientist to leverage causal inference for enhancing safety and customer care experiences.

Lyft Toronto, Canada CA$108k–CA$135k/yr Published 1 month ago
Flexible on stack
Preference Model

Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.

Preference Model San Francisco Published 2 days ago
Flexible on stack