"inference optimization" Jobs

214 open tech roles matching “inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 214 results

Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
baseten

Join Baseten as a Technical Program Manager to build and optimize the core algorithms for high-performance AI inference.

baseten San Francisco Published 2 weeks ago
baseten

Join Baseten as a Forward Deployed Engineer to solve complex AI challenges for leading companies.

baseten San Francisco Published 3 weeks ago
Flexible on stack
krea.ai

Join Krea to build innovative AI tools in a hands-on role focused on supercomputing and distributed systems.

krea.ai San Francisco Published 5 months ago
Flexible on stack
Inferact

Join Inferact as a Founding Product Designer to shape the visual identity and user experience of our AI inference engine.

Inferact San Francisco Published 1 month ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 1 year ago
krea.ai

Join Krea as an ML Researcher to train diffusion models for image and video generation in a creative AI-focused environment.

krea.ai San Francisco Published 1 week ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA $315k–$560k/yr Published 10 months ago
Goodfire

Join Goodfire as a Machine Learning Engineer to build interpretable AI systems with a world-class team.

Goodfire San Francisco, CA & New York, NY $200k–$400k/yr Published 9 months ago
OpenRouter

Join OpenRouter as a Forward Deployed Engineer to help customers implement and scale AI solutions effectively.

OpenRouter San Francisco Bay Area, California Published 3 months ago
Flexible on stack 70% coding
Together AI

Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.

Together AI San Francisco $220k–$280k/yr Published 3 months ago
Flexible on stack 60% coding
baseten

Join Baseten as a Software Engineer to lead GPU Networking efforts and optimize distributed systems for AI applications.

baseten San Francisco Published 6 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 2 years ago
Together AI

Design and deliver multi-petabyte storage systems for AI workloads at Together AI, optimizing performance and cost.

Together AI San Francisco $250k–$300k/yr Published 3 months ago
Flexible on stack
baseten

Join Baseten as a Cloud Platform Engineer to build scalable infrastructure for deploying machine learning models.

baseten San Francisco Published 11 months ago
Flexible on stack
Omnifold

Join Omnifold's Infrastructure Team to build robust systems for AI model training and deployment in a fast-paced environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
Harvey AI

Join Harvey AI as a Research Engineer to drive post-training experiments and enhance legal AI models.

Harvey AI San Francisco $231k–$340k/yr Published 2 months ago
Flexible on stack
baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
baseten

Join Baseten as a Marketing Analytics Manager to shape data-driven marketing strategies in a rapidly growing AI company.

baseten San Francisco Published 1 week ago
Flexible on stack