"ml model serving" Jobs

434 open tech roles matching “ml model serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 434 results

baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
Lyft

Join Lyft as a Machine Learning Engineer to design and deploy cutting-edge ML systems in a collaborative environment.

Lyft San Francisco, CA $140.8k–$176k/yr Published 3 weeks ago
Flexible on stack 70% coding
MaintainX

Lead the technical direction for predictive maintenance and asset intelligence initiatives at MaintainX, leveraging deep ML expertise.

MaintainX San Francisco Published 1 month ago
Flexible on stack
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
MaintainX

Join MaintainX as a Senior Applied Scientist to develop AI-powered inventory optimization models for enterprise maintenance teams.

MaintainX San Francisco Published 1 week ago
Flexible on stack
Harvey AI

Lead a high-performing team to develop and manage the model infrastructure platform at Harvey AI.

Harvey AI San Francisco $272k–$355k/yr Published 1 month ago
Heavy meetings
Chai Discovery

Join Chai Discovery as a Software Engineer to optimize AI models for drug discovery in a fast-paced, innovative environment.

Chai Discovery San Francisco office Published 9 months ago
MaintainX

Own the LLMX roadmap as a Staff Product Manager at MaintainX, driving AI platform adoption and quality.

MaintainX Toronto Published 1 month ago
MaintainX

Own the LLMX roadmap as a Staff Product Manager at MaintainX, driving AI platform adoption and quality in a hybrid role.

MaintainX San Francisco Published 3 months ago
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a staff engineer to build distributed systems for AI inference at global scale.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Abridge

Join Abridge as a Machine Learning Infrastructure Engineer to optimize AI model inference infrastructure in a fast-paced healthcare startup.

Abridge SF Office Published 1 year ago
Flexible on stack
Cursor

Join Cursor as a Software Engineer on the ML Platform to build infrastructure that enhances machine learning models and supports product engineers.

Cursor San Francisco Published 2 weeks ago
Flexible on stack
Lyft

Join Lyft's ETA team as a Machine Learning Engineer to enhance ride experience through accurate ETA predictions.

Lyft San Francisco, CA $140.8k–$176k/yr Published 1 month ago
Flexible on stack
Inferact

Join Inferact as a Site Reliability Engineer to enhance the reliability and performance of AI inference systems at scale.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack AI-first team
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Omnifold

Join Omnifold's Infrastructure Team to build robust systems for AI model training and deployment in a fast-paced environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
Patreon

Join Patreon as a Senior Machine Learning Engineer to architect and maintain high-throughput ML infrastructure for creator discovery.

Patreon New York Published 4 weeks ago
Flexible on stack
Inferact

Join Inferact as a cloud orchestration engineer to build reliable systems for AI model deployment at scale.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Airbnb

Join Airbnb as a Staff Machine Learning Engineer to innovate customer service with cutting-edge AI technologies.

Airbnb San Francisco, CA $212k–$260k/yr Published 2 months ago