"ml model serving" Jobs

1272 open tech roles matching “ml model serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 1272 results

Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Palantir

Join Palantir's software engineering team to enable ML models in production for critical customers.

Palantir Washington, D.C. Published 3 months ago
Abridge

Join Abridge as a Machine Learning Infrastructure Engineer to optimize AI model inference infrastructure in a fast-paced healthcare startup.

Abridge SF Office Published 1 year ago
Flexible on stack
MaintainX

Join MaintainX as a Senior Applied Scientist to develop AI-powered inventory optimization models for enterprise maintenance teams.

MaintainX Canada Published 3 months ago
Flexible on stack
Reddit

Lead a high-performing machine learning team to innovate and optimize engagement prediction models for Reddit's Ads.

Reddit Remote - United States $230k–$322k/yr Published 3 months ago
Flexible on stack Heavy meetings
Cursor

Join Cursor as a Software Engineer on the ML Platform to build infrastructure that enhances machine learning models and supports product engineers.

Cursor San Francisco Published 1 week ago
Flexible on stack
Lyft

Join Lyft's ETA team as a Machine Learning Engineer to enhance ride experience through accurate ETA predictions.

Lyft San Francisco, CA $140.8k–$176k/yr Published 1 month ago
Flexible on stack
BloomReach

Join Bloomreach as a Senior AI/ML Engineer to build and maintain ML-powered features in a dynamic, hybrid work environment.

BloomReach Czechia $1300k–$1600k/yr Published 1 year ago
Flexible on stack
Lyft

Join Lyft as a Machine Learning Engineer to develop AI agents that enhance safety and customer care for riders and drivers.

Lyft Toronto, Canada CA$118.8k–CA$148.5k/yr Published 1 month ago
Flexible on stack 70% coding
BloomReach

Join Bloomreach as a Senior AI/ML Engineer to build and maintain ML-powered features that revolutionize marketing.

BloomReach Slovakia €40k–€49.5k/yr Published 1 year ago
Flexible on stack
Inferact

Join Inferact as a Site Reliability Engineer to enhance the reliability and performance of AI inference systems at scale.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack AI-first team
Lyra Health

Spearhead the design and deployment of next-generation AI capabilities for a leading mental health care platform.

Lyra Health United States Published 3 months ago
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Omnifold

Join Omnifold's Infrastructure Team to build robust systems for AI model training and deployment in a fast-paced environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
Patreon

Join Patreon as a Senior Machine Learning Engineer to architect and maintain high-throughput ML infrastructure for creator discovery.

Patreon New York Published 3 weeks ago
Flexible on stack
Inferact

Join Inferact as a cloud orchestration engineer to build reliable systems for AI model deployment at scale.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Airbnb

Own the technical strategy for ML serving and API interface in the Host Pricing org at Airbnb.

Airbnb Remote - USA Published 2 months ago
Flexible on stack 60% coding
Lyra Health

Define and drive the architectural vision for Lyra’s machine learning and generative AI technology landscape.

Lyra Health United States Published 1 month ago