"ml model serving" Jobs

1271 open tech roles matching “ml model serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 1271 results

Inductive Bio

Join Inductive Bio as a Machine Learning expert to innovate on drug discovery models and collaborate with scientists.

Inductive Bio New York City, United States Published 2 days ago
Flexible on stack
Perplexity

Join Perplexity as a Senior ML Engineer to design and optimize recommendation systems for personalized user experiences.

Perplexity San Francisco Published 3 months ago
70% coding
Cursor

Lead a team of engineers to build infrastructure for training and evaluating ML models in a flat, innovative organization.

Cursor San Francisco Published 2 months ago
Heavy meetings
Lyra Health

Join Lyra Health as a Senior ML Engineer to build impactful ML and generative AI products in a collaborative environment.

Lyra Health United States Published 1 month ago
Perplexity AI

Join Perplexity AI as a senior ML Engineer to design and optimize recommendation systems for personalized user experiences.

Perplexity AI San Francisco Published 3 months ago
Lyft

Join Lyft as a Machine Learning Engineer to design and deploy cutting-edge ML systems in a collaborative hybrid environment.

Lyft New York, NY $140.8k–$176k/yr Published 3 weeks ago
Flexible on stack
SentiLink

Lead a team in model risk and governance for a rapidly growing identity verification company with a focus on financial institutions.

SentiLink United States $210k–$240k/yr Published 1 month ago
Flexible on stack Heavy meetings
Benchling

Join Benchling as a senior Product Manager to drive the adoption of scientific AI models in biotech R&D.

Benchling San Francisco, CA Published 2 months ago
Fal

Own the reliability and security of fal's generative media model APIs in a hybrid ML Engineering/SRE role.

Fal Remote - APAC Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a staff engineer to build distributed systems for AI inference at global scale.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
Lyft

Join Lyft as a Machine Learning Engineer to design and deploy cutting-edge ML systems in a collaborative environment.

Lyft San Francisco, CA $140.8k–$176k/yr Published 3 weeks ago
Flexible on stack 70% coding
Palantir

Join Palantir's software engineering team to enable ML models in production across various environments.

Palantir Palo Alto, CA Published 3 months ago
Inworld AI

Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic startup environment.

Inworld AI Germany Published 5 months ago
Flexible on stack
Wizard

Own the end-to-end lifecycle of production ML serving systems for a top-performing AI Shopping Agent.

Wizard Remote - USA Published 5 months ago
Flexible on stack
Airbnb

Fine-tune state-of-the-art LLMs and develop AI products to enhance the travel experience at Airbnb.

Airbnb Remote - USA $248k–$310k/yr Published 4 months ago
Flexible on stack 60% coding
Palantir

Join Palantir's software engineering team to enable ML models in production for critical customers.

Palantir New York, NY Published 3 months ago
Cloudflare

Join Cloudflare as a Senior Machine Learning Engineer to optimize and productionize ML models for a global serverless inference platform.

Cloudflare Hybrid Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
MaintainX

Lead the technical direction for predictive maintenance and asset intelligence initiatives at MaintainX, leveraging deep ML expertise.

MaintainX San Francisco Published 1 month ago
Flexible on stack