"ml model serving" Jobs

442 open tech roles matching “ml model serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 442 results

Sesame

Join Sesame as an ML Model Serving Engineer to enhance our serving layer for voice agents with cutting-edge techniques.

Sesame San Francisco Published 1 year ago
Flexible on stack
Chef Robotics

Join Chef Robotics as a Senior ML Engineer to develop and deploy foundation models for food robotics in production kitchens.

Chef Robotics San Francisco Published 3 months ago
Mercury

Build and operate real-time inference services for risk decisioning in a fast-growing fintech startup.

Mercury San Francisco, CA, New York, NY, Portland, OR, or Remote within Canada or United States $166.6k–$208.3k/yr Published 4 days ago
Flexible on stack
baseten

Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.

baseten San Francisco Published 11 months ago
Sprinter Health

Join Sprinter Health as an ML Engineer to build reliable production systems for machine learning models in a hybrid work environment.

Sprinter Health San Francisco, CA Published 1 month ago
Flexible on stack
Chef Robotics

Join Chef Robotics as a Senior ML Engineer to develop learning systems for autonomous robots in food preparation.

Chef Robotics San Francisco Published 3 months ago
Databricks

Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI workloads.

Databricks San Francisco, California $190k–$265k/yr Published 1 month ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.

baseten San Francisco Published 8 months ago
Flexible on stack
Sprinter Health

Join Sprinter Health as a Staff Machine Learning Engineer to build and lead the ML engineering function in a hybrid work environment.

Sprinter Health San Francisco, CA Published 1 month ago
Inferact

Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco Published 3 days ago
Flexible on stack
Databricks

Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.

Databricks San Francisco, California $190k–$265k/yr Published 1 month ago
Perplexity

Join Perplexity as a Senior ML Engineer to design and optimize recommendation systems for personalized user experiences.

Perplexity San Francisco Published 3 months ago
70% coding
Cursor

Lead a team of engineers to build infrastructure for training and evaluating ML models in a flat, innovative organization.

Cursor San Francisco Published 2 months ago
Heavy meetings
Perplexity AI

Join Perplexity AI as a senior ML Engineer to design and optimize recommendation systems for personalized user experiences.

Perplexity AI San Francisco Published 3 months ago
Benchling

Join Benchling as a senior Product Manager to drive the adoption of scientific AI models in biotech R&D.

Benchling San Francisco, CA Published 2 months ago