"ml serving" Jobs

536 open tech roles matching “ml serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 536 results

baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
Insitro

Join insitro as a Senior ML Scientist to develop machine learning methods for biological data analysis in a vibrant biotech startup.

Insitro South San Francisco, CA $183k–$238k/yr Published 4 months ago
Flexible on stack
Together AI

Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.

Together AI San Francisco $220k–$280k/yr Published 3 months ago
Flexible on stack 60% coding
Mariana Minerals

Own a critical product domain at Mariana Minerals, driving software solutions for the minerals supply chain with a focus on operational performance.

Mariana Minerals San Francisco HQ Published 2 months ago
Benchling

Join Benchling as a senior Product Manager to drive the adoption of scientific AI models in biotech R&D.

Benchling San Francisco, CA Published 2 months ago
Sesame

Join Sesame as a Backend Software Engineer to tackle complex challenges in building reliable, scalable systems for innovative voice agents.

Sesame San Francisco Published 2 weeks ago
Flexible on stack 70% coding
Databricks

Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.

Databricks San Francisco, California $190k–$265k/yr Published 1 month ago
Kodiak Robotics

Join Kodiak Robotics as a Staff Machine Learning Engineer to design and deploy ML systems for autonomous trucking.

Kodiak Robotics San Francisco Bay Area $200k–$265k/yr Published 4 months ago
Flexible on stack
Inferact

Join Inferact as a cloud orchestration engineer to build reliable systems for AI model deployment at scale.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Lyft

Join Lyft as a Staff Applied Scientist to develop ML and optimization models that enhance pricing and ETA decisions.

Lyft San Francisco, CA $193.6k–$242k/yr Published 1 month ago
Flexible on stack
Omnifold

Join Omnifold's Infrastructure Team to build robust systems for AI model training and deployment in a fast-paced environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
Cursor

Join Cursor as a Software Engineer on the ML Platform to build infrastructure that enhances machine learning models and supports product engineers.

Cursor San Francisco Published 2 weeks ago
Flexible on stack
Sesame

Join Sesame as a Data Engineer to build and maintain data pipelines for AI models in a team of experts from leading tech companies.

Sesame San Francisco Published 2 months ago
Flexible on stack
Lumafield

Join Lumafield as a Forward Deployed Engineer to work directly with customers, applying expertise in manufacturing technology on-site.

Lumafield Boston, MA Published 2 months ago
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco Published 4 days ago
Flexible on stack
Airbnb

Join Airbnb as a Staff Machine Learning Engineer to innovate customer service with cutting-edge AI technologies.

Airbnb San Francisco, CA $212k–$260k/yr Published 2 months ago