"ml training infra" Jobs

842 open tech roles matching “ml training infra”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 842 results

Sarvam AI

Own the data infrastructure for foundational models at Sarvam AI, focusing on large-scale data pipelines and quality systems.

Sarvam AI Bengaluru Published 3 months ago
Flexible on stack
Mirelo AI

Join Mirelo AI as a Training Infrastructure Engineer to optimize and design scalable systems for training generative AI models.

Mirelo AI Berlin Published 9 months ago
Flexible on stack
Preference Model

Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.

Preference Model San Francisco, United States Published 1 day ago
Flexible on stack
Periodic Labs

Join Periodic Labs as an ML Systems Engineer to build and optimize large-scale training and reinforcement learning infrastructure.

Periodic Labs Menlo Park, CA $250k–$350k/yr Published 4 months ago
Flexible on stack
Fundamental

Join Fundamental as an MLOps Engineer to tackle technical challenges in AI and transform enterprise decision-making.

Fundamental Europe Published 1 month ago
Flexible on stack
Preference Model

Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.

Preference Model San Francisco Published 3 days ago
Flexible on stack
krea.ai

Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.

krea.ai San Francisco Published 1 week ago
Flexible on stack
Genesis Molecular AI

Join Genesis Molecular AI as a Machine Learning Infrastructure Engineer to build innovative data systems for drug discovery.

Genesis Molecular AI San Mateo, CA Published 9 months ago
Flexible on stack
Triomics

Join Triomics as an MLOps & Data Engineer to build infrastructure for ML workflows in oncology, impacting patient outcomes.

Triomics India Office Published 2 months ago
Flexible on stack
Omnifold

Join Omnifold's Infrastructure Team to build robust systems for AI model training and deployment in a fast-paced environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
Fundamental

Lead a team of MLOps engineers at an AI company transforming enterprise decision-making.

Fundamental Europe Published 2 months ago
Flexible on stack
Sprinter Health

Join Sprinter Health as an ML Engineer to build reliable production systems for machine learning models in a hybrid work environment.

Sprinter Health San Francisco, CA Published 1 month ago
Flexible on stack
Abridge

Join Abridge as a Machine Learning Infrastructure Engineer to optimize AI model inference infrastructure in a fast-paced healthcare startup.

Abridge SF Office Published 1 year ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to architect and develop scalable infrastructure for ML training platforms.

baseten San Francisco Published 1 year ago
Flexible on stack
Hippocratic AI

Own the serving infrastructure for healthcare AI, optimizing LLM inference systems to enhance patient experiences.

Hippocratic AI Menlo Park, CA Published 3 weeks ago
Flexible on stack
Sesame

Join Sesame as an ML Model Serving Engineer to enhance our serving layer for voice agents with cutting-edge techniques.

Sesame San Francisco Published 1 year ago
Flexible on stack
Applied Intuition

Join Applied Intuition as a Senior Software Engineer to design and implement ML infrastructure for deep learning model training.

Applied Intuition Sunnyvale $215k–$285k/yr Published 3 years ago
Flexible on stack
Iambic Therapeutics

Join Iambic Therapeutics as a Machine Learning Scientist to innovate AI-based drug discovery with multimodal models.

Iambic Therapeutics UK Office Published 2 weeks ago
Flexible on stack
Fundamental

Join Fundamental as an ML Researcher to tackle groundbreaking challenges in AI model development for enterprise decision-making.

Fundamental Barcelona Published 10 months ago
Flexible on stack
Mercury

Build and operate real-time inference services for risk decisioning in a fast-growing fintech startup.

Mercury San Francisco, CA, New York, NY, Portland, OR, or Remote within Canada or United States $166.6k–$208.3k/yr Published 3 days ago
Flexible on stack