"real time ml inference" Jobs

115 open tech roles matching “real time ml inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 115 results

baseten

Join Baseten as a Software Engineer to build an AI developer platform that enhances productivity for engineers.

baseten San Francisco Published 1 month ago
Flexible on stack
Lyft

Join Lyft as a Staff Applied Scientist to develop ML and optimization models that enhance pricing and ETA decisions.

Lyft San Francisco, CA $193.6k–$242k/yr Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic as a Staff+ Site Reliability Engineer to ensure safe AI model launches and automate deployment processes.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $320k–$485k/yr Published 1 week ago
Flexible on stack
Databricks
Databricks New York City, New York; San Francisco, California $190k–$270k/yr Published 4 months ago
Databricks

Develop and run the research stack that powers Databricks AI Research, enabling rapid large-scale experiments.

Databricks New York City, New York; San Francisco, California $199k–$270k/yr Published 4 months ago
Flexible on stack
Hilbert

Join Hilbert as an AI Engineer to build production-grade AI systems that drive enterprise outcomes in a fast-paced startup environment.

Hilbert San Francisco Published 6 months ago
Flexible on stack 70% coding
Rox

Join Rox as a Founding Applied Research Engineer to shape the future of applied AI with a focus on real-world production challenges.

Rox San Francisco Published 3 months ago
Twelve Labs

Lead and build a new team focused on developing Jockey Core, a reasoning LLM for video understanding at Twelve Labs.

Twelve Labs Seoul, South Korea Published 1 month ago
Flexible on stack
Peregrine

Lead the development of AI-powered features for an end-to-end intelligence platform in public safety.

Peregrine San Francisco, CA $225k–$320k/yr Published 7 months ago
Sprinter Health

Join Sprinter Health as a Staff Machine Learning Engineer to build and lead the ML engineering function in a hybrid work environment.

Sprinter Health San Francisco, CA Published 1 month ago
Decagon

Join Decagon as a Research Engineer to build next-generation AI voice agents in a collaborative, onsite environment.

Decagon San Francisco $200k–$400k/yr Published 1 week ago
Flexible on stack 70% coding
krea.ai

Join Krea as a Backend Software Engineer to build AI creative tools in a collaborative, innovative environment.

krea.ai San Francisco Published 1 month ago
Flexible on stack
baseten

Join Baseten as a Senior Frontend Engineer to craft user-friendly interfaces for AI systems in a collaborative environment.

baseten San Francisco Published 2 months ago
Harvey AI

Join Harvey AI as a Research Engineer to drive post-training experiments and enhance legal AI models.

Harvey AI San Francisco $231k–$340k/yr Published 2 months ago
Flexible on stack
Descript

Join Descript as a Senior Software Engineer to own and enhance our infrastructure platform, impacting AI model training and deployment.

Descript San Francisco, CA or Remote, US $220k–$292k/yr Published 1 week ago
Flexible on stack
Xaira Therapeutics

Join Xaira Therapeutics as an AI in Residence to apply advanced AI in drug discovery and development.

Xaira Therapeutics South San Francisco, California, United States $10k–$15k/mo Published 5 months ago
Twitch

Join Twitch's Monetization team as a Data Scientist to drive product decisions through rigorous analysis and causal inference.

Twitch San Francisco, CA $136k–$212.8k/yr Published 2 months ago
Flexible on stack
Twitch

Join Twitch's Monetization team as a Data Scientist to apply causal inference methods and optimize revenue for creators.

Twitch New York City $136k–$212.8k/yr Published 2 months ago
Flexible on stack