"real time ml inference" Jobs

289 open tech roles matching “real time ml inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, PyTorch, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 289 results

Sprinter Health

Join Sprinter Health as an Applied Scientist, AI to develop machine learning models that enhance healthcare access and operational efficiency.

Sprinter Health San Francisco, CA Published 1 month ago
Flexible on stack
Ambiq

Join Ambiq as a Principal Edge AI Firmware Engineer to develop and optimize embedded software for real-time, battery-powered AI applications.

Ambiq Austin, Texas, United States Published 3 months ago
Flexible on stack
Fiddler AI

Lead applied research and development for AI models at Fiddler AI, ensuring safety and compliance in production applications.

Fiddler AI Palo Alto $220k–$260k/yr Published 1 month ago
Flexible on stack 60% coding
Harvey AI

Lead a high-performing team to develop and manage the model infrastructure platform at Harvey AI.

Harvey AI San Francisco $272k–$355k/yr Published 1 month ago
Heavy meetings
baseten

Join the Base Labs Fellowship to conduct cutting-edge AI research with mentorship and funding in San Francisco.

baseten San Francisco $15k–$15k/mo Published 2 months ago
Dialpad

Lead product direction for AI models and infrastructure at Dialpad, focusing on real-time customer experience solutions.

Dialpad Bay Area, US $210.5k–$266k/yr Published 2 months ago
Truecaller

Join Truecaller as a Senior Data Engineer to build and own the data foundation for recommendation and advertising ML systems.

Truecaller Stockholm, Sweden Published 2 months ago
Flexible on stack
Omada Health

Lead high-impact research and applied algorithm development in healthcare as a Principal Applied ML Scientist at Omada Health.

Omada Health Remote, USA $270.5k–$338.1k/yr Published 2 months ago
Flexible on stack
Atoms

Join Atoms as a Senior Machine Learning Engineer to bridge AI research and real-world physical actuation in autonomous transport.

Atoms San Francisco, CA $208k–$263.5k/yr Published 2 months ago
Flexible on stack
DeepL

Lead scientific innovation in speech and translation models for real-time voice products at DeepL.

DeepL London Published 1 month ago
Flexible on stack 70% coding
Dialpad

Join Dialpad as a Sr. AI Engineer to lead the development of next-generation AI voice agents in a collaborative environment.

Dialpad Bay Area, US $224.5k–$256k/yr Published 1 week ago
Flexible on stack
AKASA

Join AKASA as a Sr. Machine Learning Engineer to develop state-of-the-art ML solutions that empower healthcare professionals.

AKASA Remote - United States $175k–$230k/yr Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff+ Software Engineer to build and scale Claude Managed Agents in a rapidly growing team.

Anthropic San Francisco, CA | New York City, NY $405k–$485k/yr Published 3 weeks ago
DeepL

Lead a senior engineering team to develop DeepL's real-time AI voice products in a hands-on leadership role.

DeepL London Published 3 weeks ago
Heavy meetings
Coreweave

Join CoreWeave as a Senior Engineer to build performance insights and observability systems for AI infrastructure.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 1 month ago
Flexible on stack
Elastic

Join Elastic as an AI QA & Evaluation Engineer to validate and test AI infrastructure and solutions in a fully remote environment.

Elastic Bangalore, India Published 2 weeks ago
Flexible on stack
Decagon

Lead the development of models and algorithms for Decagon's real-time voice agents in a collaborative, onsite environment.

Decagon San Francisco $200k–$400k/yr Published 4 months ago
Flexible on stack 70% coding
baseten

Join Baseten as a foundational member of the GTM team, driving revenue strategy and operations in the AI infrastructure space.

baseten San Francisco Published 2 months ago
Coreweave

Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 2 months ago
70% coding