"real time ml inference" Jobs

122 open tech roles matching “real time ml inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, PyTorch, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 122 results

Dexterity

Develop and optimize computer vision models and ML pipelines for robotics at Dexterity.

Dexterity Redwood City Published 1 year ago
Flexible on stack
Cartesia

Join Cartesia as a Researcher in London to advance AI through innovative neural network architecture design.

Cartesia London Published 11 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 11 months ago
baseten

Join Baseten as a Sr. Analyst in Revenue Strategy & Operations to shape GTM strategies for AI infrastructure.

baseten San Francisco Published 1 week ago
Tigera

Own the ML and applied AI aspects of a new product tackling security challenges in enterprise infrastructure.

Tigera Vancouver CA$160k–CA$180k/yr Published 1 month ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Senior Reliability Engineer to ensure dependable AI systems and cloud infrastructure.

Fireworks AI San Mateo Published 4 weeks ago
Flexible on stack
Sarvam AI

Own the architecture of Sarvam's vision models serving harness, ensuring high-quality document intelligence at national scale.

Sarvam AI Bengaluru Published 3 weeks ago
Flexible on stack 70% coding
Harvey AI

Lead a high-performing team to develop and manage the model infrastructure platform at Harvey AI.

Harvey AI San Francisco $272k–$355k/yr Published 1 month ago
Heavy meetings
Truecaller

Join Truecaller as a Senior Data Engineer to build and own the data foundation for recommendation and advertising ML systems.

Truecaller Stockholm, Sweden Published 2 months ago
Flexible on stack
Atoms

Join Atoms as a Senior Machine Learning Engineer to bridge AI research and real-world physical actuation in autonomous transport.

Atoms San Francisco, CA $208k–$263.5k/yr Published 2 months ago
Flexible on stack
Dialpad

Join Dialpad as a Sr. AI Engineer to lead the development of next-generation AI voice agents in a collaborative environment.

Dialpad Bay Area, US $224.5k–$256k/yr Published 1 week ago
Flexible on stack
Lyft

Lead technical initiatives in data science to enhance Lyft's enterprise offerings and drive growth through algorithm development.

Lyft San Francisco, CA $136.2k–$170.2k/yr Published 2 months ago
Flexible on stack
AKASA

Join AKASA as a Sr. Machine Learning Engineer to develop state-of-the-art ML solutions that empower healthcare professionals.

AKASA Remote - United States $175k–$230k/yr Published 3 months ago
Flexible on stack
Lyft

Lead technical initiatives in data science to enhance Lyft's enterprise offerings and drive growth.

Lyft New York, NY $136.2k–$170.2k/yr Published 2 months ago
Flexible on stack
DeepL

Lead a senior engineering team to develop DeepL's real-time AI voice products in a hands-on leadership role.

DeepL London Published 2 weeks ago
Heavy meetings
Coreweave

Join CoreWeave as a Senior Engineer to build performance insights and observability systems for AI infrastructure.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 1 month ago
Flexible on stack
Decagon

Lead the development of models and algorithms for Decagon's real-time voice agents in a collaborative, onsite environment.

Decagon San Francisco $200k–$400k/yr Published 3 months ago
Flexible on stack 70% coding
baseten

Join Baseten as a foundational member of the GTM team, driving revenue strategy and operations in the AI infrastructure space.

baseten San Francisco Published 2 months ago
Coreweave

Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 2 months ago
70% coding
Imprint

Deliver analytical projects that influence product decisions and marketing campaigns in a fast-paced startup environment.

Imprint New York City Published 1 month ago
Flexible on stack