"inference systems" Jobs

1011 open tech roles matching “inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 1011 results

Inworld AI

Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic startup environment.

Inworld AI Germany Published 5 months ago
Flexible on stack
Applied Intuition

Lead the perception model team for autonomous vehicles at a rapidly growing AI infrastructure company.

Applied Intuition Sunnyvale Published 3 months ago
Flexible on stack
Together AI

Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.

Together AI San Francisco $200k–$290k/yr Published 2 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an AI Performance Engineer to optimize large-scale machine learning workloads in a collaborative environment.

Applied Intuition Sunnyvale Published 1 month ago
Flexible on stack
Dialpad

Join Dialpad as a Sr. AI Engineer to shape real-time speech systems for AI voice agents in a collaborative environment.

Dialpad Vancouver, Canada CA$184.5k–CA$213.8k/yr Published 1 week ago
Flexible on stack 70% coding
Applied Intuition

Join Applied Intuition as an Embedded AI Engineer to develop on-device intelligence for Android Automotive platforms.

Applied Intuition Sunnyvale Published 5 months ago
Flexible on stack
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
baseten

Lead the finance systems strategy and implementation at a rapidly growing AI company.

baseten San Francisco Published 1 week ago
Insitro

Lead and grow a team of machine learning researchers to develop methods for extracting insights from in-vitro microscopy datasets.

Insitro South San Francisco, CA $247k–$262k/yr Published 3 months ago
Heavy meetings
Inferact

Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.

Inferact Remote Published 7 months ago
Flexible on stack
Inferact

Join Inferact as a cloud orchestration engineer to build reliable systems for AI model deployment at scale.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Reflection AI

Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.

Reflection AI San Francisco, CA Published 5 months ago
Flexible on stack
Dialpad

Join Dialpad as a Senior Software Engineer to build and improve the AI/ML inference platform for enterprise-scale applications.

Dialpad Buenos Aires, Argentina Published 5 days ago
Flexible on stack 70% coding
Inworld AI

Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and AI applications.

Inworld AI Mountain View, California, USA $270k–$500k/yr Published 3 years ago
baseten

Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.

baseten San Francisco Published 8 months ago
Flexible on stack
baseten

Join Baseten as an AI Engineer to design and automate workflows for AI-powered capacity management in a fast-growing company.

baseten San Francisco, United States Published 2 days ago
Flexible on stack
Inferact

Join Inferact as an IT Support & Operations Engineer to enhance internal technology and security for a growing AI startup.

Inferact San Francisco, CA, United States $125k–$170k/yr Published 2 days ago
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack