"inference optimization" Jobs

586 open tech roles matching “inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 586 results

Fireworks AI

Join Fireworks AI as a Backend Software Engineer to design and develop core backend systems for a high-performance generative AI platform.

Fireworks AI Singapore Published 2 months ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $206k–$333k/yr Published 8 months ago
DeepL

Lead the Production Inference team at DeepL, focusing on performance-critical model serving systems in a fast-paced AI environment.

DeepL London Published 1 month ago
Heavy meetings
Twelve Labs

Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.

Twelve Labs Seoul, South Korea Published 3 weeks ago
Flexible on stack
Fireworks AI

Lead paid marketing programs at Fireworks AI, managing multi-million dollar budgets to drive demand generation for cutting-edge AI infrastructure.

Fireworks AI San Mateo $160k–$190k/yr Published 3 months ago
Cantina

Join Cantina as a Machine Learning Intern to work on advanced video generation models in a hands-on research environment.

Cantina Singapore Published 6 days ago
Flexible on stack
Reddit

Join Reddit as a Staff Data Scientist to enhance advertising measurement and identity resolution in a fully remote role.

Reddit Remote - United States $217k–$303.9k/yr Published 1 month ago
Flexible on stack
Fundamental

Join Fundamental as a Senior Applied Research Engineer to tackle technical challenges in AI model development for enterprise decision-making.

Fundamental Barcelona Published 5 months ago
Flexible on stack
Dialpad

Join Dialpad as a Sr. AI Engineer to shape real-time speech systems for AI voice agents in a collaborative environment.

Dialpad Vancouver, Canada CA$184.5k–CA$213.8k/yr Published 1 week ago
Flexible on stack 70% coding
Reddit

Join Reddit as a Staff Data Scientist to lead innovative measurement and identity solutions in advertising.

Reddit Toronto, Canada Published 1 month ago
Instacart

Lead a team of engineers to build an AI-native platform for catalog enrichment in a fully remote environment.

Instacart Canada - Remote (ON, AB, BC, or NS Only) CA$196k–CA$207k/yr Published 4 months ago
Heavy meetings
Fireworks AI

Design and maintain large-scale backend infrastructure for a leading generative AI platform at Fireworks AI.

Fireworks AI New York Published 3 months ago
Flexible on stack
Faire

Join Faire as a Senior Applied AI/ML Scientist to drive brand growth through innovative machine learning solutions.

Faire San Francisco, CA $211k–$290.5k/yr Published 2 days ago
Flexible on stack
Cantina

Join Cantina as a Machine Learning Engineer to build advanced speech systems and contribute to innovative AI technology.

Cantina Remote (U.S. or Europe) $200k–$220k/yr Published 1 month ago
Flexible on stack 70% coding
baseten

Join Baseten as a Forward Deployed Engineer to architect and deploy high-scale AI applications while collaborating with customers.

baseten San Francisco Published 2 years ago
Flexible on stack 70% coding
Fundamental

Join Fundamental as an MLOps Engineer to tackle technical challenges in AI and transform enterprise decision-making.

Fundamental Europe Published 1 month ago
Flexible on stack
Vercel

Lead consumption forecasting at Vercel, architecting ML systems for financial planning and decision-making.

Vercel Hybrid - San Francisco, New York City $250k–$330k/yr Published 2 weeks ago
Flexible on stack
Graphcore

Join Graphcore as an AI Research Engineer to advance AI research with a focus on hardware-aware algorithms and impactful implementations.

Graphcore Bristol, UK Published 2 months ago
Flexible on stack
Graphcore

Join Graphcore as an AI Research Engineer to advance AI research and collaborate on next-gen AI hardware.

Graphcore Cambridge, UK Published 2 months ago
Flexible on stack