"inference infrastructure" Jobs

934 open tech roles matching “inference infrastructure”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 934 results

ChipAgents

Join ChipAgents as an ML Systems Engineer to optimize LLM inference systems for leading semiconductor companies.

ChipAgents San Jose $150k–$350k/yr Published 3 months ago
Flexible on stack
Perplexity

Join Perplexity as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions in a fast-paced environment.

Perplexity San Francisco Published 1 week ago
Hedra

Design and scale backend systems for Hedra’s developer platform, focusing on Python services and cloud infrastructure.

Hedra San Francisco Published 3 months ago
Applied Intuition

Join Applied Intuition as an ML Runtime Optimization Engineer to optimize ML models for embedded environments in a collaborative team.

Applied Intuition Sunnyvale Published 1 year ago
Flexible on stack
Fireworks AI

Own foundational capabilities for enterprise AI, designing data models and APIs while ensuring security and compliance for large-scale customers.

Fireworks AI New York Published 1 month ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a staff software engineer to enhance our model serving platform for AI products and infrastructure.

Perplexity AI San Francisco Published 2 weeks ago
Flexible on stack
Arena

Explore and analyze large datasets to uncover insights about AI model behavior in a mission-driven team.

Arena Bay Area Published 1 week ago
Flexible on stack
Skydio

Join Skydio as a Deep Learning Infrastructure Engineer to build and scale AI solutions for autonomous drones.

Skydio San Mateo, California, United States $170k–$236.5k/yr Published 9 months ago
Applied Intuition

Join Applied Intuition as a Software Engineer to optimize application-layer software for embedded systems in autonomous driving.

Applied Intuition Sunnyvale Published 1 year ago
Together AI

Design and deliver multi-petabyte storage systems for AI workloads at Together AI, optimizing performance and cost.

Together AI San Francisco $250k–$300k/yr Published 3 months ago
Flexible on stack
Applied Intuition

Lead the perception model team for autonomous vehicles at a rapidly growing AI infrastructure company.

Applied Intuition Sunnyvale Published 3 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and AI applications.

Inworld AI Serbia Published 5 months ago
Fireworks AI

Join Fireworks AI as a senior AI Field Engineer to build production systems for innovative AI-native companies.

Fireworks AI San Mateo Published 3 months ago
Flexible on stack 70% coding
baseten

Lead a team of cloud platform engineers to build scalable and reliable infrastructure for AI products at Baseten.

baseten San Francisco Published 3 months ago
Flexible on stack Heavy meetings
baseten

Join Baseten as an AI Engineer to build AI-driven product features and enhance internal workflows.

baseten San Francisco Published 3 weeks ago
Flexible on stack
Palantir

Join Palantir's software engineering team to enable ML models in production across various environments.

Palantir Palo Alto, CA Published 3 months ago
baseten

Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.

baseten San Francisco Published 11 months ago
Cartesia

Join Cartesia as a Software Engineer in India to build real-time multimodal intelligence with cutting-edge AI technologies.

Cartesia Bangalore Published 2 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Software Engineer to architect and build scalable cloud infrastructure for generative AI workloads.

Fireworks AI San Mateo Published 1 year ago
Flexible on stack