"cuda q" Jobs

118 open tech roles matching “cuda q”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, CUDA, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 118 results

Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco Published 2 days ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $92k–$135k/yr Published 10 months ago
Coreweave

Join CoreWeave as a Staff Software Engineer to lead the development of a Kubernetes-native inference platform for AI workloads.

Coreweave Sunnyvale, CA / Bellevue, WA $188k–$275k/yr Published 4 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve top-ranked realtime voice models.

Inworld AI UK £140k–£200k/yr Published 5 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Member of Technical Staff to design and build systems infrastructure for AI workloads at scale.

Fireworks AI San Mateo Published 4 days ago
Flexible on stack
baseten

Join Baseten as a Software Engineer focused on ML performance to optimize large language models in a fast-paced startup environment.

baseten San Francisco Published 2 years ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an ML Runtime Optimization Engineer to optimize ML models for embedded environments in a collaborative team.

Applied Intuition Sunnyvale Published 1 year ago
Flexible on stack
Together AI

Join Together AI as a Staff Software Engineer to build systems that automate infrastructure management for AI clusters.

Together AI India Published 3 weeks ago
Flexible on stack
Inferact

Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.

Inferact Remote Published 7 months ago
Flexible on stack
World Labs

Join World Labs as a Performance Engineer to optimize AI models for speed and efficiency in a cutting-edge research environment.

World Labs San Francisco $200k–$300k/yr Published 4 months ago
Flexible on stack 70% coding
Fireworks AI

Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.

Fireworks AI San Mateo Published 1 year ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve state-of-the-art voice models in a fully remote role.

Inworld AI Switzerland Published 5 months ago
Flexible on stack
Together AI

Join Together AI as a Staff Software Engineer to build systems that automate infrastructure management for AI clusters.

Together AI London & Amsterdam Published 3 weeks ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Hawthorne, CA $165k–$230k/yr Published 2 months ago
Flexible on stack
SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Starbase, TX Published 2 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an AI Performance Engineer to optimize large-scale machine learning workloads in a collaborative environment.

Applied Intuition Sunnyvale Published 1 month ago
Flexible on stack