"ml hardware accelerators" Jobs

159 open tech roles matching “ml hardware accelerators”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 159 results

Fireworks AI

Join Fireworks AI as a Member of Technical Staff to design and build systems infrastructure for AI workloads at scale.

Fireworks AI San Mateo Published 4 days ago
Flexible on stack
Graphcore

Join Graphcore as an Infrastructure and MLOps Engineer to scale and manage AI compute infrastructure and tools.

Graphcore Bristol, UK Published 3 months ago
Flexible on stack
Graphcore

Join Graphcore as an Infrastructure and MLOps Engineer to scale and manage AI compute infrastructure and tools.

Graphcore Cambridge, UK Published 1 month ago
Flexible on stack
Graphcore

Join Graphcore as an Infrastructure and MLOps Engineer to scale and manage AI compute infrastructure and tools.

Graphcore London, UK Published 1 month ago
Flexible on stack
Inferact

Join Inferact as a staff engineer to build distributed systems for AI inference at global scale.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Inferact

Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Graphcore

Join Graphcore's Triton team to enhance AI frameworks and improve performance on cutting-edge hardware.

Graphcore Bristol, UK; Gdańsk, Pomeranian Voivodeship, Poland Published 2 weeks ago
Flexible on stack
ChipAgents

Join ChipAgents as a Research Scientist to advance AI-assisted electronic design automation in a high-impact, mission-driven team.

ChipAgents San Jose $150k–$350k/yr Published 3 months ago
Flexible on stack
Graphcore

Join Graphcore as a Senior Software Engineer to optimize AI hardware and software performance in a collaborative environment.

Graphcore Gdańsk, Pomeranian Voivodeship, Poland PLN 260.4k–PLN 352.2k/yr Published 4 months ago
Flexible on stack
Inferact

Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.

Inferact Remote Published 1 week ago
Flexible on stack
Inferact

Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Graphcore

Lead the architectural vision of the software stack for Graphcore's ML accelerator, inspiring and mentoring a dedicated team.

Graphcore Bristol, UK Published 7 months ago
Flexible on stack
ChipAgents

Join ChipAgents as a Full-Stack AI Engineer to create innovative AI technologies for chip design and verification.

ChipAgents San Jose $150k–$350k/yr Published 3 months ago
Flexible on stack
Skydio

Join Skydio as a Senior Autonomy Engineer to enhance deep learning infrastructure for autonomous drones.

Skydio San Mateo, California, United States $170k–$277.5k/yr Published 9 months ago
Flexible on stack
Worldcoin

Own face machine learning projects and improve biometric identification models in a collaborative AI team.

Worldcoin Munich Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Engineer to lead the technical direction of the Inference Runtime for AI systems serving millions of users.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $405k–$485k/yr Published 3 months ago
Flexible on stack
Atoms

Join Atoms as a Staff Machine Learning Infrastructure Engineer to design and build large-scale ML training infrastructure for autonomous transport models.

Atoms San Francisco, CA $224k–$280k/yr Published 2 months ago
Flexible on stack
Sesame

Join Sesame as a Senior Electrical Engineer to drive the development of innovative wearable electronics in a dynamic startup environment.

Sesame San Francisco Published 1 year ago
baseten

Lead the capacity management for Baseten's TPU fleet, ensuring optimal performance and reliability in AI workloads.

baseten San Francisco, United States Published 2 days ago
Flexible on stack