"llm based systems" Jobs

346 open tech roles matching “llm based systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 346 results

Inferact

Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
MaintainX

Own the LLMX roadmap as a Staff Product Manager at MaintainX, driving AI platform adoption and quality in a hybrid role.

MaintainX San Francisco Published 3 months ago
MaintainX

Own the LLMX roadmap as a Staff Product Manager at MaintainX, driving AI platform adoption and quality.

MaintainX Toronto Published 1 month ago
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Profound

Join Profound as a Machine Learning Engineer to build and deploy large scale NLP and LLM systems in a fast-paced environment.

Profound New York, New York $180k–$260k/yr Published 1 year ago
Flexible on stack
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Omnifold

Join Omnifold as an ML Research Engineer to tackle complex supply chain challenges with innovative AI models.

Omnifold San Francisco HQ Published 5 months ago
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Horizon3.ai

Join Horizon3.ai as a Staff Attack Engineer to develop automated attacks for AI/LLM systems in a fully remote environment.

Horizon3.ai US, Remote $223k–$275k/yr Published 5 months ago
Flexible on stack
Coinbase

Join Coinbase as a Senior Staff Software Engineer to architect and build AI-driven legal automation systems.

Coinbase Remote - USA $253.9k–$298.7k/yr Published 2 months ago
Thumbtack

Join Thumbtack as a Staff Software Engineer to lead the development of AI/ML infrastructure for innovative home improvement solutions.

Thumbtack Remote, United States $212.5k–$275k/yr Published 3 days ago
TRM Labs

Join TRM Labs as a Staff Software Engineer to build AI-powered solutions for crime investigation in a fully remote environment.

TRM Labs United States $200k–$275k/yr Published 6 months ago
Flexible on stack
MoonPay

Own the decisioning system for real-time transaction scoring in a high-impact fraud detection role at MoonPay.

MoonPay London - Hybrid Published 2 weeks ago
Flexible on stack 70% coding
Thumbtack

Join Thumbtack as a Staff Software Engineer to lead the development of AI/ML infrastructure for innovative home improvement solutions.

Thumbtack Remote, Ontario $212.5k–$275k/yr Published 3 days ago
Bluefish AI

Lead the vision and execution of LLM-powered products at Bluefish AI, shaping the future of AI-driven marketing technologies.

Bluefish AI London - Remote in the UK Published 2 months ago
Flexible on stack
Block

Build production ML systems to transform customer behavior into trusted signals for decision-making at Block.

Block Bay Area, CA, United States of America $276.8k–$415.2k/yr Published 11 months ago
Flexible on stack