← Back to results

Applied AI Engineer, Kernel Performance

Build AI systems that autonomously optimize model architectures for production-ready implementations at Etched.

Location
San Jose
Compensation
$150k–$225k/yr
Level
senior
Type
full time · On-site

Posted by employer 2 months ago

First seen on Joblaze 4 days ago

Last verified on the company career page 13 hours ago

Apply at Etched → Save job Scanned from etched.com

Skills & Technologies

What you'll build

  • Own the system that turns new model architectures into verified kernels
  • Build agents that understand Etched hardware
  • Design evals covering correctness and efficiency
  • Turn profiler traces into structured signals
  • Ship model-generated improvements to production

Must have

  • Comfort with both Python and low-level code
  • Kernel experience
  • Fluency using AI to learn new problems

Nice to have

  • First principles thinking on accelerator performance
  • Hands-on experience building LLM-based agents
  • Eval-driven mindset

AI in the day-to-day

You will teach models using proprietary performance signals and iterate on their proposals.

Not disclosed in this posting: years of experience, visa sponsorship.

Benefits

Wellness Benefits Housing Subsidy Unlimited Compute Budget Daily lunch and dinner Health Insurance Relocation Assistance

Joblaze summary

The Applied AI Engineer at Etched focuses on developing AI systems that autonomously convert new model architectures into optimized, production-ready kernels tailored for the company's hardware. This role requires proficiency in both Python and low-level coding, alongside a strong background in kernel development and performance optimization. Ideal candidates are those who thrive in complex environments, possess a blend of research and practical execution skills, and are comfortable navigating ambiguity. Etched fosters a collaborative atmosphere where engineering and research intersect, enhancing the innovation process.

Joblaze insights

  • Listed 4 days ago — first seen on Joblaze September 21, 2026. Last confirmed on Etched's careers page September 25, 2026.
  • Salary band is below the typical range for AI/ML roles (median ~$190,000).
  • Starts above 14% of 205 comparable senior ai/ml roles in United States that list Python we track (median $180,000 across 68 companies). See Python salary trends
  • Python appears in 52.3% of 545 comparable senior ai/ml roles in United States; AI/ML appears in 45.9% of 545 comparable senior ai/ml roles in United States.

Quick facts

Is the Applied AI Engineer, Kernel Performance role remote?
No — this is an on-site role in San Jose.
What's the salary range?
Etched lists $150,000–$225,000 for this role.
Where is the role based?
Etched is hiring for this position in San Jose.
What's the tech stack?
Joblaze extracted these technologies from the posting: AI/ML, Python, kernel engineering.
What seniority level is this role?
Etched targets senior candidates for this position.
Is this full-time or contract?
Full-time for this Applied AI Engineer, Kernel Performance role at Etched.

From the original posting

Job Summary

Every model release presents a new opportunity to push the frontier on kernel engineering. Future performance breakthroughs will come from AI systems that can understand model architectures and hardware, run thousands of experiments, learn from profiler feedback, and discover the most performant implementations faster than the best engineers.

You will build that system. Your mandate is to build AI systems that autonomously turn newly released model architectures into correct, production-ready implementations optimized for Etched hardware. These systems should explore broader design spaces, learn from every experiment, and reach peak performance faster than any traditional kernel-development workflows.

Etched offers a uniquely tight research loop: proprietary hardware, runtime, kernels, production workloads, and dedicated in-office compute under one roof. You will teach models using proprietary performance signals, iterate on their proposals, and make every experiment improve both the performance optimization system and the hardware it runs on.

Key Responsibilities

  • Own the system that turns new model architectures into verified, production-ready kernels and model mappings.

  • Build agents that understand Etched hardware, design experiments, generate implementations, profile them, diagnose bottlenecks, and iterate with our teams, to the limits of model autonomy.

  • Design evals covering correctness, numerical stability, latency and efficiency.

  • Turn profiler traces, simulation, hardware counters, and expert judgment into structured signals models can learn from.

  • Curate proprietary datasets and optimization memory from complete trajectories, expert demonstrations, counterexamples, and production outcomes.

  • Build fast, reproducible experiment infrastructure and observability so experiments remain interpretable, trustworthy, and high-throughput.

  • Ship model-generated improvements to production and quantify their impact on end-to-end system performance.

  • Partner deeply with other architecture teams to shape new abstractions and Etched’s hardware-software roadmap.

  • Continuously evaluate new model releases and deploy the best for each stage of the optimization loop.

You may be a good fit if you have

  • A track record of solving hard problems across stacks and domains — you enjoy being dropped into unfamiliar territory and figuring it out

  • Comfort with both Python and low-level code: you can read it, modify it, debug it, and direct AI to write it well. We do not care whether you write code from scratch — we care whether you ship things that work.

  • Kernel experience: you've written or tuned kernels and can explain the mechanisms and performance impact of optimizations you’ve shipped

  • Fluency using AI to learn and ramp on new problems — agentic coding tools, deep research, and frontier models are how you work, not an add-on

  • Moving fluidly between research exploration, agentic experimentation, low-level debugging, and production execution.

Strong candidates may also have experience with

  • First principles thinking on accelerator performance: memory hierarchy, data movement, parallelism, synchronization, and low-precision computation.

  • Hands-on experience building and shipping LLM-based agents or AI tooling that real users depend on in production environments (beyond calling an API — context engineering, tool integration, orchestration, failure analysis)

  • An eval-driven mindset: you measure whether AI systems work before scaling them

  • Fine-tuning or post-training, RAG over proprietary data, and/or multi-agent orchestration

  • High agency and comfort with ambiguity — you find the real problem to solve

Benefits

  • Medical, dental, and vision packages with generous premium coverage

    • $500 per month credit for waiving medical benefits

  • Housing subsidy of $2,500 per month for those living within walking distance of the office

  • Daily lunch and dinner in our office

  • Unlimited compute budget subject to ROI justification

Base Compensation Range

  • $150,000 – $225,000

 

Standard company text repeated across Etched's postings is omitted here.

Similar positions

Etched
Applied AI Engineer - Hardware
Etched · San Jose, United States
Etched