← Back to results

Node Systems Lead

Lead the Node Systems team to optimize software for inference clusters at gigawatt-scale in a startup focused on frontier intelligence.

Location
San Jose, United States
Compensation
Not disclosed
Level
lead
Type
full time · On-site

Posted by employer 19 hours ago

First seen on Joblaze 17 hours ago

Last verified on the company career page 17 hours ago

Apply at Etched → Save job Scanned from etched.com

Skills & Technologies

C++ Rust Linux Flexible on stack

What you'll build

  • Lead and develop the Node Systems team
  • Set the technical direction for host software
  • Guide performance work across CPU scheduling
  • Ensure system optimizations become tested software
  • Partner with inference software and hardware teams

Must have

  • Strong experience developing production systems software in C, C++ or Rust on Linux
  • Experience leading technical projects and mentoring engineers
  • Deep understanding of operating systems fundamentals
  • Demonstrated ability to investigate performance problems
  • Experience taking system software from prototype to production

Nice to have

  • Experience with low-latency systems
  • Experience with Linux kernel development
  • Familiarity with PCIe, DMA, RDMA
  • Experience bringing up new hardware platforms
  • Experience building systems at an early-stage startup

Role intensity

40% coding

Not disclosed in this posting: compensation, years of experience, visa sponsorship.

Benefits

Wellness Benefits Housing Subsidy Daily lunch and dinner Health Insurance Relocation Assistance

Joblaze summary

The Node Systems Lead at Etched is responsible for overseeing the software that enables the deployment of large-scale inference clusters, focusing on optimizing system performance across various components. This role requires expertise in C, C++, or Rust on Linux, along with a strong grasp of operating systems and hardware interactions. Ideal candidates will have a background in leading technical projects and a hands-on approach to problem-solving, particularly in production environments. Etched emphasizes a collaborative culture where engineering and research intersect, fostering innovation in frontier intelligence.

Joblaze insights

  • Listed today — first seen on Joblaze October 7, 2026. Last confirmed on Etched's careers page October 7, 2026.
  • C++ appears in 10% of 100 comparable lead ai/ml roles in United States; Rust appears in 3% of 100 comparable lead ai/ml roles in United States.

Quick facts

Is the Node Systems Lead role remote?
No — this is an on-site role in San Jose, United States.
Where is the role based?
Etched is hiring for this position in San Jose, United States.
What's the tech stack?
Joblaze extracted these technologies from the posting: C++, Linux, Rust.
What seniority level is this role?
Etched targets lead candidates for this position.
Is this full-time or contract?
Full-time for this Node Systems Lead role at Etched.

From the original posting

Job Summary

We’re hiring a Node Systems Lead to join our Supercomputing organization. This team builds the software that enables Etched to deploy inference clusters at gigawatt-scale. This role presents an opportunity to shape how frontier inference hardware is configured and managed for our customers.

We co-design chips, racks, software, and manufacturing methods so frontier models can run with best-in-class throughput, latency, cost, and power efficiency for both prefill and decode workloads. Node Systems owns the software layer of each individual Etched node, from host software and system configuration to the interfaces between rack components. It tunes CPU scheduling, memory, networking, and host-to-accelerator communication so models get every bit of performance and reliability our hardware can deliver, then turns those gains into tested software and configurations that ship in every system.

We’re looking for a leader who can set the technical direction and dive deep into difficult systems problems. Someone who has taken hardware from early bring up into production, can build the best team in the industry, and wants to stay close to architecture and code. The team’s work is central in enabling Etched to deploy inference clusters at massive scale.

Key Responsibilities

  • Lead and develop the Node Systems team, setting priorities and giving engineers clear ownership.

  • Set the technical direction for host software, system configuration and rack component interfaces.

  • Guide the team’s performance work across CPU scheduling, memory, networking and host-to-accelerator communication.

  • Ensure system optimizations become tested, maintainable software and configurations ready for deployment.

  • Lead the development of rack simulation and diagnostic capabilities that support platform development and debugging.

  • Partner with inference software, firmware, and hardware teams to architect our system design for current and next-gen products

  • Align with Fleet Software on the configurations and interfaces needed to manage and monitor deployed systems at the cluster-level

  • Work with manufacturing and test engineering to establish the software baseline and test coverage needed to ship reliable systems.

  • Stay close to the implementation through design reviews, code contributions and hands-on debugging.

You may be a good fit if you have (Must-have qualifications)

  • Strong experience developing and debugging production systems software in C, C++ or Rust on Linux.

  • Experience leading technical projects and mentoring engineers while staying hands on.

  • Deep understanding of operating systems fundamentals, including scheduling, concurrency, memory management and I/O.

  • Demonstrated ability to investigate performance problems on real hardware and validate improvements through profiling and measurement.

  • Experience taking system software or performance improvements from prototype through production deployment.

  • Understanding of hardware/software interactions and the ability to debug across application, kernel and device boundaries.

  • Ability to work closely with hardware and software teams and translate workload requirements into concrete system changes.

Strong candidates may also have experience with (Nice-to-have qualifications)

  • Experience with low-latency systems, high-frequency trading, HPC, or accelerator-based compute platforms.

  • Experience with Linux kernel development or debugging, CPU isolation, NUMA, interrupt affinity, and performance tuning.

  • Familiarity with PCIe, DMA, RDMA, device drivers, or high-performance networking.

  • Experience bringing up new hardware platforms or working closely with firmware and hardware engineers.

  • Exposure to manufacturing diagnostics, factory testing, or production system validation.

  • Experience building systems or engineering teams at an early-stage startup.

Benefits

  • Medical, dental, and vision packages with generous premium coverage

    • $500 per month credit for waiving medical benefits

  • Housing subsidy of $2,500 per month for those living within walking distance of the office

  • Daily lunch and dinner in our office

  • Unlimited compute budget subject to ROI justification

How we’re different

Etched believes in the Bitter Lesson. We are the first inference-focused frontier AI system, betting early on transformer and transformer-like architectures and on increasing model sizes. Our addressable market is the entirety of inference, unlike many of our competitors.

Standard company text repeated across Etched's postings is omitted here.

Similar positions

Etched
Head of Supercomputing
Etched · San Jose
Etched
Fleet Software Lead
Etched · San Jose, United States
Etched
Supercomputing Engineer
Etched · San Jose
Etched
Core Software Engineer
Etched · San Jose