← Back to results

AI Engineer - Cloud Infrastructure

Join Traversal as a senior AI Engineer to design and operate core systems for AI products in a fast-paced, collaborative environment.

Location
New York
Compensation
$175k–$300k/yr
Level
senior
Type
full time · On-site

Posted by employer 2 months ago

First seen on Joblaze 1 week ago

Last verified on the company career page 1 day ago

Apply at Traversal → Save job Scanned from traversal.com

Skills & Technologies

AI in the day-to-day

Traversal's AI SRE agent autonomously understands and reasons across complex production environments.

Requirements

Experience
7+ years

Not disclosed in this posting: visa sponsorship.

Benefits

Equity/Stock Options Flexible Time Off Health Insurance

Joblaze summary

In the role of AI Engineer - Cloud Infrastructure at Traversal, the individual will design and manage the core systems that underpin the company's AI products, focusing on scalable and secure infrastructure. Key skills include expertise in AWS, Kubernetes, Terraform, and Python, with a strong emphasis on maintaining high availability and resilience. This position is suited for seasoned professionals with over seven years of experience in technically demanding environments, particularly those familiar with cloud-native operations. Traversal's team is composed of top-tier talent from prestigious institutions and companies, fostering a collaborative culture aimed at tackling complex AI challeng

Joblaze insights

Quick facts

Is the AI Engineer - Cloud Infrastructure role remote?
No — this is an on-site role in New York.
What's the salary range?
Traversal lists $175,000–$300,000 for this role.
How much experience is required?
At least 7 years of relevant experience for this AI Engineer - Cloud Infrastructure role.
Where is the role based?
Traversal is hiring for this position in New York.
What's the tech stack?
Joblaze extracted these technologies from the posting: AWS, Helm, Kubernetes, Python, Terraform.
What seniority level is this role?
Traversal targets senior candidates for this position.
Is this full-time or contract?
Full-time for this AI Engineer - Cloud Infrastructure role at Traversal.

From the original posting

About Traversal

Traversal is the AI Site Reliability Engineer (AI SRE) for the enterprise.

Production complexity was already outpacing what engineering teams could manage manually, and AI-generated code is accelerating that gap. Traversal is built for that challenge, autonomously understanding and reasoning across even the largest, most complex production environments to diagnose, fix, and prevent incidents. Our mission is to free engineers from endless firefighting and give them more time to focus on creative, high-impact work.

Today, Traversal operates in mission-critical environments at some of the world’s largest enterprises. Our roots remain deeply embedded in AI research, and we’ve brought together researchers from institutions including MIT, Harvard, Berkeley, Columbia, and Cornell with world-class technical staff and operators from companies like Google, Meta, Datadog, ServiceNow, and Citadel Securities to take on one of the hardest problems for AI to solve. Traversal is backed by Sequoia Capital, Kleiner Perkins, Hanabi, NFDG, and American Express Ventures.

The Role

As an AI Engineer - Cloud Infrastructure on Traversal’s Infrastructure team, you’ll design, secure, and operate the core systems that power Traversal’s AI products. We already serve Fortune 50 enterprises with large-scale, multi-tenant environments, BYOC deployments, and SOC 2 Type II controls, and we’re rapidly scaling.

You’ll focus on the building blocks of our Terraform-defined infrastructure and Kubernetes environments, while supporting the complex needs of operating the highly-available, highly-resilient, and cost-efficient platform that supports the Traversal AI SRE agent.

This is a senior, high-impact role: you’ll own foundational systems, work across AWS-native infrastructure, cloud networking, Kubernetes environments, Terraform, Helm, Python, and more, shaping how enterprise AI reliability is built and scaled.

Responsibilities

  • System Design & Architecture: Design scalable, reliable infrastructure for AI workloads, inference, data pipelines, and agentic workflows

  • CI/CD: Build and deliver best-in-class developer experience and software development lifecycle tooling for our growing engineering team

  • Autoscaling: Scale on real signals (queue lag, in-flight requests, latency); add burst capacity and safe drains

  • Infrastructure as Code: Evolve Terraform+Helm for multi-environment deployments, secrets, policy-as-code, and workload identity

  • Observability: Build and deliver end-to-end visibility into our infrastructure, systems, and applications, and connect it to Traversal’s AI SRE agent for self-driving production

  • Security: Partner with our cloud security lead to improve Traversal’s security and compliance posture, implementing least privilege principles, JIT access workflows, default-deny egress, auditability, and policy-as-code

Requirements

  • 7+ years of experience at technically rigorous companies or teams

  • Proven experience operating cloud and Kubernetes native infrastructure and applications at scale with >99.9% availability

  • Demonstrated hands-on experience with AWS, EKS, Terraform, Helm

  • Experience designing idempotent systems (outbox, dedupe keys, safe replay)

  • Incident response, chaos testing, capacity planning

  • Strong debugging skills across infrastructure, compute, network, runtime, storage, and auth layers

Nice to Have

  • Service mesh (Envoy/Istio), Cilium/eBPF

  • GPU workload operations, inference servers, token streaming gateways

  • Production experience building and maintaining systems in Python, Rust, and TypeScript

  • Data governance (PII discovery/redaction), lineage, tokenization

  • Experience designing, implementing, and deploying cross-region active/active architectures

  • Familiarity with other cloud providers (GCP, Azure, Oracle Cloud)

Compensation

We offer competitive compensation, startup equity, health insurance, and additional benefits. The U.S. base salary range for this full-time, in-person role in New York is $175,000–$300,000, plus equity and benefits. Our salary ranges are based on location, level, and role. Individual compensation is determined by experience, skills, and job-related knowledge.

Why You Should Join Us

Traversal is a place to take on hard, meaningful problems with real ownership from day one. You’ll work alongside people who challenge you to grow, learn constantly, and help define a new category of infrastructure software. We think long term, move quickly, and hold a high bar without taking ourselves too seriously.

We offer competitive salary and equity packages, health insurance, fertility benefits, a great tech setup stipend and flexible time off. Plus in-office snacks, team happy hours and outings, an annual company offsite, and plenty of built in time to collaborate across teams.

Traversal is fully in-office, 5 days a week, based in New York near Madison Square Park. We have a collaborative, hard-working culture and are energized by building the future of AI-powered software maintenance.

Similar positions

Traversal
AI Engineer - AI Platform
Traversal · New York
Traversal
AI Engineer - Backend
Traversal · New York
Traversal
AI Research Engineer
Traversal · New York
Traversal
Traversal
Solutions Engineer - East
Traversal · Remote