← Back to results

Senior Site Reliability Engineer

Join Airbyte as a Senior Site Reliability Engineer to enhance infrastructure and reliability for a data replication platform.

Location
San Francisco
Compensation
Not disclosed
Level
senior
Type
full time · On-site

Posted by employer 1 month ago

First seen on Joblaze 3 months ago

Last verified on the company career page 9 hours ago

AI in the day-to-day

Engineers actively use AI as a force multiplier to automate toil, augment incident response, and build smarter internal tooling.

Requirements

Experience
7+ years

Not disclosed in this posting: compensation, visa sponsorship.

Benefits

401k Match Education Budget Flexible PTO Commuter Benefits Meals Provided Health Insurance Parental Leave

Joblaze summary

In this role, the Senior Site Reliability Engineer will focus on maintaining and enhancing the infrastructure for Airbyte's Data Replication platform, ensuring reliability and efficiency across millions of sync jobs. Key skills include expertise in Kubernetes, Terraform, and observability tools, along with a strong understanding of CI/CD processes. This position is ideal for seasoned professionals with a background in infrastructure or DevOps who thrive in fast-paced, startup environments. The team emphasizes the use of AI to streamline operations and improve tooling.

Joblaze insights

  • Listed about 3 months ago — first seen on Joblaze July 1, 2026. Last confirmed on Airbyte's careers page October 9, 2026.
  • Kubernetes appears in 67.9% of 190 comparable senior devops/sre roles in United States; Datadog appears in 13.7% of 190 comparable senior devops/sre roles in United States.

Quick facts

Is the Senior Site Reliability Engineer role remote?
No — this is an on-site role in San Francisco.
How much experience is required?
At least 7 years of relevant experience for this Senior Site Reliability Engineer role.
Where is the role based?
Airbyte is hiring for this position in San Francisco.
What's the tech stack?
Joblaze extracted these technologies from the posting: AWS, Datadog, GCP, Grafana, Kubernetes, Prometheus.
What seniority level is this role?
Airbyte targets senior candidates for this position.
Is this full-time or contract?
Full-time for this Senior Site Reliability Engineer role at Airbyte.

From the original posting

 

The Role:

You'll be the infrastructure and reliability engineer on the Data Replication team - a full-stack product team running over 3 million sync jobs a week powering thousands of data use cases across multiple regions and clouds. You’ll build and maintain the infrastructure, set reliability standards, drive down incidents, and make it easier and safer for engineers to ship through tooling. You're equally comfortable in a Terraform file, a Kubernetes cluster, and a postmortem doc.


We expect engineers here to actively use AI as a force multiplier - agentic tools to automate toil, augment incident response, and build smarter internal tooling. If you're not already doing this, you should be excited to start. We care as much about how you work as what you build. Trust, directness, and craftsmanship matter here.

 

What You’ll Do:

  • Own the infrastructure underpinning the Data Replication platform - Kubernetes clusters, CI/CD pipelines, secrets management, networking, and cloud resource configuration across AWS and GCP.

  • Partner with product engineers to reliably integrate product features with infrastructure.

  • Maintain and enhance observability, alerting, and anomaly detection with an eye towards LLM automation.

  • Maintain and enhance AI-augmented release and internal tooling: canary deployments, progressive rollouts, automated release qualification, and rollback automation - with an eye towards LLM automation.

  • Set the infrastructure bar for the team - build self-serve tooling, write runbooks, and coach engineers to own more of their stack.

 

What You’ll Need:

  • 7+ years in infrastructure, platform engineering, SRE, or DevOps.

  • Hands-on ownership of Kubernetes, Helm, and Terraform in production environments.

  • Deep experience with observability stacks (Prometheus, Grafana, Datadog) and on-call operations.

  • Experience with CI/CD pipeline ownership and developer tooling.

  • Ability & willingness to read backend code to understand how systems break and instrument them correctly.

  • Fluency with AI tools - LLMs and agentic frameworks to automate, debug faster, and reduce toil.

  • A startup-ready mindset: comfortable with ambiguity, moving fast, and owning problems end-to-end.

 

Nice To Have:

  • Data pipelines, replication systems, or ETL/ELT platforms.

  • Control plane / data plane architectures or internal developer platforms.

  • Experience with Airbyte, CDKs, or connector-based architectures.

 

Location:

  • Onsite 4 days/week in San Francisco, CA

  • Flexible PTO with a culture that encourages at least 25 days off annually

  • 401(k) retirement plan

  • Commuter benefits and monthly internet reimbursement

Standard company text repeated across Airbyte's postings is omitted here.

Similar positions

Airbyte
Engineering Manager, Platform
Airbyte · San Francisco
Sprinter Health
AI Automation Team - Software Engineer (Senior)
Sprinter Health · San Francisco, CA
Airbyte
Senior Solutions Architect
Airbyte · New York
Sprinter Health
Product Engineering Team - Software Engineer (Senior)
Sprinter Health · San Francisco, CA