← Back to results

Site Reliability Engineer, Data Center Infrastructure

Join SpaceX as a Site Reliability Engineer to enhance the reliability and scalability of manufacturing systems for interplanetary exploration.

Location
Bastrop, TX, United States
Compensation
Not disclosed
Level
mid
Type
full time · On-site

Posted by employer 15 hours ago

First seen on Joblaze 9 hours ago

Last verified on the company career page 9 hours ago

What you'll build

  • Deploy, upgrade, operate, maintain, and scale compute, storage, and networking for manufacturing systems
  • Manage infrastructure as code and use observability to provide a complete picture of platform health
  • Design for reliability, stability, and scale
  • Practice proactive maintenance: capacity planning and lifecycle management
  • Provide high-quality support to manufacturing and engineering users

Must have

  • Bachelor’s degree in computer science, information systems, or an engineering discipline; OR 3+ years of professional experience in SRE or DevOps
  • 1+ years of software development experience
  • Experience with Linux operating systems

Nice to have

  • Experience with compute, storage, and/or networking infrastructure in production
  • Infrastructure as code (Terraform, Ansible, Puppet, or similar)
  • Containers and virtualization (Docker, Kubernetes, vSphere, QEMU, KVM, etc.)
  • Databases and data modeling (Postgres, Clickhouse, etc.)
  • Comfort with mission-critical systems

Practical constraints

  • Must be able to work extended hours and weekends as needed
  • Must be able to travel to different sites

Requirements

Experience
3+ years
Education
Bachelor's degree
Visa
No sponsorship (stated in posting)

Not disclosed in this posting: compensation.

Joblaze summary

The Site Reliability Engineer for Data Center Infrastructure at SpaceX focuses on maintaining and scaling the compute, storage, and networking systems that support critical manufacturing operations for projects like Starship and Starlink. Key skills include software development, infrastructure as code, and experience with Linux systems, alongside a proactive approach to reliability and incident management. This role is suited for engineers with a solid foundation in software and infrastructure, who thrive in collaborative environments and are eager to tackle complex challenges. The position emphasizes ownership and communication, ensuring that manufacturing processes run smoothly and efficie

Joblaze insights

  • Listed today — first seen on Joblaze October 9, 2026. Last confirmed on SpaceX's careers page October 9, 2026.
  • Terraform appears in 61.5% of 117 comparable mid devops/sre roles in United States; KVM appears in 2.6% of 117 comparable mid devops/sre roles in United States.

Quick facts

Is the Site Reliability Engineer, Data Center Infrastructure role remote?
No — this is an on-site role in Bastrop, TX, United States.
How much experience is required?
At least 3 years of relevant experience for this Site Reliability Engineer, Data Center Infrastructure role.
Where is the role based?
SpaceX is hiring for this position in Bastrop, TX, United States.
What's the tech stack?
Joblaze extracted these technologies from the posting: Ansible, ClickHouse, Docker, KVM, Kubernetes, Linux.
What seniority level is this role?
SpaceX targets mid-level candidates for this position.
Is this full-time or contract?
Full-time for this Site Reliability Engineer, Data Center Infrastructure role at SpaceX.

From the original posting

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.

SITE RELIABILITY ENGINEER, DATA CENTER INFRASTRUCTURE

The application software team is the central nervous system of SpaceX. Manufacturing is how SpaceX turns designs into hardware. The compute, storage, and networking that run our factories must be as reliable as the products we build. This team owns infrastructure supporting Starship, Starlink, Starshield, and Terafab. This position will have a direct impact on factory uptime, throughput, and production scale across programs.

The ideal candidate has strong software engineering fundamentals and a passion for infrastructure: reliability, stability, proactive maintenance, and scalability. You understand the system before you change it, solve hard problems, communicate clearly with stakeholders and teammates, and take ownership of work that manufacturing depends on.

Aerospace experience is not required. We value smart, motivated, collaborative engineers who treat teammates with fairness, respect, and support, and who want to take full ownership of challenging problems to help make humanity multi-planetary.

RESPONSIBILITIES:

  • Deploy, upgrade, operate, maintain, and scale compute, storage, and networking for manufacturing systems across Starship, Starlink, Starshield, and Terafab
  • Manage infrastructure as code and use observability to provide a complete picture of platform health
  • Design for reliability, stability, and scale; find and remove bottlenecks with measurement and engineering
  • Practice proactive maintenance: capacity planning, lifecycle management, and reducing toil before it becomes an incident
  • Partner with software engineers, manufacturing stakeholders, and site teams to build operable, maintainable systems
  • Improve the full lifecycle—from design through deployment, operation, and continuous refinement
  • Practice sustainable incident response and blameless postmortems
  • Provide high-quality support to manufacturing and engineering users
  • Communicate clearly with stakeholders and teammates
  • Participate in on-call and travel to sites as needed for deployments, incidents, and cross-site reliability

BASIC QUALIFICATIONS:

  • Bachelor’s degree in computer science, information systems, or an engineering discipline; OR 3+ years of professional experience in SRE or DevOps in lieu of a degree
  • 1+ years of software development experience
  • Experience with Linux operating systems

PREFERRED SKILLS AND EXPERIENCE:

  • Experience with compute, storage, and/or networking infrastructure in production
  • Infrastructure as code (Terraform, Ansible, Puppet, or similar)
  • Containers and virtualization (Docker, Kubernetes, vSphere, QEMU, KVM, etc.)
  • Databases and data modeling (Postgres, Clickhouse, etc.)
  • Ability to translate high-level requirements into implementations from first principles
  • Comfort with mission-critical systems and appropriate urgency and care
  • Skillful communication with customers, peers, and management
  • Comfort operating across multiple sites and manufacturing programs

ADDITIONAL REQUIREMENTS:

  • Must be able to work extended hours and weekends as needed
  • Must be able to travel to different sites (Hawthorne, CA; Redmond, WA; Cape Canaveral, FL; Starbase, TX)
  • Ability to pass Air Force background check for Cape Canaveral
  • This role requires you to be onsite. Remote and/or hybrid work will not be considered

ITAR REQUIREMENTS:

Standard company text repeated across SpaceX's postings is omitted here.