Join Databricks as an IT Software Engineer to build scalable infrastructure solutions in a rapidly growing environment.
Posted by employer 18 hours ago
First seen on Joblaze 5 hours ago
Last verified on the company career page 5 hours ago
Not disclosed in this posting: compensation, work arrangement, visa sponsorship.
Joblaze summary
In the role of IT Software Engineer for Infrastructure at Databricks, the individual will focus on enhancing the company's core infrastructure and observability platforms, ensuring they are resilient and scalable. Key skills include proficiency in Python, experience with Infrastructure as Code tools like Terraform, and familiarity with cloud environments such as AWS, alongside container orchestration using Kubernetes. This position is suited for professionals with 3-5 years of software engineering experience who can independently manage technical projects. The team operates within a product-led organization that emphasizes collaboration across various technical domains.
Quick facts
- How much experience is required?
- 3–5 years of relevant experience for this IT Software Engineer, Infrastructure role.
- What's the tech stack?
- Joblaze extracted these technologies from the posting: AWS, Datadog, GitHub Actions, Grafana, Kubernetes, Prometheus.
- What seniority level is this role?
- Databricks targets mid-level candidates for this position.
- Is this full-time or contract?
- Full-time for this IT Software Engineer, Infrastructure role at Databricks.
From the original posting
GAQ327R265
About the Role
At Databricks Information Technology, we are a product-led organisation transforming the way we work, from how easy it is to use our IT services to the applications we develop to help us scale seamlessly amid incredible growth.
As an IT Software Engineer (Infrastructure), you will be a core technical contributor on the IT CorpEng Infrastructure team, owning and driving the evolution of our core infrastructure and observability platforms. This role requires a strong software engineering mindset, deep technical breadth across SRE and infrastructure worlds, and the ability to deliver high-quality, scalable solutions for currently "immature" system problems. You will be responsible for building resilient, scalable, and automated infrastructure that empowers our development teams. As a member of the team, you will bridge the gap between software engineering and systems architecture, ensuring our AWS environment is cost-optimised, secure, and highly available.
The impact you will have:
- Automate: Develop and deploy production-grade infrastructure on cloud platforms (AWS preferred) using Infrastructure as Code tools like Terraform or Pulumi.
- Orchestration: Deploy, manage, and scale containerised workloads using Kubernetes (EKS/AKS), focusing on security, reliability, and resource efficiency.
- CI/CD Excellence: Build and maintain robust deployment pipelines using GitHub Actions, including both cloud-hosted and self-hosted runners for specialised build requirements.
- Drive "Observable by Default" Frameworks: Implement infrastructure standards to ensure new internal applications are secure, with logging, metrics, and tracing configured out of the box.
- Tooling, Scripting & AI: Build internal CLI tools, AI-assisted workflows, and automation scripts to streamline developer workflows and reduce manual toil.
- Partner Cross-Functionally: Collaborate with Security, Engineering, Infrastructure, and Support teams to execute technical projects aligned with business outcomes.
- Document & Knowledge Share: Participate in peer code reviews, maintain operational playbooks/failure triage guides, and help onboard team members onto platforms you maintain.
What we look for:
- Software Engineering Expertise: 3–5 years of hands-on software engineering experience, with strong proficiency in Python for scripting, automation, and tooling (required).
- Infrastructure as Code (IaC): Solid working knowledge of Terraform (modules, state management) or Pulumi (preferred).
- Cloud & Containerization: Hands-on experience building and supporting production workloads in AWS (or Azure/GCP), along with Kubernetes and Docker fundamentals.
- CI/CD & Automation: Practical experience designing, maintaining, and troubleshooting pipelines using GitHub Actions or similar modern CI/CD tools.
- Observability Foundations: Practical understanding of observability pillars (logging, metrics, tracing) and hands-on experience integrating tools such as Datadog, Prometheus, or Grafana.
- Independent Execution: Proven ability to own end-to-end tasks with guidance from Tech Leads, breaking down broader technical goals into actionable technical solutions.
Standard company text repeated across Databricks's postings is omitted here.