xelys jobs xelys jobs

Site Reliability Engineer

Haystack

seniorpermanentdevops United States Yesterday via LinkedIn
230,000 - 250,000 USD/annual

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

Site Reliability EngineeringSLI/SLO/Error BudgetsIncident ResponseRoot Cause AnalysisKubernetesTerraformPythonGoLinuxObservability

About the role

Role: Site Reliability Engineer

Design, build, and maintain highly available production systems for mission-critical environments.

Responsibilities

  • Define and manage reliability metrics: SLIs, SLOs, and error budgets
  • Automate operations to remove manual processes
  • Build monitoring, alerting, and observability solutions
  • Improve system performance, capacity, and resilience
  • Lead incident response and perform root cause analysis (RCA)

Requirements

  • 12+ years of infrastructure/cloud engineering experience
  • 5–10+ years of engineering experience with strong Linux and Windows systems background
  • Kubernetes and container platform expertise
  • Experience working with cloud infrastructure environments
  • Proficiency in Python and Go
  • Hands-on experience with Terraform and automation tools

Compensation

  • $230,000 – $250,000 per year

About Haystack

Haystack appears to be a hiring partner or platform for a technology solutions provider focused on delivering innovative and secure solutions for critical government missions. The role centers on operating and improving highly available infrastructure and cloud systems to support mission-critical workloads.

Scraped 7/27/2026