Site Reliability Engineer
Haystack
seniorpermanentdevops United States Yesterday via LinkedIn
230,000 - 250,000 USD/annual
See how well this job matches your profile
Sign up to get an AI match score and generate a tailored application in seconds.
Get your match scoreTags
Site Reliability EngineeringSLI/SLO/Error BudgetsIncident ResponseRoot Cause AnalysisKubernetesTerraformPythonGoLinuxObservability
About the role
Role: Site Reliability Engineer
Design, build, and maintain highly available production systems for mission-critical environments.
Responsibilities
- Define and manage reliability metrics: SLIs, SLOs, and error budgets
- Automate operations to remove manual processes
- Build monitoring, alerting, and observability solutions
- Improve system performance, capacity, and resilience
- Lead incident response and perform root cause analysis (RCA)
Requirements
- 12+ years of infrastructure/cloud engineering experience
- 5–10+ years of engineering experience with strong Linux and Windows systems background
- Kubernetes and container platform expertise
- Experience working with cloud infrastructure environments
- Proficiency in Python and Go
- Hands-on experience with Terraform and automation tools
Compensation
- $230,000 – $250,000 per year
About Haystack
Haystack appears to be a hiring partner or platform for a technology solutions provider focused on delivering innovative and secure solutions for critical government missions. The role centers on operating and improving highly available infrastructure and cloud systems to support mission-critical workloads.
Scraped 7/27/2026