xelys jobs xelys jobs

Staff Site Reliability Engineer

Sprinter Health

full-remoteleadpermanentdevopssecurity Full remote 58 days ago via WTTJ

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

Site Reliability Engineering (SRE)Platform EngineeringAWSGCPTerraformInfrastructure as CodeObservabilityMonitoringCloud SecurityIncident Response

About the role

Role Overview

Join Sprinter Health as a Staff Site Reliability Engineer. You will build the reliability, infrastructure, and security foundations that power large-scale, last-mile healthcare delivery. The role has broad ownership across reliability, cloud infrastructure, security, observability, automation, and platform design.

Key Missions

  • Design, build, and improve the infrastructure powering Sprinter’s patient care, clinician operations, internal tooling, and partner-facing systems.
  • Raise the security baseline across cloud infrastructure, including access controls, secrets management, identity, and operational workflows.
  • Partner with engineering teams to improve system architecture, deployment practices, monitoring, logging, and alerting.

Responsibilities

  • Automate infrastructure, deployment, and operational workflows using scripting languages.
  • Operate and improve production systems in cloud environments.
  • Make technical decisions balancing short-term business needs with long-term scalability, reliability, and maintainability.
  • Troubleshoot production issues across application, infrastructure, networking, and deployment layers.
  • Lead end-to-end infrastructure, reliability, platform, or security projects with minimal oversight.
  • Improve observability and incident response practices across teams.
  • Contribute to establishing reliability standards and platform patterns across the org.

Requirements

  • 8+ years in Site Reliability Engineering, platform engineering, infrastructure engineering, security engineering, or related roles.
  • Experience automating workflows using Python, Bash, or TypeScript.
  • Built and operated production systems in AWS and/or GCP.
  • Demonstrated experience leading high-impact infrastructure/reliability/security initiatives end to end.
  • Troubleshooting experience across application, infrastructure, networking, and deployment layers.
  • Experience improving cloud security posture (access management, secrets management, networking, operational controls).
  • Deep experience with Infrastructure as Code, ideally Terraform.
  • Comfort working in environments requiring reliability, security, ambiguity, and speed.
  • Strong observability/monitoring/logging/alerting and incident response improvements.
  • Experience working in mid- or growth-stage startups and operationally complex domains.
  • Experience improving security posture in a practical, engineering-friendly way.
  • Mentoring experience and ability to raise the operational bar.
  • Comfortable partnering with product engineers, data teams, operations, security stakeholders, and technical leadership.
  • Experience in regulated environments with privacy, security, and compliance best practices.

Nice-to-haves

  • People management experience or interest in growing into broader technical leadership.

About Sprinter Health

Sprinter Health is a high-growth healthcare company building infrastructure for last-mile healthcare delivery. The role focuses on operational reliability, security, and platform foundations that support patient care, clinician workflows, and internal/partner-facing systems.

Scraped 6/11/2026