Engineering Manager, Site Reliability
Radar
hybridmidpermanentengineering-managementdevops United States Today via LinkedIn
See how well this job matches your profile
Sign up to get an AI match score and generate a tailored application in seconds.
Get your match scoreTags
Site Reliability Engineering (SRE)Engineering ManagementAWS EKSTerraformCloudWatchGrafanaPagerDutyCircleCIMulti-RegionInfrastructure Reliability
About the role
Role Overview
Engineering Manager, Site Reliability (SRE) to lead a distributed SRE team responsible for Radar’s production infrastructure. You’ll drive a roadmap that improves scalability and high availability worldwide, including multi-region deployment.
Responsibilities
- Lead and support a distributed SRE team.
- Drive infrastructure roadmap to make Radar services scalable and highly available.
- Help enhance deployments from multi-availability-zone to multi-region.
- Ensure production reliability for a high-throughput, data-intensive platform processing 1B+ API calls per day.
Tech & Infrastructure Stack
- Infrastructure as Code: Terraform
- Orchestration/Compute: AWS EKS
- Database: MongoDB (Atlas)
- CI/CD: CircleCI
- Monitoring/Alerting: CloudWatch, Grafana, Pingdom, PagerDuty
- DNS: Cloudflare
- Primary server languages: TypeScript, Rust
- Data pipelines: Airflow, Scala Spark
Requirements
- Experience leading SRE/operations or reliability-focused engineering teams.
- Ability to drive reliability and infrastructure roadmaps in a production environment.
Nice-to-haves
- Experience with multi-region or high-availability architectures.
- Familiarity with the stack and tools listed above (AWS/EKS, Terraform, MongoDB Atlas, CI/CD, monitoring/alerting).
About Radar
Radar is a geolocation company providing geofencing SDKs, maps APIs, and AI-enabled solutions for marketing, fraud, and operations teams. It serves large-scale, high-throughput infrastructure used across hundreds of millions of devices, supporting worldwide production systems.
Scraped 7/31/2026