xelys jobs xelys jobs

Site Reliability Engineer

Origami Risk

hybridseniorpermanentdevopsbackend United States Today via LinkedIn
100,000 - 120,000 USD/annual

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

Site Reliability EngineeringIncident ManagementObservabilityNew RelicDatadogSumo LogicAWSAzureCI/CDSQL

About the role

Role Overview

Origami Risk is hiring a Site Reliability Engineer (SRE) to improve time to resolution, strengthen site reliability and scalability, and drive prevention through incident and performance investigations.

Responsibilities

  • Lead post-incident investigations and perform in-depth root-cause analysis
  • Draft clear RCAs for customer delivery
  • Develop and share preventive strategies to reduce future disruptions
  • Cross-train teammates on leveraging observability tools during incident and performance work
  • Provide end-to-end visibility to stakeholders across the SRE process
  • Collaborate cross-functionally to implement scalability and stability enhancements
  • Build client-focused dashboards/alerts to proactively identify performance challenges
  • Monitor and continuously improve time to resolution metrics
  • Maintain and configure observability tools to ensure key data is available for investigations
  • Create an actionable feedback loop to improve MELT and development patterns
  • Contribute to automation tools to streamline incident response
  • Proactively prevent incidents and reduce platform impact

Requirements

  • Bachelor’s degree in Computer Science (or related) or equivalent experience
  • 5+ years of proven SRE experience
  • Strong knowledge of SRE best practices and incident management protocols
  • Deep experience with and/or configuring observability tools such as New Relic, Datadog, Sumo Logic (or similar)
  • Ability to read/write code: JavaScript, .NET, SQL
  • Familiarity with cloud platforms (e.g., AWS, Azure) and architectural patterns
  • Excellent problem-solving and data-driven incident analysis
  • Prior experience operating in public cloud environments (AWS strongly preferred)
  • Experience troubleshooting C#/.NET web applications
  • Solid knowledge of SaaS operations
  • Ability to succeed under ambiguity and varying operational maturity
  • Strong written and verbal communication skills

Nice to Have

  • Windows and SQL Server troubleshooting skills
  • Experience with CI/CD pipelines
  • Experience with Infrastructure as Code (IaC)

About Origami Risk

Origami Risk is a SaaS company focused on helping organizations manage risk. It operates and supports its platforms in a public-cloud environment, with an emphasis on reliability, observability, and scalable operations.

Scraped 7/28/2026