xelys jobs xelys jobs

DevOps Engineer

Santa Clara University Leavey School of Business

seniorpermanentdevopsbackend Washington, DC 74 days ago via LinkedIn
155,000 - 175,000 USD/annual

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

DevOpsSite Reliability EngineeringInfrastructure as CodeCI/CDMonitoringAlertingSLOs/SLIsHigh AvailabilitySecurity Best PracticesRegTech

About the role

Role Overview

DevOps Engineer to drive DevOps and Site Reliability Engineering initiatives at a growing RegTech SaaS company. You will design automated deployment processes and improve infrastructure reliability, scalability, security, and performance.

Responsibilities

  • Design, implement, and maintain automated deployment and configuration management systems
  • Build infrastructure as code (IaC) for provisioning and managing infrastructure
  • Continuously optimize deployment processes for efficiency and reliability
  • Implement and maintain monitoring, alerting, and logging for service availability and reliability
  • Collaborate to define SLOs and SLIs and support error budgets
  • Lead blameless post-incident reviews and drive reliability improvements
  • Implement high-availability systems and influence system design for scalability/performance
  • Build and manage CI/CD pipelines and ensure seamless integration of automated testing
  • Partner with security teams to implement and monitor security best practices and compliance
  • Perform capacity planning and identify performance bottlenecks
  • Create and maintain infrastructure/deployment documentation, runbooks, and knowledge base articles
  • Stay current with emerging technologies and industry trends

Requirements

  • Bachelor’s degree in Computer Science, Software Engineering, or related field
  • 8+ years of DevOps experience with a focus on Site Reliability Engineering (or equivalent combination of education/experience)
  • Strong IaC experience for highly available, fault-tolerant, scalable infrastructure
  • Strong CI/CD pipeline expertise (automation of testing, integration, and deployment)
  • Experience managing cloud and on-prem infrastructure (performance, scalability, cost, and security compliance)
  • Ability to produce clear documentation and regular reporting (health/performance/cost metrics)
  • Excellent communication skills across technical and non-technical stakeholders
  • Strong analytical/problem-solving skills and ability to meet tight deadlines

Nice-to-haves

  • Experience with security/regulatory compliance standards related to infrastructure and deployment
  • Evidence of mentoring or leading reliability practices (runbooks, error budgets, incident review discipline)

Scraped 5/13/2026