Site Reliability Engineer (6266)
itD
full-remotemidfixed-termdevopsbackend United States Yesterday via LinkedIn
See how well this job matches your profile
Sign up to get an AI match score and generate a tailored application in seconds.
Get your match scoreTags
Site Reliability EngineeringSREAnsibleRubyRSpecCI/CDGitLab CILinuxAWSKubernetes
About the role
Role Overview
Site Reliability Engineer to improve reliability, scalability, and operational efficiency of large-scale cloud infrastructure. You will build automation and resilient operational tooling, and support scalable deployment pipelines for production environments.
Responsibilities
- Develop and maintain infrastructure automation solutions using Ansible.
- Design, implement, and enhance CI/CD pipelines, testing frameworks, and operational tooling.
- Troubleshoot complex Linux-based infrastructure and distributed systems issues to maintain high availability and performance.
- Build automation for rapid, repeatable deployment of regional, sovereign, and purpose-built cloud environments.
- Collaborate with engineering, product management, and cross-functional stakeholders to identify operational improvements.
- Monitor infrastructure performance and implement enhancements that reduce overhead and improve scalability.
- Contribute to automation and engineering best practices for reliable cloud platform operations.
Required Qualifications
- 2+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering supporting cloud production.
- Hands-on experience building and maintaining infrastructure automation with Ansible.
- Programming experience in Ruby and automated testing with RSpec (or comparable frameworks).
- Ability to administer and troubleshoot Linux systems and distributed infrastructure.
- CI/CD experience including GitLab CI.
- Experience supporting large-scale infrastructure (hundreds to thousands of systems).
- Must be eligible to work on FedRAMP projects.
- Must be a U.S. citizen working from U.S. soil.
Preferred Qualifications
- Experience with AWS or other public cloud platforms and hybrid infrastructure.
- Knowledge of monitoring/observability and SRE practices/tools.
- Familiarity with Kubernetes and containerized platforms.
- Experience using AI-assisted development tools to improve automation and engineering productivity.
Logistics / Employment Details
- 100% remote within the United States.
- Duration: 24 months.
- Direct W2 only; no sponsorship.
- Benefits include comprehensive medical, 401(k), paid holidays, and more.
About itD
itD is a consulting and software development company focused on delivering real business results with diversity, innovation, and integrity. The firm is woman- and minority-led and emphasizes a low-hierarchy culture to achieve strong customer outcomes.
Scraped 8/1/2026