Site Reliability Engineer (FedRAMP / Security)
Coralogix
full-remoteseniorpermanentbackend Full remote - New York, US Today via WTTJ
See how well this job matches your profile
Sign up to get an AI match score and generate a tailored application in seconds.
Get your match scoreTags
KubernetesAWSFedRAMPSREGrafanaPrometheusTerraformInfrastructure as CodeIncident ManagementNetwork Protocols
About the role
Role overview
You will be a Site Reliability Engineer on Coralogix’s Cloud Infrastructure Team, focused on Enterprise FedRAMP Cloud Infrastructure. The role combines production operations, reliability engineering, and roadmap ownership for FedRAMP cloud products.
Key missions & responsibilities
- Partner with R&D to improve system stability and reliability.
- Operate FedRAMP cloud products, including:
- Deployments
- On-call support
- Incident management
- Build internal tools to expand platform capabilities and own improvements end-to-end.
- Adopt cutting-edge technologies to enhance reliability and operational excellence.
- Lead the product roadmap for the FedRAMP cloud offering.
Requirements
- In-depth Kubernetes experience with strong operating and monitoring skills.
- FedRAMP compliance experience (High/Moderate), including:
- Vulnerability management
- Continuous monitoring (e.g., scanning, patching, reporting)
- Familiarity with monitoring tools such as Grafana and Prometheus (Coralogix monitoring is a plus).
- Networking understanding across networking layers and protocols (e.g., HTTP, gRPC, SSL/TLS).
- Experience with AWS (or other cloud providers).
- Experience with Infrastructure as Code (e.g., Terraform, Crossplane).
- 5+ years as a DevOps Engineer / SRE in production environments.
Nice to have
- Familiarity with Apache Kafka.
- Experience operating data pipelines.
- Software engineering experience, preferably in Golang.
About Coralogix
Coralogix is a fast-growing technology company recognized as a unicorn in its industry. It builds cloud infrastructure and related platforms used in high-scale environments, collaborating closely with R&D to improve reliability and system stability.
Scraped 8/5/2026