DevOps Engineer
Evlo AI
midpermanentdevopsbackend Washington, DC 2 days ago via LinkedIn
See how well this job matches your profile
Sign up to get an AI match score and generate a tailored application in seconds.
Get your match scoreTags
TerraformKubernetesAWSGCPEKSGKECI/CDGitHub ActionsPrometheusGrafanaOpenTelemetryDatadogLinuxIAMNetworkingArgoCDGitOpsOn-callObservability
About the role
DevOps Engineer (Cloud Platform Reliability)
You’ll be responsible for the availability, latency, performance, efficiency, and capacity management of a high-throughput cloud-native platform.
Responsibilities
- Design, provision, and maintain multi-region cloud infrastructure on AWS or GCP using Terraform and Infrastructure-as-Code (IaC) best practices
- Manage and scale production Kubernetes clusters (EKS/GKE), including networking, ingress controllers, and IAM integrations
- Build and optimize CI/CD pipelines using GitHub Actions, GitLab CI, or ArgoCD to automate build, test, and deployment
- Implement observability frameworks and monitoring using Prometheus, Grafana, OpenTelemetry, and ELK/Datadog
- Participate in a blameless on-call rotation, run post-mortems, and implement preventative automation to reduce operational toil
- Partner with security to enforce IAM roles, network security policies, and vulnerability scanning across build and runtime environments
Requirements
- 3–6 years of experience in DevOps, SRE, or Infrastructure Engineering supporting high-traffic production systems
- Strong proficiency with IaC, specifically Terraform, and container orchestration with Kubernetes
- Solid software engineering skills in at least one language: Python, Go, or Bash
- Deep understanding of Linux systems administration, networking fundamentals (TCP/IP, DNS, VPCs), and cloud security best practices
- Experience configuring CI/CD and modern observability pipelines (Prometheus, Grafana, Datadog)
Nice to Have
- Service meshes (Istio/Linkerd)
- GitOps workflows (ArgoCD/Flux)
- Managing distributed databases at scale
Scraped 7/26/2026