xelys jobs xelys jobs

DevOps Engineer

Careflow

full-remoteseniorpermanentdevops Georgetown, SC 48 days ago via LinkedIn

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

Google Cloud Platform (GCP)DevOpsSite Reliability Engineering (SRE)CI/CDMonitoringLoggingObservabilityIncident ResponseSecurityInfrastructure Automation

About the role

Role Overview

Careflow is hiring an experienced DevOps Engineer to own and improve cloud infrastructure, security, observability, and operational reliability. You’ll build and automate infrastructure, enhance deployment processes, monitor production systems, and troubleshoot across the stack when needed. The position is fully remote, with flexible scheduling that includes Saturday coverage and a weekday day off in return.

Responsibilities

  • Cloud Infrastructure & Operations
    • Manage and improve Google Cloud Platform (GCP) environments
    • Design infrastructure for scalability, reliability, and cost efficiency
    • Own networking, compute, databases, storage, and cloud services
    • Monitor health and proactively address performance bottlenecks
  • Monitoring, Logging & Observability
    • Implement centralized logging and monitoring
    • Create dashboards and alerts for system health and key workflows
    • Define operational metrics and usage tracking
    • Lead incident response and root cause analysis
    • Improve monitoring and alerting over time
  • Security & Compliance
    • Apply security best practices to infrastructure and applications
    • Manage identity/access controls and secrets management
    • Support security reviews and vulnerability remediation
    • Assist with compliance initiatives and audit readiness
  • CI/CD & Automation
    • Improve deployment pipelines and release processes
    • Automate infrastructure provisioning and operational workflows
    • Enhance deployment reliability and reduce manual operational work
  • Reliability Engineering
    • Improve uptime, resiliency, backups, and disaster recovery
    • Establish SLOs and operational standards
    • Drive platform stability and performance improvements
  • Cross-Functional Support
    • Partner with engineering, product, and leadership teams
    • Provide technical guidance on infrastructure/operations
    • Participate in on-call and operational support rotation
  • Bonus / Hands-on Support
    • Troubleshoot and fix application-level issues when needed
    • Contribute bug fixes and code improvements
    • Support performance optimization and debugging

Requirements

  • 5+ years in DevOps, Site Reliability Engineering, Cloud Engineering, or related roles
  • Strong hands-on experience with Google Cloud Platform (GCP)
  • Experience building and maintaining CI/CD pipelines
  • Strong understanding of monitoring, logging, and alerting

Success in the First 90 Days

  • Take ownership of GCP infrastructure and environments
  • Establish visibility into performance, reliability, and usage metrics
  • Improve monitoring/alerting and incident response processes
  • Identify and address security and operational risks
  • Reduce infrastructure-related issues and deployment friction
  • Become a trusted resource for platform reliability and operational excellence

About Careflow

Careflow is a software company building and operating a cloud-hosted platform. The role focuses on maintaining a secure, scalable, and reliable infrastructure across Google Cloud Platform with strong monitoring, CI/CD automation, and operational excellence.

Scraped 6/17/2026