xelys jobs xelys jobs

DevOps Engineer

Careflow

full-remoteseniorpermanentdevops New York, NY 46 days ago via LinkedIn

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

Google Cloud Platform (GCP)DevOpsSite Reliability EngineeringCI/CDMonitoringLoggingObservabilityCloud SecurityIncident ResponseDisaster Recovery

About the role

Role Overview

Careflow is seeking an experienced DevOps Engineer to own and improve its cloud infrastructure, security, observability, and operational reliability. You’ll build and maintain production systems end-to-end, automate deployments, monitor services, and troubleshoot issues across the stack (including the application code when needed).

This is a fully remote role with flexible scheduling and Saturday weekend coverage plus a weekday day off in exchange.

Responsibilities

Cloud Infrastructure & Operations

  • Manage and improve the Google Cloud Platform (GCP) environment
  • Design infrastructure for scalability, reliability, and cost efficiency
  • Own networking, compute, databases, storage, and other cloud services
  • Monitor health and proactively address performance bottlenecks

Monitoring, Logging & Observability

  • Build and maintain centralized logging and monitoring
  • Create dashboards and alerts for system health, application performance, and critical workflows
  • Define operational metrics and track usage across the platform
  • Lead incident response and root cause analysis

Security & Compliance

  • Implement and maintain security best practices for infrastructure and applications
  • Manage identity and access controls, secrets management, and environment security
  • Perform security reviews and remediate vulnerabilities
  • Support compliance initiatives and audit readiness

CI/CD & Automation

  • Improve deployment pipelines and release processes
  • Automate infrastructure provisioning and operational workflows
  • Enhance development environments and deployment reliability
  • Reduce manual operational work through automation

Reliability Engineering

  • Improve uptime, resiliency, backups, and disaster recovery
  • Define service-level objectives (SLOs) and operational standards
  • Drive improvements in platform stability and performance

Cross-Functional Support

  • Partner with engineering, product, and leadership teams
  • Provide technical guidance on infrastructure/operations
  • Participate in on-call and operational support rotation

Bonus Responsibilities

  • Troubleshoot and fix application-level issues
  • Contribute code improvements and bug fixes
  • Assist with performance optimization and debugging

Required Qualifications

  • 5+ years experience in DevOps, SRE, Cloud Engineering, or related work
  • Strong hands-on experience with Google Cloud Platform (GCP)
  • Experience building and maintaining CI/CD pipelines
  • Solid understanding of monitoring, logging, and alerting systems
  • Experience with cloud security best practices

Position Details

  • Title: DevOps Engineer
  • Employment Type: Full-Time
  • Location: Fully Remote
  • Schedule: Flexible; Saturday coverage with a weekday day off
  • Reports To: Lead Architect

About Careflow

Careflow is building and operating a cloud-based software platform that serves its users reliably at scale. The role focuses on DevOps, site reliability, security, and observability to support growth and operational excellence.

Scraped 6/18/2026