xelys jobs xelys jobs

DevOps Engineer

Careflow

full-remoteseniorpermanentdevopsbackend New York, NY 47 days ago via LinkedIn

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

Google Cloud Platform (GCP)DevOpsSite Reliability Engineering (SRE)CI/CDMonitoring & ObservabilityIncident ResponseCloud SecurityAutomationReliability EngineeringRoot Cause Analysis

About the role

Role overview

Careflow is hiring an experienced DevOps Engineer to own and continuously improve cloud infrastructure, security, observability, and operational reliability. This is a fully remote role for someone who can “wear multiple hats”: building infrastructure, improving deployment processes, monitoring production, and troubleshooting across the stack (including diving into application code when needed).

Responsibilities

Cloud Infrastructure & Operations

  • Manage and maintain Google Cloud Platform (GCP) environments
  • Design and improve infrastructure for scalability, reliability, and cost efficiency
  • Oversee networking, compute, databases, storage, and related cloud services
  • Monitor system health and proactively address performance bottlenecks

Monitoring, Logging & Observability

  • Build and maintain centralized logging and monitoring
  • Create dashboards and alerts for system health and key workflows
  • Establish operational metrics and usage tracking
  • Lead incident response and root cause analysis
  • Monitor and manage platform spend

Security & Compliance

  • Implement and maintain security best practices across infrastructure and applications
  • Manage identity and access controls, secrets management, and environment security
  • Conduct security reviews and drive vulnerability remediation
  • Support compliance initiatives and audit readiness

CI/CD & Automation

  • Improve deployment pipelines and release processes
  • Automate infrastructure provisioning and operational workflows
  • Enhance development environments and deployment reliability
  • Reduce manual operational tasks through automation

Reliability Engineering

  • Improve uptime, resiliency, backups, and disaster recovery
  • Establish service-level objectives (SLOs) and operational standards
  • Drive stability and performance improvements

Cross-functional support

  • Partner with engineering, product, and leadership teams
  • Provide technical guidance on infrastructure and operational considerations
  • Participate in an on-call / operational support rotation

Bonus responsibilities

  • Troubleshoot and fix application-level issues when needed
  • Contribute bug fixes and code improvements
  • Help with performance optimization and debugging

Schedule / coverage

  • Flexible schedule, with availability for Saturday coverage and an additional weekday off in exchange.

Required qualifications

  • 5+ years in DevOps, SRE, Cloud Engineering, or related experience
  • Strong hands-on experience with Google Cloud Platform (GCP)
  • Experience building and maintaining CI/CD pipelines
  • Strong understanding of monitoring, logging, and alerting systems
  • Experience with cloud security best practices

Role details

  • Title: DevOps Engineer
  • Employment type: Full-time
  • Location: Fully Remote
  • Reports to: Lead Architect

About Careflow

Careflow is a software company building and operating a growing platform that requires reliable, secure, and scalable infrastructure. The role focuses on owning cloud infrastructure and production operations, including security, observability, and deployment automation on Google Cloud Platform.

Scraped 6/18/2026