Site Reliability Engineer (SRE)
Rocket.net
midpermanentdevops Jupiter, FL 2 days ago via LinkedIn
See how well this job matches your profile
Sign up to get an AI match score and generate a tailored application in seconds.
Get your match scoreTags
Site Reliability EngineeringSRELinuxNGINXApachePHP-FPMMySQL/MariaDBRedisDNSIncident Response
About the role
Role Overview
Rocket.net is hiring a Site Reliability Engineer (SRE) for the Platform Operations team. You’ll help maintain the health, stability, and performance of Rocket.net’s hosting platform while providing advanced, customer-facing technical support as an escalation layer between WordPress Support and Engineering.
Responsibilities
Platform Monitoring & Reliability
- Monitor health, availability, and performance of servers, services, and customer environments
- Proactively identify infrastructure issues, performance degradation, and potential service disruptions
- Investigate alerts and operational events to maintain platform stability
- Run regular platform health checks to ensure critical systems are operating correctly
- Participate in incident response and coordinate troubleshooting during customer-impacting events
- Communicate platform issues, updates, and resolutions to internal teams
Advanced Technical Support & Escalations
- Provide advanced technical support for VIP customers and complex hosting issues
- Act as a senior escalation point for WordPress Support Engineers
- Troubleshoot complex problems across servers, websites, networking, DNS, performance, caching, and hosting infrastructure
- Investigate and resolve issues involving server resources, application performance, connectivity, and platform behavior
- Work directly with customers when expert-level technical assistance is required
Infrastructure Operations
- Troubleshoot and maintain Linux-based production environments
- Investigate issues involving NGINX, Apache, PHP-FPM, MySQL/MariaDB, Redis, and other platform services
- Support server maintenance, configuration changes, and operational improvements
- Contribute to security updates, system hardening, and infrastructure best practices
- Monitor resource usage and identify capacity/performance concerns
- Improve monitoring, automation, and operational workflows
Team Collaboration
- Collaborate with WordPress Support Engineers, Shift Leads, SREs, and Engineering teams
- Provide technical guidance and knowledge sharing to support teams
- Create internal documentation, troubleshooting guides, and knowledge base articles
- Identify recurring issues and recommend improvements to reduce future incidents
- Participate in incident reviews and conduct root-cause analysis (RCA)
Requirements
- 3+ years of experience in SRE, DevOps, Platform Engineering, or similar roles
- Strong troubleshooting experience with Linux production environments
- Experience supporting customer-facing technical environments
- Strong understanding of web hosting technologies: NGINX, Apache, PHP-FPM, MySQL/MariaDB, Redis
- Advanced troubleshooting across WordPress, servers, DNS, networking, and performance issues
- Proficiency with Linux command line (SSH)
- Understanding of DNS, HTTP/HTTPS, SSL/TLS, CDN, and caching
- Experience with Cloudflare/WAF and web performance optimization
- Familiarity with monitoring tools and incident response processes
- Ability to troubleshoot independently and communicate technical solutions clearly
Location
- Jupiter, FL
About Rocket.net
Rocket.net is a managed WordPress hosting platform focused on reliability, performance, and customer experience. Its Platform Operations function bridges customer support and engineering to keep servers and hosted customer environments stable and highly available.
Scraped 7/31/2026