Manager, Software Engineering
About Us
Sophos is a cybersecurity leader defending 600,000 organizations globally with an AI-driven platform and expert-led services. Sophos meets organizations wherever they are in their security maturity and grows with them to defeat cyberattacks. Its solutions combine machine learning, automation, and real-time threat intelligence with frontline human expertise from Sophos X-Ops to deliver advanced, 24/7 threat monitoring, detection, and response.
Sophos offers industry-leading managed detection and response (MDR) alongside a comprehensive portfolio of cybersecurity technologies — including endpoint, network, email, and cloud security, extended detection and response (XDR), identity threat detection and response (ITDR), and next-gen SIEM. Together with expert advisory services, these capabilities help organizations proactively reduce risk and respond faster, with the visibility and scalability needed to stay ahead of evolving threats.
Sophos goes to market with a global partner ecosystem, including Managed Service Providers (MSPs), Managed Security Service Providers (MSSPs), resellers and distributors, marketplace integrations, and cyber risk partners, giving organizations the flexibility to choose trusted relationships when securing their business. Sophos is headquartered in Oxford, U.K. More information is available at .
Role Summary
We are seeking an experienced Manager, Software Engineering (SRE) to lead a team focused on improving the reliability, scalability, and operational maturity of Sophos cloud services and engineering delivery systems. In this role, you will lead a team distributed across the U.S. and Canada. The team is responsible for production reliability, operational readiness, incident response, escalation management, observability, automation, release reliability, and continuous improvement across cloud-based platforms.
This role requires a leader with a strong background spanning site reliability engineering, platform engineering, DevOps practices, release engineering, infrastructure management, and cloud operations. You should have enough technical depth to guide discussions around AWS, Kubernetes/EKS, CI/CD, Linux, infrastructure as code, observability, cost optimization, security practices, compliance needs, and automation, while primarily focusing on team leadership, execution, stakeholder alignment, prioritization, and improving how engineering teams deliver and operate services at scale.
What You Will Do
Leadership and Team Management
- Lead, coach, and develop a team of engineers focused on site reliability, platform operations, release engineering, and automation. Set clear priorities, expectations, and delivery goals aligned to business and engineering needs.
- Support hiring, onboarding, performance management, and career development for team members.
- Foster a culture of ownership, collaboration, continuous improvement, and operational excellence.
SRE Strategy and Operational Maturity
- Partner with senior leadership to define and execute the roadmap for SRE and operational improvements.
- Drive improvements in reliability, availability, scalability, deployment confidence, and production readiness.
- Drive operational efficiency and cost optimization practices across cloud platforms, balancing reliability, performance, and responsible cloud spend.
- Help establish best practices for incident response, escalation management, observability, runbooks, post-incident reviews, and toil reduction.
- Promote the use of reliability metrics, operational health indicators, and continuous improvement practices.
Platform, Release, and Automation Enablement
- Partner with engineering, architecture, security, release, and operations teams to improve shared tooling and delivery workflows.
- Support improvements across CI/CD pipelines, release automation, deployment reliability, and environment stability. Guide team efforts involving cloud infrastructure, Kubernetes/EKS, Linux sys
Get remote developer jobs like this by email
One weekly digest. No spam, unsubscribe anytime.