Manager of Site Reliability Engineering at Docebo, overseeing platform health and leading engineering teams. Focus on operational efficiency and incident management for reliable SaaS delivery.
Responsibilities
Lead a team of Site Reliability Engineers to ensure operational health of the Docebo platform
Manage incident responses and day-to-day production operations
Conduct post-incident reviews and implement corrective actions
Collaborate with Product, Engineering, and Support teams for reliable practices
Requirements
6 to 10 years of experience in SRE, DevOps, or systems operations
Managing or coordinating critical incident responses and production operations in SaaS environments
Familiarity with AWS and advanced monitoring/observability tools
Collaboration with cross-functional teams to deliver results
Benefits
Employee Share Purchase Plan (ESPP) at a 15% discount
Manager leading global SRE teams for Akamai's distributed Cloud IAM services. Improving reliability, scalability, security, and usability through cloud - native tooling and software.
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.
Team Leader guiding Remote’s global SRE platform for compliant international employment. Leading engineers and reliability across Kubernetes, AWS, observability, and infrastructure.