Senior SRE Engineer with software development background to automate cloud infrastructure using Azure, Terraform, Kubernetes. Hybrid role in Mississauga, ON.
Responsibilities
Design and manage Kubernetes clusters and containerized workloads in Azure (AKS). Automate infrastructure using Terraform, Ansible, ARM, Helm, Docker. Build and optimize CI/CD pipelines, including PowerShell/Shell scripting. Monitor production environments using Grafana, Kibana, Prometheus, ensuring uptime and performance. Collaborate with Dev, Systems, and Networking teams to simplify and automate operational processes. Maintain technical documentation, DR plans, and dashboards for production visibility.
Requirements
6+ years as an SRE supporting production infrastructure. 6+ years in software development (automation QA, SDET, app support, or build/release experience). 2+ years managing Kubernetes production systems. Deep experience with Azure cloud, AKS, Terraform, Docker, CI/CD pipelines, microservices. Strong analytical, problem-solving, and communication skills. Ability to read and understand code (loops, conditionals, classes, object structures).
Benefits
Hands-on work with enterprise-grade cloud infrastructure. Opportunity to lead automation and DevOps best practices. Hybrid flexibility and exposure to senior leadership.
DevOps Engineer building AWS platforms, CI/CD pipelines, and infrastructure automation for SWTCH’s North American EV charging solutions. Improving deployment reliability, observability, and cloud architecture.
Reliability Engineering Manager leading SAP PM and asset - reliability strategy. Improving equipment performance across Apotex’s global pharmaceutical manufacturing sites.
DevOps Intern supporting CI/CD, cloud infrastructure, and automation for Ludia’s mobile game studio. Improving reliability and developer tools in production game environments.
Manager leading global SRE teams for Akamai's distributed Cloud IAM services. Improving reliability, scalability, security, and usability through cloud - native tooling and software.
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.