Site Reliability Engineer automating deployment and operation of software for dental software provider. Collaborating with developers and mentoring peers to streamline processes and improve efficiencies.
Responsibilities
Stewarding Infrastructure as Code (IaC)
Mentoring peers
Shaping the technical direction of the platform
Collaborate closely with developers to support a wide range of applications
Automate repetitive tasks to improve deployment and operation of software
Responding quickly to incidents while prioritizing long-term solutions
Helping create a positive culture that embraces learning and curiosity
Requirements
Bachelor’s degree in Computer Science or equivalent experience
Deep experience with AWS in production environments, including EC2, S3, IAM, EKS, and RDS
Hands-on experience managing Kubernetes clusters and deploying applications
Strong Linux expertise and system-level troubleshooting skills
Solid understanding of system administration, security best practices, and managing mission-critical data
Proven experience monitoring and optimizing large-scale enterprise web applications
Familiarity with infrastructure automation tools such as Ansible, Packer, and AWS CloudFormation
Proficient in at least one programming language and scripting for automation
Strong analytical skills and experience with root cause analysis
Comfortable working in agile environments, including participation in code reviews and testing
DevOps Intern supporting CI/CD, cloud infrastructure, and automation for Ludia’s mobile game studio. Improving reliability and developer tools in production game environments.
Manager leading global SRE teams for Akamai's distributed Cloud IAM services. Improving reliability, scalability, security, and usability through cloud - native tooling and software.
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.