Site Reliability Engineer managing deployments and optimizations for CBC/Radio-Canada's CMS projects. Collaborating with teams to enhance infrastructure and software efficiencies.
Responsibilities
Manage and perform all deployment operations.
Improve deployment practices to achieve continuous delivery (CD) objectives.
Maintain and improve log reports for the Radio‑Canada CMS ecosystem.
Participate in code reviews with a focus on logging practices and server‑side code efficiency.
Identify issues and prevent waste of cloud resources.
Re-engineer software functionalities to optimize resource usage, in API code or databases.
Document processes and influence developers on best practices.
Present collected data to the relevant teams.
Requirements
3 to 5 years of backend development experience or a university degree
Proficiency with Git and CI/CD principles
Proficiency with Kibana
Advanced knowledge of monorepo concepts
Advanced knowledge of Azure Cloud
Advanced knowledge of Azure Functions
Advanced knowledge of Node.js (TypeScript)
Advanced knowledge of .NET Core
Advanced knowledge of MongoDB and Elasticsearch
Familiarity with the Microsoft Azure DevOps suite
Familiarity with Rancher
Familiarity with Kubernetes
Strong command of French; working knowledge of spoken and written English is an asset.
Benefits
Flexible work schedules, allowing you to prioritize yourself, your family, and your work.
Work-from-home opportunities.
Competitive total rewards package.
Opportunities to work with cutting-edge technology.
Opportunities for continued learning and professional development.
Opportunities to become a member of our Employee Resource Groups.
Pair programming and mentorship opportunities to learn from industry experts and help coach new talent.
A creative and dynamic work environment where your ideas and contributions are heard, valued, and respected.
A supportive management team committed to upholding high standards of diversity and inclusivity.
An environment that favors experimentation and iterative approaches to achieve technical innovation.
DevOps Intern supporting CI/CD, cloud infrastructure, and automation for Ludia’s mobile game studio. Improving reliability and developer tools in production game environments.
Manager leading global SRE teams for Akamai's distributed Cloud IAM services. Improving reliability, scalability, security, and usability through cloud - native tooling and software.
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.