Observability / DevOps Advisor role overseeing reliability and performance of applications. Support teams by implementing observability platforms, focusing on CI/CD pipelines and AI.
Responsibilities
Support teams during production incidents and help prevent future incidents by leveraging our observability platforms and industry best practices;
Onboard new teams, applications, services, and infrastructure components onto our observability and SRE platforms, with a strong focus on CI/CD pipelines and the use of artificial intelligence;
Enhance our observability platforms, tools, pipelines, practices, documentation, training materials, and our use of artificial intelligence.
Requirements
At least 5 years of experience in a developer, DevOps, and/or observability role
Strong experience and interest in support and operations, ideally in a large enterprise environment with multiple cross-functional teams
Solid experience designing CI/CD pipelines with Terraform, including designing and maintaining modules that facilitate pipeline adoption and reuse
Experience using artificial intelligence in an enterprise or software development context (ML, GenAI, and AI agents)
Bilingual (French and English): Need to interact on a regular basis with an English-speaking clientele and colleagues across the country
No Canadian work experience required however must be eligible to work in Canada
Strong assets: Experience with Dynatrace or similar application and service observability (APM) solutions; Experience with Elasticsearch or similar log management solutions; Experience producing and working with telemetry signals (traces, logs, metrics, events, etc.); Experience with IT service management (ITSM), configuration management (CMDB), incident management, and notification platforms; Experience defining KPIs and service level objectives (SLA, SLO, SLI); Experience with AWS, Azure, GCP, Kubernetes, and OpenShift; Experience with networking, routing/switching, and cybersecurity infrastructure; Experience in Java, Python, and SQL; Experience with diagramming, dashboards, and reporting; Experience delivering team training and knowledge transfer; Experience in cybersecurity and vulnerability detection.
Benefits
Flexible work arrangements and a hybrid work model
Possibility to purchase up to 5 extra days off per year
Multiple benefits offered to support physical and mental wellbeing, including telemedicine, Wellness account and much more
Share plan & other savings: up to 12% of salary or even more (ask how you could earn guaranteed income for life)
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.
Team Leader guiding Remote’s global SRE platform for compliant international employment. Leading engineers and reliability across Kubernetes, AWS, observability, and infrastructure.
AWS and DevOps Engineer establishing secure, automated environments for a bilingual nonprofit digital platform. Managing deployment, monitoring, recovery, and operational handover.