Observability / DevOps Advisor role overseeing reliability and performance of applications. Support teams by implementing observability platforms, focusing on CI/CD pipelines and AI.
Responsibilities
Support teams during production incidents and help prevent future incidents by leveraging our observability platforms and industry best practices;
Onboard new teams, applications, services, and infrastructure components onto our observability and SRE platforms, with a strong focus on CI/CD pipelines and the use of artificial intelligence;
Enhance our observability platforms, tools, pipelines, practices, documentation, training materials, and our use of artificial intelligence.
Requirements
At least 5 years of experience in a developer, DevOps, and/or observability role
Strong experience and interest in support and operations, ideally in a large enterprise environment with multiple cross-functional teams
Solid experience designing CI/CD pipelines with Terraform, including designing and maintaining modules that facilitate pipeline adoption and reuse
Experience using artificial intelligence in an enterprise or software development context (ML, GenAI, and AI agents)
Bilingual (French and English): Need to interact on a regular basis with an English-speaking clientele and colleagues across the country
No Canadian work experience required however must be eligible to work in Canada
Strong assets: Experience with Dynatrace or similar application and service observability (APM) solutions; Experience with Elasticsearch or similar log management solutions; Experience producing and working with telemetry signals (traces, logs, metrics, events, etc.); Experience with IT service management (ITSM), configuration management (CMDB), incident management, and notification platforms; Experience defining KPIs and service level objectives (SLA, SLO, SLI); Experience with AWS, Azure, GCP, Kubernetes, and OpenShift; Experience with networking, routing/switching, and cybersecurity infrastructure; Experience in Java, Python, and SQL; Experience with diagramming, dashboards, and reporting; Experience delivering team training and knowledge transfer; Experience in cybersecurity and vulnerability detection.
Benefits
Flexible work arrangements and a hybrid work model
Possibility to purchase up to 5 extra days off per year
Multiple benefits offered to support physical and mental wellbeing, including telemedicine, Wellness account and much more
Share plan & other savings: up to 12% of salary or even more (ask how you could earn guaranteed income for life)
Senior DevOps Engineer improving Sherweb’s cloud - based IT solution delivery through CI/CD, IaC, GitOps, and AI automation. Supporting secure platforms, operational transitions, and developer self - service.
Senior DevOps / Cloud Infrastructure Engineer needed for hybrid role in North York, ON. Requires 10+ years experience with GCP, AWS, Kubernetes, Terraform, and CI/CD.
Staff Site Reliability Engineer strengthening AWS and Kubernetes resilience for Caseware, a fintech company building audit and accounting software. Driving secure delivery, observability, and incident management.
Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.
Mozilla release engineer optimizing Firefox build, test, and deployment pipelines at global scale. Improving developer experience, maintaining automation, and responding to critical service outages without an on - call rotation.