Senior DevOps Engineer operating AWS and Kubernetes infrastructure for Modern Campus’s higher-education platform. Automating deployments, improving reliability, and supporting customer-facing products.
Responsibilities
Design, build, and operate AWS infrastructure, including production workloads on Amazon EKS and ECS/Fargate
Own infrastructure projects through design, testing, deployment, rollback planning, documentation, and operational handoff
Develop infrastructure as code, deployment automation, and reusable platform capabilities using Terraform, Ansible, Helm, and Argo CD
Build and maintain deployment pipelines and automation, partnering with Development on application modernization, scaling, validation, production readiness, and rollback procedures
Troubleshoot infrastructure and application issues across operating systems, networking, containers, application runtimes, and database dependencies
Improve monitoring, performance, and cost efficiency; implement security controls in partnership with Information Security; maintain platform upgrades and patching; test backup restoration and disaster recovery procedures
Contribute to technical standards and design reviews, mentor engineers, and work with management to prioritize risks and improvements
Participate in the on-call rotation, lead incident diagnosis, and carry out planned maintenance
Communicate clearly during incidents and use follow-up work, automation, and runbooks to reduce recurring issues
Requirements
Substantial infrastructure and DevOps experience, typically eight or more years in infrastructure and five or more years in automation-focused roles, or equivalent demonstrated expertise
Strong hands-on production Kubernetes experience, including upgrades, networking, access controls, scaling, and troubleshooting
Amazon EKS experience is strongly preferred
Strong AWS experience across networking, IAM, compute, storage, load balancing, and managed databases
Practical experience with Terraform, containers, Helm, and Git-based deployment workflows, including building and maintaining CI/CD pipelines using GitHub Actions, GitLab CI/CD, or comparable tools
Strong Linux administration and troubleshooting skills, including DNS, HTTP/TLS, and system performance
Strong practical networking and troubleshooting skills, including routing, firewalls, and secure connectivity such as site-to-site VPNs or SSH tunnels
Ability to write maintainable automation in Python, Go, PowerShell, or another suitable language, alongside shell scripting
Experience diagnosing application and database issues using metrics, logs, traces, and monitoring tools
Demonstrated ability to independently deliver complex work, make safe production changes, and communicate technical decisions across teams
A degree in Computer Science or a related field, or equivalent practical experience
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.
Team Leader guiding Remote’s global SRE platform for compliant international employment. Leading engineers and reliability across Kubernetes, AWS, observability, and infrastructure.
AWS and DevOps Engineer establishing secure, automated environments for a bilingual nonprofit digital platform. Managing deployment, monitoring, recovery, and operational handover.
Senior Site Reliability Engineer securing AuthZed’s cloud infrastructure and authorization platform, including SpiceDB. Building Kubernetes guardrails, supply - chain security, vulnerability management, and incident response.