Platform DevOps managing the Enterprise Data and AI Platform across AWS and Kubernetes. Implementing Infrastructure as Code with Terraform and maintaining CI/CD pipelines for secure solutions.
Responsibilities
Manage and maintain the Enterprise Data and AI Platform across AWS and on-prem Kubernetes OKE infrastructure
Implement Infrastructure as Code with Terraform
Build and maintain CI/CD pipelines using Gitlab, GitHub Actions, and Jenkins
Collaborate with cross-functional teams to design and deliver scalable, secure solutions
Provide operational support and issue resolution
Enable frequent, reliable deployments in partnership with software developers
Promote DevOps best practices across teams
Implement observability and define KPIs for our platforms and services
Embed security best practices throughout the development lifecycle
Contribute to proof-of-concept projects
Implement high-availability and resilience, including disaster recovery and backups
Follow Agile/Scrum processes to deliver high-quality solutions on time
Requirements
Bachelor’s degree in computer science, computer engineering, or equivalent experience
5–8 years of experience, including DevOps in AWS cloud environments
5+ years of AWS experience with strong knowledge of S3, Lambda, CloudWatch, VPC, and AWS Backup
Hands-on expertise experience with Terraform, GitHub, Gitlab GitHub Actions, and Jenkins
Proficiency in Python
Strong networking skills including VPCs, load balancers, and firewalls
Experience with Kubernetes, OpenShift, and Docker is mandatory
Passion for technology and continuous learning, particularly in GenAI and AI Agentic
Customer-oriented, with a focus on delivering high-value solutions
Creative and innovative, with strong problem-solving skills
Adaptable and flexible, thriving in a dynamic, fast-paced environment
Detail-oriented, ensuring high standards of quality and precision in your work
A self-starter, who can drive and follow through the initiatives from the beginning to the end
Excellent communication skills, spoken and written
For candidates in Quebec: bilingualism required to collaborate with English-speaking colleagues across Canada
No work experience in Canada required, but you must have authorization to work in Canada
Benefits
Flexible work arrangements and a hybrid work model
Possibility to purchase up to 5 extra days off per year
Multiple benefits offered to support physical and mental wellbeing, including telemedicine, Wellness account and much more
Share plan & other savings: up to 12% of salary or even more (ask how you could earn guaranteed income for life)
Senior DevOps Engineer improving Sherweb’s cloud - based IT solution delivery through CI/CD, IaC, GitOps, and AI automation. Supporting secure platforms, operational transitions, and developer self - service.
Senior DevOps / Cloud Infrastructure Engineer needed for hybrid role in North York, ON. Requires 10+ years experience with GCP, AWS, Kubernetes, Terraform, and CI/CD.
Staff Site Reliability Engineer strengthening AWS and Kubernetes resilience for Caseware, a fintech company building audit and accounting software. Driving secure delivery, observability, and incident management.
Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.
Mozilla release engineer optimizing Firefox build, test, and deployment pipelines at global scale. Improving developer experience, maintaining automation, and responding to critical service outages without an on - call rotation.