Staff DevOps Engineer at Agiloft leading DevOps initiatives and designing scalable systems for CloudOps. Collaborating with cross-functional teams to drive departmental goals and architectural standards.
Responsibilities
help drive the overall architecture for the Agiloft CloudOps systems
design systems that are scalable and resilient utilizing DevOps and security best practices
collaborate cross-functional and lead DevOps initiatives
provide expertise to drive the departmental goals
assume a team lead role for various projects and initiatives
advise on complex technical issues
Requirements
Bachelor's degree in Computer Science, Information Technology, or related field (or equivalent experience).
7+ years of related experience
Expertise in DevOps principles, practices, and technologies including Amazon Web Services (AWS) and Terraform or other Infrastructure as Code (IaC)
Advanced knowledge of Linux operating systems and troubleshooting OS issues
Highly skilled in setting up and managing monitoring tools (such as Prometheus, Grafana, Datadog, Nagios, Open Telemetry, ELK, or similar tools)
Advanced knowledge of scripting languages and automation utilizing Python, Bash or Ruby
Proficiency in using relevant AI Tools in the SLDC (for example, GitHub Copilot, JetBrains AI Assistant)
Deep understanding of:
Networking concepts and principles
Version Control Systems (such as Git)
CI/CD tools such as Jenkins, Gitlab CI/CD, Github, or similar tool
Containerization and orchestration (Docker, Kubernetes)
Expertise with cloud platforms (AWS, Azure, or Google Cloud).
Superior problem-solving, troubleshooting/debugging skills, and communication skills.
Proven experience in security best practices including identity and access management, encryption, and vulnerability assessments
Eagerness to learn and adapt to new technologies and tools.
Participate in on-call rotation schedule with the rest of the team
Experience with adopting AI to supplement code development (Cursor, Copilot)
Experience with managing and optimizing vector databases for AI/ML (e.g., PostgreSQL vectordb)
DevOps Intern supporting CI/CD, cloud infrastructure, and automation for Ludia’s mobile game studio. Improving reliability and developer tools in production game environments.
Manager leading global SRE teams for Akamai's distributed Cloud IAM services. Improving reliability, scalability, security, and usability through cloud - native tooling and software.
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.