DevOps Engineer leading strategic cloud infrastructure initiatives at SpryPoint. Focused on AWS migrations and infrastructure as code for a multi-tenant SaaS platform.
Responsibilities
Lead infrastructure migrations from Elastic Beanstalk to ECS/Fargate and RDS to Aurora Serverless v2 with zero downtime
Design and implement AWS CDK (Python) patterns for infrastructure as code, establishing team standards and best practices
Optimize PostgreSQL and Aurora databases for performance through query tuning, connection pool management, and capacity planning
Architect and standardize network infrastructure using VPCs, Transit Gateway, security groups, and routing aligned with AWS Well-Architected Framework
Build automation workflows using Python, AWS Lambda, and Step Functions to improve operational efficiency
Implement observability and monitoring systems using AWS native tools (CloudWatch, Application Signals, X-Ray, etc.) to proactively identify and resolve issues
Create comprehensive runbooks and documentation that enable team self-service and reduce dependencies
Mentor DevOps engineers on infrastructure fundamentals, AWS best practices, and effective use of AI-assisted development tools
Collaborate with engineering and security teams to maintain SOC2/PCI compliance while enabling rapid delivery
Drive operational excellence through systematic troubleshooting, incident response, and continuous improvement
Requirements
5+ years managing production AWS infrastructure at scale, including ECS, Fargate, and container orchestration
Proven track record leading cloud migrations or infrastructure modernization projects from planning through production
Strong experience with PostgreSQL or Aurora including performance tuning, query optimization, and connection management
Proficient in Python for infrastructure automation, CLI tools, and operational scripting
Experience with infrastructure as code using CloudFormation, Terraform, or AWS CDK
Deep understanding of AWS networking including VPCs, Transit Gateway, security groups, NACLs, and routing
Solid Linux system administration skills and systematic troubleshooting methodology
Experience designing and maintaining CI/CD pipelines with Jenkins, GitHub Actions, or similar tools
Strong written and verbal communication skills with ability to create clear technical documentation
Familiarity with AWS Well-Architected Framework principles and their application in production environments
Benefits
Comprehensive compensation package that grows with you
Health, dental, vision, and life insurance from day one
Generous PTO, Summer Friday half-days, and unlimited sick days
RRSP (Canada) and 401k (US) matching programs
$2,500 annual development fund, tuition assistance, and Book Bounty program
Annual company events and team offsites that bring us together
Reliability Engineering Co - Op testing optical switching products and components at Lumentum. Analyzing failures, reliability data, and regulatory qualification results in Ottawa.
Senior DevOps Engineer operating secure AWS infrastructure and CI/CD for SOVRA’s public procurement platform. Improving reliability, observability, compliance, and AI - enabled operations.
Forward Deployed Engineer deploying CruxOCM’s heavy - industry automation software in complex operational technology environments. Integrating customer systems, commissioning solutions, and translating field learnings into product improvements.
Site Reliability Engineer improving AXON Networks’ AI - driven ISP orchestration platform and high - speed router services. Enhancing cloud - to - device reliability, observability, automation and incident response.
We’re looking for an Azure & Databricks DevOps Engineer (12 - month renewable contract) to support a large - scale Azure data platform initiative. 📍 Hybrid role in
Senior DevOps Engineer building secure, scalable Azure platforms for CARET’s legal and accounting practice - management software. Leading infrastructure, Kubernetes, CI/CD, security, observability, and reliability initiatives.