Senior DevOps Engineer at HTG managing CI/CD pipelines and cloud infrastructure. Collaborating with a team to ensure reliable software releases and system availability.
Responsibilities
Build, optimize, and manage Continuous Integration and Continuous Deployment (CI/CD) pipelines.
Automate build, testing, and deployment processes.
Ensure faster, reliable, and repeatable software releases.
Troubleshoot pipeline failures and improve performance.
Design and manage infrastructure using code instead of manual configuration.
Automate provisioning of servers, networks, and environments.
Ensure consistency across development, staging, and production environments.
Implement version control for infrastructure changes.
Architect, deploy, and manage cloud-based systems.
Optimize scalability, availability, and cost efficiency.
Monitor cloud performance and resource utilization.
Implement backup, recovery, and disaster recovery strategies.
Implement monitoring and alerting systems for infrastructure and applications.
Ensure high system availability and performance.
Manage incident response and root cause analysis.
Maintain logging solutions for troubleshooting and auditing.
Integrate security practices into development and deployment pipelines.
Manage access control, secrets, and vulnerability scanning.
Ensure systems comply with organizational and regulatory standards.
Implement automated security checks.
Requirements
5-10 years of DevOps experience
Strong experience working in an AWS environment
Hands-on expertise with AWS Lambda (serverless applications)
Deep understanding of event-based messaging systems such as Solace, RabbitMQ, or Kafka
Proficient with containers and orchestration tools — ECS, Docker, Kubernetes
Expert-level knowledge of Terraform for infrastructure as code (IaC)
Skilled in scripting using Bash and/or Python
Experience with traceability and monitoring tools such as Dynatrace
Familiar with Service Mesh technologies including Kong or Istio
Strong background in building and maintaining CI/CD pipelines
Benefits
Equal Opportunity Employer with a commitment to inclusive teams
Senior DevOps / Cloud Infrastructure Engineer needed for hybrid role in North York, ON. Requires 10+ years experience with GCP, AWS, Kubernetes, Terraform, and CI/CD.
Staff Site Reliability Engineer strengthening AWS and Kubernetes resilience for Caseware, a fintech company building audit and accounting software. Driving secure delivery, observability, and incident management.
Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.
Mozilla release engineer optimizing Firefox build, test, and deployment pipelines at global scale. Improving developer experience, maintaining automation, and responding to critical service outages without an on - call rotation.