Site Reliability Engineer managing AWS accounts and cloud infrastructure deployment. Collaborating with teams to ensure security and efficiency of cloud operations at Ping Identity.
Responsibilities
Manage multiple AWS accounts with tools like AWS Control Tower and Terraform
You will deploy and debug cloud stacks, educating teams on new cloud projects, and ensuring the security of the cloud infrastructure
As a SRE, you can identify the most optimal cloud-based solutions for our internal users, and maintain cloud infrastructures following industry leading practices and company security policies
This is a 24/7 on-call position with a rotation schedule
Requirements
3+ years of experience with Linux/UNIX systems administration
3+ years of experience with Amazon Web Services (AWS)
2+ years experience in scripting skills in Python / Ruby / Bash / Go
Experience provisioning public cloud resources using frameworks such as CloudFormation and Terraform
Solid experience with server configuration with Ansible / Puppet / Chef / Salt
Experience using Git in a team environment (merge requests, branching, push, and pulls)
Senior Site Reliability Engineer deploying Kubernetes - based AI infrastructure on NVIDIA - certified hardware for Mirantis. Ensuring reliable, secure, scalable cloud operations and customer delivery.
DevOps Engineer building and maintaining cloud infrastructure, automation, and CI/CD pipelines for Calliere's software platform. Operating containers, observability tooling, and production workloads across public clouds.
Senior DevOps Engineer improving Sherweb’s cloud - based IT solution delivery through CI/CD, IaC, GitOps, and AI automation. Supporting secure platforms, operational transitions, and developer self - service.
Senior DevOps / Cloud Infrastructure Engineer needed for hybrid role in North York, ON. Requires 10+ years experience with GCP, AWS, Kubernetes, Terraform, and CI/CD.
Staff Site Reliability Engineer strengthening AWS and Kubernetes resilience for Caseware, a fintech company building audit and accounting software. Driving secure delivery, observability, and incident management.
Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.