Senior Site Reliability Engineer responsible for designing scalable systems at Euna Solutions. Collaborating with developers and mentoring juniors while driving automation and reliability.
Responsibilities
Design & implement highly available, scalable, and fault-tolerant systems with a programming-driven approach to problem-solving.
Partner closely with software developers, applying your multi-language programming skills (e.g., Python, Go, Java, or others) to build tools, services, and automation that improve reliability.
Drive adoption of Infrastructure as Code (IaC) using Terraform and other technologies, ensuring repeatable, version-controlled deployments.
Design, build, and maintain CI/CD pipelines — integrating automated testing, linting, and deployment strategies informed by software development best practices.
Implement and manage observability solutions (monitoring, logging, tracing) that provide actionable insights into application performance and infrastructure health.
Participate in code reviews for infrastructure-related services, promoting high-quality, maintainable, and secure code.
Mentor junior engineers on both SRE principles and coding standards across languages.
Participate in incident response activities, perform root cause analysis, and implement long-term preventative measures — often via code-driven solutions.
Evaluate and integrate new tools, frameworks, and programming techniques to improve operational efficiency and team productivity.
Contribute to the technical direction of the SRE team, shaping priorities with a developer’s mindset.
Requirements
Bachelor’s degree in Computer Science, Software Engineering, or equivalent practical experience.
6+ years of combined experience in SRE, DevOps, or software engineering roles.
Proven expertise in designing and supporting distributed systems at scale.
Solid professional experience in multiple programming languages (e.g., Python, Go, Java, C#, or JavaScript/TypeScript) with strong debugging and code optimization skills.
Hands-on experience with IaC tools — especially Terraform.
Senior DevOps Engineer improving Sherweb’s cloud - based IT solution delivery through CI/CD, IaC, GitOps, and AI automation. Supporting secure platforms, operational transitions, and developer self - service.
Senior DevOps / Cloud Infrastructure Engineer needed for hybrid role in North York, ON. Requires 10+ years experience with GCP, AWS, Kubernetes, Terraform, and CI/CD.
Staff Site Reliability Engineer strengthening AWS and Kubernetes resilience for Caseware, a fintech company building audit and accounting software. Driving secure delivery, observability, and incident management.
Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.
Mozilla release engineer optimizing Firefox build, test, and deployment pipelines at global scale. Improving developer experience, maintaining automation, and responding to critical service outages without an on - call rotation.