Junior Site Reliability Engineer supporting system reliability and performance for accessible digital experiences at Fable. Collaborating with engineers to enhance infrastructure and developer experience.
Responsibilities
Support the reliability, performance, and scalability of Fable’s systems and infrastructure
Monitor systems, respond to alerts, and assist in troubleshooting production issues
Work with engineers and SREs to debug issues across infrastructure and application layers
Contribute to improving system observability, including logging, monitoring, and alerting
Assist in maintaining and improving CI/CD pipelines and deployment processes
Support infrastructure and configuration changes under guidance (e.g., cloud resources, environments)
Help identify opportunities to improve system performance, stability, and cost efficiency
Support efforts to reduce operational complexity and technical debt
Participate in incident response and post-incident reviews (postmortems)
Document systems, processes, and troubleshooting steps
Learn and apply best practices in reliability engineering, system design, and platform operations
Requirements
0–2 years of experience in software engineering, DevOps, SRE, or related fields (internships included)
Degree in Computer Science, Engineering, or a related field (or equivalent experience)
Basic understanding of software systems, APIs, and web applications
Familiarity with at least one programming or scripting language (e.g., Python, JavaScript, Bash)
Exposure to cloud platforms (e.g., AWS, GCP, Azure) or strong interest in learning
Basic understanding of Linux/Unix environments
Familiarity with version control (e.g., Git)
Strong problem-solving skills and willingness to learn
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.
Team Leader guiding Remote’s global SRE platform for compliant international employment. Leading engineers and reliability across Kubernetes, AWS, observability, and infrastructure.
AWS and DevOps Engineer establishing secure, automated environments for a bilingual nonprofit digital platform. Managing deployment, monitoring, recovery, and operational handover.