Staff Site Reliability Engineer, Global Security

Posted 9 hours ago

Apply Now

Resume Score

Check how well your resume matches this job before you apply.

Sign in to check score

About the role

  • Staff SRE securing and scaling IAM systems at RBC, a Canadian bank. Designing resilient infrastructure, automating operations, and leading incident response.

Responsibilities

  • Serve as the senior-most technical voice for IAM reliability, setting architecture direction and reliability standards
  • Own end-to-end service reliability for IAM systems by defining and maintaining SLOs, SLIs, and error budgets
  • Design and implement resilient, highly available IAM infrastructure and services across multi-region and hybrid-cloud architectures
  • Write and review production-grade services, APIs, automation frameworks, and internal tools
  • Build self-service platforms, reusable modules, and golden-path automation for IAM services
  • Champion Infrastructure as Code and GitOps practices using Terraform, Ansible, Puppet, Kubernetes, and Helm
  • Own and improve CI/CD pipelines and release engineering for IAM services
  • Lead high-severity incident response, root cause analysis, blameless postmortems, and structural remediation
  • Build and evolve observability, monitoring, and alerting pipelines using metrics, logs, and traces
  • Develop and test failover strategies, backup validation, chaos engineering exercises, and disaster recovery simulations
  • Orchestrate workload automation, scheduling, and release pipelines across enterprise systems
  • Partner with security, infrastructure, application, and compliance teams on business continuity and resilience strategy
  • Mentor and coach engineers on reliability, automation-first thinking, and software engineering practices

Requirements

  • 5+ years of experience in Site Reliability Engineering, DevOps or Platform Engineering with demonstrated staff/senior-level technical leadership across SLOs/SLIs, error budgets, and driving continuous reliability improvement at scale
  • Proficient in at least one modern language (Python, Go, Java, or similar)
  • Experience designing, implementing, and operating highly available, fault-tolerant, and scalable systems in production, including hybrid and multi-cloud environments
  • Experience owning CI/CD pipelines and release engineering practices
  • Track record of turning recurring operational work into self-service tooling, reusable modules, or golden-path automation
  • Experience with Docker and Kubernetes in production environments
  • Deep knowledge of monitoring, alerting, and observability platforms
  • Proven incident management skills, including high-severity incident response, root cause analysis, and postmortems
  • Proficient in disaster recovery, failover strategies, and resilience testing
  • Understanding of AWS, Azure, and hybrid environments
  • Excellent collaboration and communication skills
  • Nice to have: Experience with IAM platforms and security-focused systems
  • Nice to have: Infrastructure as Code and configuration management with Terraform, Ansible, and Puppet; scripting with Python, PowerShell, and Bash
  • Nice to have: Familiarity with OAuth2, OIDC, SAML, LDAP, and SCIM
  • Nice to have: Knowledge of enterprise security architecture and compliance frameworks
  • Nice to have: Exposure to AIOps or ML-based anomaly detection

Benefits

  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable
  • Leaders who support your development through coaching and managing opportunities
  • Ability to make a difference and lasting impact
  • Work in a dynamic, collaborative, progressive, and high-performing team
  • A world-class training program in financial services
  • Opportunities to do challenging work

Job type

Full Time

Experience level

Lead

Salary

Not specified

Degree requirement

No Education Requirement

Tech skills

AnsibleAWSAzureCloudDockerJavaKubernetesPuppetPythonTerraformGo

Location requirements

OnsiteTorontoCanada

Report this job

Found something wrong with the page? Please let us know by submitting a report below.