Senior operations leader needed to own reliability, stability, and performance of a large-scale Pega ecosystem. 6-month contract, hybrid in Toronto.
Responsibilities
This role sits at the core of platform health — where incident response, governance, and continuous improvement directly impact business continuity.
Requirements
10-15+ years in platform operations, application support, or technology operations; 5+ years leading L2/L3 support, SRE, or platform operations teams; strong experience with Pega platform operations (Pega 8.x+)
Backend Platform Engineer scaling Hapiko’s real - time APIs, queues, infrastructure and observability. Building reliable, secure cloud systems for Spin Master’s voice - activated kids’ sticker printer.
Azure Platform Engineer operating secure, reliable AKS infrastructure for Smile Digital Health’s FHIR - based healthcare data platform. Owning Kubernetes networking, CI/CD, observability, security, and production operations.
Senior Platform Engineer building scalable backend and cloud infrastructure for ExaCare’s AI - powered post - acute care platform. Improving reliability, developer velocity, and healthcare admission workflows.
Principal Platform Engineer leading SkyWatch’s satellite - data platform and AI agent infrastructure. Owning architecture, customer - driven roadmap delivery, and platform engineering leadership.
Platform Engineer securing Just Eat Takeaway.com’s global food - delivery edge infrastructure. Building gateways, automation, and resilient traffic routing across production environments.
Director leading Blackpoint Cyber’s cloud - based Unified Security Posture data platform for cybersecurity solutions. Driving platform roadmap, reliability, APIs, data engineering, and team growth.
Ingénieur logiciel principal intégrant des plateformes, API et solutions IA chez EDC. Gouvernance technique, sécurité, résilience et mentorat dans une société canadienne de financement du commerce.
Senior AI Platform Developer building scalable AI services and LLM workflows for MaintainX’s industrial work execution platform. Improving reliability, observability, performance, and cost efficiency.
Infrastructure team lead building and operating Spare’s GCP and Kubernetes platform for on - demand transit. Leading developers while improving reliability, security, AI SRE, and cloud cost efficiency.