Senior Platform Engineer building and evolving internal platforms for Shakepay to enhance engineering productivity. Collaborating with teams to improve system reliability and performance.
Responsibilities
Design, build, and maintain infrastructure that supports Shakepay’s growing platform
Develop internal tooling and automation that improve developer productivity and operational efficiency
Implement infrastructure-as-code and GitOps practices to manage our systems
Build and maintain observability systems including monitoring, alerting, and dashboards
Design and build AI-powered internal tools and agentic workflows that automate tasks and augment engineering productivity
Improve system reliability, performance, and scalability across our platform
Collaborate closely with product engineers on infrastructure design and operational best practices
Participate in incident response and contribute to improving operational processes
Mentor engineers through code reviews, design discussions, and knowledge sharing
Help shape the technical direction of our infrastructure and developer platform
Requirements
5+ years of experience in software engineering, platform engineering, SRE, or DevOps roles
Strong programming experience (e.g., Go, Python, JavaScript, or similar)
Experience with infrastructure-as-code and automated deployment systems
Strong understanding of distributed systems, reliability, and performance
Experience improving developer experience through tooling or platform improvements
Experience building tools or workflows that leverage AI/LLMs to automate tasks or augment developer productivity
Experience working in highly collaborative engineering teams
A pragmatic mindset focused on solving real business problems.
Benefits
🤖 AI Enablement: Generous AI token budget (currently unlimited)
🤝 Be an owner - Every employee has stock options as part of their total compensation.
🥅 Reach your goals - Yearly salary assessments.
🦷 Health & wellness: Access to health and dental coverage, including health and wellness spending accounts.
🌎 Remote-friendly: Work from anywhere in Canada, with optional access to our office spaces in Montreal and Toronto.
🆙 Level Up: A $2,000 annual budget for courses, certifications, and training to support your career growth.
🌴 Time off: 20 days of vacation per year. We will give you a $1,000 bonus if you use all your vacation time.
🐣 Parental leave: Enjoy a parental leave top-up to 100% of your salary for 18 weeks.
🙌 Have fun together: quarterly team-specific or company-wide offsite to connect with each other.
Backend Platform Engineer scaling Hapiko’s real - time APIs, queues, infrastructure and observability. Building reliable, secure cloud systems for Spin Master’s voice - activated kids’ sticker printer.
Azure Platform Engineer operating secure, reliable AKS infrastructure for Smile Digital Health’s FHIR - based healthcare data platform. Owning Kubernetes networking, CI/CD, observability, security, and production operations.
Senior Platform Engineer building scalable backend and cloud infrastructure for ExaCare’s AI - powered post - acute care platform. Improving reliability, developer velocity, and healthcare admission workflows.
Principal Platform Engineer leading SkyWatch’s satellite - data platform and AI agent infrastructure. Owning architecture, customer - driven roadmap delivery, and platform engineering leadership.
Platform Engineer securing Just Eat Takeaway.com’s global food - delivery edge infrastructure. Building gateways, automation, and resilient traffic routing across production environments.
Director leading Blackpoint Cyber’s cloud - based Unified Security Posture data platform for cybersecurity solutions. Driving platform roadmap, reliability, APIs, data engineering, and team growth.
Ingénieur logiciel principal intégrant des plateformes, API et solutions IA chez EDC. Gouvernance technique, sécurité, résilience et mentorat dans une société canadienne de financement du commerce.
Senior AI Platform Developer building scalable AI services and LLM workflows for MaintainX’s industrial work execution platform. Improving reliability, observability, performance, and cost efficiency.
Infrastructure team lead building and operating Spare’s GCP and Kubernetes platform for on - demand transit. Leading developers while improving reliability, security, AI SRE, and cloud cost efficiency.