Senior Platform Engineer developing infrastructure for SaaS product at Orchestry. Responsible for building production-grade, containerized systems and ensuring scalability and reliability in the platform.
Responsibilities
Convert existing development-only Dockerfiles into production-ready, multi-stage container builds with health checks and graceful shutdown handling (including draining in-flight background jobs before termination)
Stand up and administer a container registry, including access policies, image tagging conventions, and retention/cleanup automation
Extend CI/CD pipelines to add container build-and-push steps, and migrate deployment targets from traditional package-based deployment to container-based App Service deployments
Author Infrastructure-as-Code modules from scratch covering compute, database, storage, secrets, and monitoring resources
Build provisioning automation for identity/app-registration setup, database initialization, and environment bootstrapping
Implement liveness and readiness health endpoints that verify downstream dependency connectivity (database, queue, background job processing)
Build a telemetry pipeline that exports sanitized, PII-free operational metrics to a centralized monitoring system, with configurable scope and destination
Design and maintain a tracking/registry system for infrastructure deployments, capturing version, status, and configuration state across many environments
Extend deployment pipelines to support automated, tag-based promotion and staged/canary rollout strategies, with per-environment failure isolation and rollback
Build a controlled mechanism for pushing critical patches outside the normal release cadence, with audit logging, notification workflows, and approval gates
Partner with engineering to audit core subsystems (background job processing, database partitioning/sharding, secrets access patterns, external integrations, feature-flag and analytics tooling, licensing/validation logic, notifications, scheduled tasks) for portability across deployment environments
Optimize databases (Azure SQL, Cosmos DB) for performance and scalability, including sharding and partitioning strategies
Ensure horizontal scaling strategies to handle SaaS growth efficiently
Implement caching (Redis, CDN) and performance tuning techniques
Lead incident response and root cause analysis (RCA) efforts, reducing mean time to recovery (MTTR)
Participate in the on-call rotation as a first responder to production incidents and emergency scenarios, providing timely triage and resolution outside of standard business hours
Work closely with Engineering, Security, and Product teams to align platform goals with business objectives
Act as a technical mentor for junior and mid-level engineers, fostering best practices in DevOps, cloud, and automation
Staff Platform Engineer defining cloud and DevSecOps strategy for Robots & Pencils’ enterprise AI systems. Leading Kubernetes, AI/ML infrastructure, migrations, reliability, security, and platform standards remotely in Canada.
Principal platform developer designing AWS - native integrations, CI/CD pipelines, and developer tooling for Autodesk’s design software. Leading architecture, reliability, and cross - team engineering initiatives.
Senior MLOps Developer operationalizing machine learning models and scalable AI/ML infrastructure for Autodesk’s design and entertainment software. Building deployment, monitoring, governance, and recovery systems.
Senior SRE/Platform Engineer needed for global company. 6+ years SRE experience, AWS/Azure, Kubernetes, Terraform, observability tools. Contract - to - hire in Mississauga.
Senior Full Stack Engineer building GraphQL, React, and TypeScript platforms for PENN Entertainment’s online gaming and sports media products. Improving shared client tooling, server - driven UI, performance, observability, and release workflows.
Senior platform engineering lead shaping compute and virtualization strategy for BMO, a major bank. Driving modernization, architecture standards, automation and hybrid - cloud infrastructure transformation.
Software Engineer building backend AI platform systems for DraftKings’ sports entertainment and gaming technology. Developing retrieval, vector database, automation, and agent infrastructure for scalable AI applications.