Senior Data Platform Engineer developing infrastructure for data processing at Tubi in Toronto. Handling Spark-on-Kubernetes and real-time feature pipelines with responsibilities across data platform components.
Responsibilities
Spark-on-Kubernetes — EKS-based compute platform for Spark workloads: cluster configuration, Pod Identity IAM, job environment setup, Kustomize overlays, and shadow canary validation
Event ingestion — Rust services and Flink jobs processing billions of events per day over Kinesis; throughput, reliability, on-call response, and AI-assisted operational tooling to reduce toil
Platform infrastructure — Terraform modules for environment provisioning, cross-account AWS IAM, ARC runner infrastructure, and CI/CD for data platform changes
Feature store and ML compute — Flink-based real-time feature pipelines feeding a large-scale MemoryDB cluster; GPU capacity governance and Databricks multi-environment operations for ML training workloads
Workflow orchestration and CDC — Airflow-based DAG deployment, change data capture pipeline operations, and data quality monitoring
Requirements
3+ years building and operating production data platform infrastructure at the cluster or platform level, across Spark, Flink, Kinesis, Kubernetes, or equivalent
Deep experience in at least one of: Spark-on-K8s cluster operations, Rust-based data or systems engineering, Kubernetes platform engineering and IaC, or data catalog and governance tooling
Production AWS experience or equivalent: EKS, S3, Kinesis, and multi-account IAM patterns (EKS Pod Identity, KIAM, or IRSA)
You've owned a critical platform component, you wrote the runbooks, tracked the cost, and were on-call for it
Strong in at least one of: Rust, JVM (Java or Scala), or Python for data platform work.
Benefits
This role is also eligible for an annual discretionary bonus
long-term incentive plan
medical/dental/vision
insurance
vacation/paid time off
Flexible Time Off Policy to manage all personal matters
generous Parental Leave Program allowing parents twelve (12) weeks of paid bonding leave
Staff Platform Engineer defining cloud and DevSecOps strategy for Robots & Pencils’ enterprise AI systems. Leading Kubernetes, AI/ML infrastructure, migrations, reliability, security, and platform standards remotely in Canada.
Principal platform developer designing AWS - native integrations, CI/CD pipelines, and developer tooling for Autodesk’s design software. Leading architecture, reliability, and cross - team engineering initiatives.
Senior MLOps Developer operationalizing machine learning models and scalable AI/ML infrastructure for Autodesk’s design and entertainment software. Building deployment, monitoring, governance, and recovery systems.
Senior SRE/Platform Engineer needed for global company. 6+ years SRE experience, AWS/Azure, Kubernetes, Terraform, observability tools. Contract - to - hire in Mississauga.
Senior Full Stack Engineer building GraphQL, React, and TypeScript platforms for PENN Entertainment’s online gaming and sports media products. Improving shared client tooling, server - driven UI, performance, observability, and release workflows.
Senior platform engineering lead shaping compute and virtualization strategy for BMO, a major bank. Driving modernization, architecture standards, automation and hybrid - cloud infrastructure transformation.
Software Engineer building backend AI platform systems for DraftKings’ sports entertainment and gaming technology. Developing retrieval, vector database, automation, and agent infrastructure for scalable AI applications.
Staff Platform Software Engineer building reusable infrastructure for Cantina Labs’ social AI platform. Improving search, AWS operations, Go services, observability, and developer productivity.