Senior SRE/Platform Engineer needed for global company. 6+ years SRE experience, AWS/Azure, Kubernetes, Terraform, observability tools. Contract-to-hire in Mississauga.
Responsibilities
As a Senior Site Reliability/Platform Engineer, you will be responsible for supporting production infrastructure, managing container orchestration with Kubernetes, implementing infrastructure as code using Terraform and other tools, and ensuring the reliability and scalability of enterprise-grade software systems.
Requirements
6+ years of experience as an SRE supporting production infrastructure. 6+ years of overall software engineering experience. Extensive experience with AWS OR Azure. Experience with Kubernetes and container orchestration. Experience with IAC tools (Terraform, Docker, Helm, Packer, Ansible, ARM). Experience with configuration management tools (Ansible, YAML, Terraform). Experience with PowerShell and Shell scripting. Experience with observability tools (Grafana, Kibana, Prometheus). Experience with enterprise-grade software, microservices architecture, and at least two years managing Kubernetes production systems.
Benefits
Rate: $80-$95 per hour INC / up to $150k base + benefits.
Principal Platform Engineer leading SkyWatch’s satellite - data platform and AI agent infrastructure. Owning architecture, customer - driven roadmap delivery, and platform engineering leadership.
Platform Engineer securing Just Eat Takeaway.com’s global food - delivery edge infrastructure. Building gateways, automation, and resilient traffic routing across production environments.
Director leading Blackpoint Cyber’s cloud - based Unified Security Posture data platform for cybersecurity solutions. Driving platform roadmap, reliability, APIs, data engineering, and team growth.
Ingénieur logiciel principal intégrant des plateformes, API et solutions IA chez EDC. Gouvernance technique, sécurité, résilience et mentorat dans une société canadienne de financement du commerce.
Senior AI Platform Developer building scalable AI services and LLM workflows for MaintainX’s industrial work execution platform. Improving reliability, observability, performance, and cost efficiency.
Infrastructure team lead building and operating Spare’s GCP and Kubernetes platform for on - demand transit. Leading developers while improving reliability, security, AI SRE, and cloud cost efficiency.
Senior AI Platform Developer building reusable, secure AI - agent infrastructure for Petal, a Canadian healthcare orchestration and billing company. Driving Azure - based platform capabilities across orchestration, evaluation, observability, and governance.
AI Platform Developer building Azure - based agent infrastructure for Petal, a Canadian healthcare orchestration and billing company. Creating secure, observable, governed AI services for product teams.
Kubernetes/DevOps Engineer operating Kubernetes infrastructure for a petabyte - scale social media platform. Building distributed systems and machine - learning workloads for Capgemini Engineering’s client.