Senior Infrastructure Engineer rebuilding Stream’s real-time platform as it migrates from AWS to GCP. Owning Kubernetes, PostgreSQL scaling, cloud efficiency, and reliability.
Responsibilities
Design, build and operate infrastructure for real-time systems carrying millions of concurrent connections and billions of monthly API requests
Drive Kubernetes end to end, including cluster architecture, workload design and migration of existing services
Re-architect workloads for the AWS-to-GCP migration for cost and performance
Own cloud cost and efficiency work using real spend and utilisation data
Write production Go and Python for internal services, platform tooling and automation
Lead post-migration tuning and capacity planning
Collaborate with backend, video and moderation engineers on system design, reliability targets and cross-service tradeoffs
Participate in on-call, incident response and root cause analysis, turning findings into durable fixes
Own infrastructure projects end to end on a small senior team
Requirements
5+ years in infrastructure, platform, DevOps or SRE engineering, with clear depth in infrastructure over application development
A software engineering background; built systems, not only configured them
Production coding experience in Go or Python; scripting-only backgrounds are not a fit
Kubernetes experience at meaningful production scale, including cluster strategy, workload design, migration leadership, and post-migration cost and efficiency tuning
Personally led cloud cost or efficiency optimisation on AWS or GCP, with a measurable outcome
Direct experience running high-scale, high-load production systems
Strong cloud fundamentals across networking, compute, storage and IAM
Comfortable leading projects and reviewing PRs in a small team
AI tooling already in your engineering workflow; applied use, not familiarity
Both AWS and GCP, migration experience between providers
PostgreSQL at scale, including sharding, replication strategy, partitioning tradeoffs, ideally self-hosted
Experience with real-time systems such as WebSockets, WebRTC, streaming or persistent-connection workloads
Experience with CockroachDB, Redis, Terraform, and a Prometheus-based observability stack
Experience at an API-first or infrastructure company at scaleup stage
Open source contributions to infrastructure or platform tooling
Writing or talks on cloud, platform or distributed systems
Formal FinOps practice or ownership of cloud commitment and reservation strategy
Work on developer-facing API or SDK products
Benefits
A combination of 36 days per year in PTO and public holidays
Company equity
Remote work flexibility
Fitness stipend
A Macbook Pro provided
A Learning and Development budget
The opportunity to attend or present to global conferences and meetups
The chance to work on OSS projects
The possibility to visit our offices in Boulder, CO and Amsterdam, NL
Senior DevOps Engineer operating AWS, Kubernetes, and blockchain infrastructure for Startale’s onchain finance products. Owning reliability, security, deployment, and production operations for StartaleApp and Strium.
Senior Cloud Infrastructure Engineer scaling AWS networking and storage for Mecka AI’s robotics and embodied AI data infrastructure. Leading reliability, security, transfer, and cost optimization.
Senior infrastructure developer owning AWS environments, internal applications, and enterprise integrations. Supporting Benevity’s technology platform that enables companies and employees to take social action.
Senior developer owning AWS infrastructure, security, and integrations for Benevity’s technology platform. Supporting internal applications that help companies and employees take social action.
Infrastructure Engineer building secure, automated Azure environments for PLATO, Canada’s largest Indigenous - owned software testing and technology services company. Applying Terraform, IaC, and DevOps practices across cloud infrastructure.
Infrastructure support engineer handling incidents, troubleshooting, monitoring, and service requests for Genpact’s enterprise technology services. Supporting hybrid operations on rotational night shifts in Montreal.
Senior Infrastructure Engineer designing and delivering customer infrastructure solutions for Redesign Group, a technology and cybersecurity provider. Leading implementations, documentation, technical reviews, and engineer mentorship in Toronto.
Security and Infrastructure Engineer securing Nexxa’s AI systems for heavy industries. Owning cloud infrastructure, security operations, and SOC 2/ISO 27001 compliance.