Senior Platform Engineer building resilient Go and Kubernetes platforms for Virtasant’s cloud-native products. Owning reliability, infrastructure as code, observability, and developer delivery tools.
Responsibilities
Design, build, and operate production Kubernetes clusters, including networking, workload isolation, and multi-region topologies
Work with Kubernetes internals, resource quotas, scheduling, NetworkPolicy enforcement, and custom controllers or operators
Implement and operate service mesh capabilities, including mTLS, workload authentication and authorization, and traffic management
Optimize containerized workloads for performance, cost, and resource efficiency
Write, refactor, and maintain production Go services, controllers, and middleware
Build HTTP, REST, and gRPC interfaces for internal engineering teams
Write unit and integration tests
Lead incident response, investigate root causes, and author postmortems
Troubleshoot production systems using logs, metrics, traces, and profiling tools
Diagnose and resolve distributed-system performance and reliability problems
Define and drive SLOs and actionable alerting
Own infrastructure as code and maintain reusable modules
Build and improve CI/CD and GitOps delivery workflows
Balance developer velocity with reliability, security, and compliance
Build and maintain metrics, dashboards, alerting policies, and distributed tracing
Partner with product, security, and infrastructure teams on requirements and architecture
Contribute to design reviews, set technical direction, and mentor engineers
Requirements
6+ years of professional experience in software, platform, infrastructure, or site reliability engineering
Significant experience operating production distributed systems
Demonstrated experience building and operating production Kubernetes platforms
Production experience writing Go and ability to use Go as the primary day-to-day language
Experience taking ambiguous system designs through production
Degree in Computer Science, Engineering, or a related field, or equivalent practical experience
Strong understanding of Kubernetes internals, including CNI networking, NetworkPolicy, resource management, and cluster behavior under load
Hands-on experience with a service mesh such as Istio, Envoy, or Linkerd, plus mTLS and workload identity
Solid Linux fundamentals, including cgroups and resource management
Infrastructure as code at scale using Terraform or equivalent
Production experience with GCP, AWS, or Azure
Experience with Prometheus, Grafana, OpenTelemetry, and query languages such as PromQL
Production experience with relational databases, including PostgreSQL or managed Postgres-compatible services, and replication and failover
Docker and container tooling experience
Strong debugging and performance profiling skills
Strong analytical and problem-solving ability
Clear written and verbal communication
Ability to work independently within a distributed team
Comfort working in a fast-moving, highly technical environment
Senior DevOps Engineer operating AWS, Kubernetes, and blockchain infrastructure for Startale’s onchain finance products. Owning reliability, security, deployment, and production operations for StartaleApp and Strium.
Senior Cloud Infrastructure Engineer scaling AWS networking and storage for Mecka AI’s robotics and embodied AI data infrastructure. Leading reliability, security, transfer, and cost optimization.
Senior infrastructure developer owning AWS environments, internal applications, and enterprise integrations. Supporting Benevity’s technology platform that enables companies and employees to take social action.
Senior developer owning AWS infrastructure, security, and integrations for Benevity’s technology platform. Supporting internal applications that help companies and employees take social action.
Infrastructure Engineer building secure, automated Azure environments for PLATO, Canada’s largest Indigenous - owned software testing and technology services company. Applying Terraform, IaC, and DevOps practices across cloud infrastructure.
Infrastructure support engineer handling incidents, troubleshooting, monitoring, and service requests for Genpact’s enterprise technology services. Supporting hybrid operations on rotational night shifts in Montreal.
Senior Infrastructure Engineer rebuilding Stream’s real - time platform as it migrates from AWS to GCP. Owning Kubernetes, PostgreSQL scaling, cloud efficiency, and reliability.
Senior Infrastructure Engineer designing and delivering customer infrastructure solutions for Redesign Group, a technology and cybersecurity provider. Leading implementations, documentation, technical reviews, and engineer mentorship in Toronto.