Senior Infrastructure & Platform Engineer at UserTesting designing AWS infrastructure and Kubernetes clusters. Focused on building internal platform abilities and improving developer experience through self-service.
Responsibilities
Design, build, and maintain AWS infrastructure using Terraform, reusable modules, and modern IaC patterns.
Develop and operate Kubernetes (EKS) clusters, ensuring scalability, cost effectiveness, and secure-by-default configurations.
Shape and evolve internal platform capabilities: golden paths, templates, and self-service tooling that enable engineering teams to ship faster and safer.
Enhance and standardize CI/CD experiences using GitHub Actions to promote consistency, reliability, and high developer satisfaction.
Implement policy-as-code and automated guardrails that embed security, cost controls, and operational excellence into the platform layer.
Build and maintain shared monitoring, metrics, and logging foundations (Prometheus, Grafana, Datadog, etc.).
Work closely with product engineering teams to understand friction points and co-design abstractions that simplify complex infrastructure concerns.
Improve internal documentation, onboarding guides, and developer experiences that scale through self-service, not manual support.
Requirements
7+ years in Infrastructure, SRE, DevOps, or similar roles supporting cloud-native SaaS environments.
Deep proficiency with AWS services (compute, networking, IAM, storage) and Kubernetes at production scale (preferably EKS).
Expertise in Terraform, GitHub Actions, and automation-first workflows.
Strong scripting ability (Python, Bash) and hands-on experience with observability tooling.
A platform-product mindset: you care about user experience, maintainability, and measurable improvements to developer velocity.
Passion for automation, well-designed abstractions, and building systems that scale with adoption—not headcount.
Ability to partner effectively across teams, communicate clearly, and influence through technical leadership.
Commitment to reliability, security, and performance as built-in characteristics—not afterthoughts.
Senior Platform Engineer building scalable backend and cloud infrastructure for ExaCare’s AI - powered post - acute care platform. Improving reliability, developer velocity, and healthcare admission workflows.
Principal Platform Engineer leading SkyWatch’s satellite - data platform and AI agent infrastructure. Owning architecture, customer - driven roadmap delivery, and platform engineering leadership.
Platform Engineer securing Just Eat Takeaway.com’s global food - delivery edge infrastructure. Building gateways, automation, and resilient traffic routing across production environments.
Director leading Blackpoint Cyber’s cloud - based Unified Security Posture data platform for cybersecurity solutions. Driving platform roadmap, reliability, APIs, data engineering, and team growth.
Ingénieur logiciel principal intégrant des plateformes, API et solutions IA chez EDC. Gouvernance technique, sécurité, résilience et mentorat dans une société canadienne de financement du commerce.
Senior AI Platform Developer building scalable AI services and LLM workflows for MaintainX’s industrial work execution platform. Improving reliability, observability, performance, and cost efficiency.
Infrastructure team lead building and operating Spare’s GCP and Kubernetes platform for on - demand transit. Leading developers while improving reliability, security, AI SRE, and cloud cost efficiency.
AI Platform Developer building Azure - based agent infrastructure for Petal, a Canadian healthcare orchestration and billing company. Creating secure, observable, governed AI services for product teams.
Senior AI Platform Developer building reusable, secure AI - agent infrastructure for Petal, a Canadian healthcare orchestration and billing company. Driving Azure - based platform capabilities across orchestration, evaluation, observability, and governance.