Software Development Team Lead – Infrastructure Platform

Posted 5 days ago

Apply Now

Resume Score

Check how well your resume matches this job before you apply.

Sign in to check score

About the role

  • Infrastructure team lead building and operating Spare’s GCP and Kubernetes platform for on-demand transit. Leading developers while improving reliability, security, AI SRE, and cloud cost efficiency.

Responsibilities

  • Lead the Infrastructure Platform Team responsible for Spare’s core cloud infrastructure platform
  • Own design and development of infrastructure platform capabilities from inception to launch
  • Build and evolve tooling, automation, and platform services for engineering teams
  • Architect and implement scalable distributed systems on GCP and Kubernetes
  • Improve cluster reliability, application resilience, and internal access security
  • Operate and maintain Redis and PostgreSQL databases, including tuning, scaling, upgrades, backups, and disaster recovery
  • Drive AI SRE practices for incident detection, alert triage, and operational workflows
  • Manage and improve the SRE on-call rotation, escalation paths, and blameless post-mortems
  • Drive FinOps, cloud spend visibility, right-sizing, and cost optimization
  • Use AI agentic tooling daily and coach the team in its use
  • Mentor software developers and increase team capacity
  • Collaborate with product managers, designers, and software developers
  • Ensure 99.99% uptime and exceptional system performance
  • Participate in agile rituals and improve software development processes
  • Split time approximately 50/50 between hands-on technical contribution and people leadership
  • Travel to up to four customer site visits per year and attend biannual Vancouver hackathons
  • Manage four direct-report software developers

Requirements

  • 7+ years of software development experience, with at least 2+ years in a people leadership role
  • Expert backend technology and strong distributed systems experience
  • Proficiency with AI-assisted and agentic development workflows
  • Experience operating systems at scale with a strong reliability and uptime mindset
  • Experience running or managing an SRE on-call rotation, including incident response and post-mortem culture
  • Deep experience with GCP and Kubernetes
  • Experience driving cloud cost optimization and FinOps initiatives
  • Experience with infrastructure-as-code and configuration management tooling, especially Terraform
  • Understanding of security best practices, internal access control, and container security
  • Demonstrated success managing software developers and individual/team performance
  • Demonstrated ability to mentor developers and provide technical leadership
  • Strong problem-solving, debugging, and system design skills
  • Excellent communication and collaboration skills
  • Legally entitled to work in Canada without additional sponsorship
  • Able to work hybrid three days per week in downtown Vancouver
  • Nice-to-have: transit or safety-critical domain experience; internal developer platforms; AI/LLM tooling for SRE; CI/CD at scale; PostgreSQL and Redis production operations

Benefits

  • Equity options
  • Competitive salary
  • Opportunity to work on challenging technical problems with real-world impact
  • Fast-paced, high-impact role in a rapidly growing startup
  • Ownership of core systems and opportunity to drive innovation
  • Dynamic, collaborative, and supportive team culture
  • Participation in biannual software development hackathons in Vancouver

Job type

Full Time

Experience level

Senior

Salary

CA$150,000 - CA$260,000 per year

Degree requirement

No Education Requirement

Tech skills

CloudDistributed SystemsGoogle Cloud PlatformKubernetesPostgresRedisTerraform

Location requirements

HybridVancouverCanada

Report this job

Found something wrong with the page? Please let us know by submitting a report below.