Resume Score

Check how well your resume matches this job before you apply.

Sign in to check score

About the role

  • Platform engineer maintaining Best Buy’s enterprise monitoring, logging, and reliability platforms. Automating infrastructure operations and improving platform stability with DevOps and cloud technologies.

Responsibilities

  • Design, develop, and enhance self-service capabilities for enterprise reliability and observability platforms
  • Develop and maintain monitoring, logging, and alerting capabilities to proactively identify and address platform and application issues
  • Deploy, administer, and maintain reliability platforms using Infrastructure as Code and configuration management practices
  • Own platform lifecycle management, including installation, configuration, upgrades, patching, vulnerability remediation, and decommissioning
  • Identify and implement automation opportunities to improve platform operations and engineering productivity
  • Collaborate with platform development engineers, DevOps, infrastructure, network, and product teams
  • Participate in incident reviews, root cause analysis, and continuous improvement initiatives
  • Conduct capacity planning and resource optimization with various teams
  • Maintain documentation of platform architectures, configurations, operational procedures, and automation workflows
  • Evaluate and implement platform improvements using industry best practices, emerging technologies, observability capabilities, and AI-assisted operational tools

Requirements

  • 3+ years of experience in Platform Engineering, DevOps, Site Reliability Engineering, or Infrastructure Engineering
  • Strong scripting and automation experience using Python, PowerShell, Bash, or similar technologies
  • 2+ years of experience using Infrastructure as Code (IaC) and configuration management tools, such as Terraform and Ansible
  • 2+ years of experience administering, developing, or supporting enterprise monitoring and observability platforms, such as Dynatrace, Prometheus, Grafana, Azure Monitor, and SolarWinds
  • 2+ years of experience working with logging, alerting, and operational platforms, such as Elastic, PagerDuty, GoAlert, and Splunk
  • 2+ years of experience working with cloud platforms such as Azure, AWS, or GCP
  • 2+ years of experience deploying and supporting containerized applications and orchestration platforms, such as OpenShift, Kubernetes, and Docker
  • Experience with platform upgrades, patching, vulnerability remediation, and lifecycle management of enterprise platforms
  • Experience working with Linux operating systems
  • Strong analytical, troubleshooting, problem-solving, and collaborative skills

Benefits

  • Remote-first work environment
  • Employee discounts on awesome tech from day one
  • Flexible health benefits and wellness program
  • TFSA and RRSP programs
  • 100% matched company pension plan
  • Training programs to build new and transferable skills

Job title

Job type

Full Time

Experience level

Mid levelSenior

Salary

CA$90,000 - CA$110,000 per year

Degree requirement

No Education Requirement

Tech skills

AnsibleAWSAzureCloudDockerGoogle Cloud PlatformGrafanaKubernetesLinuxOpenShiftPrometheusPythonSplunkTerraform

Location requirements

HybridVancouverCanada

Report this job

Found something wrong with the page? Please let us know by submitting a report below.