Senior Software Engineer improving reliability and security across Coinbase's services. Design and deliver projects for resilience and safe deployments within a fast-paced remote work environment.
Responsibilities
Own the design and delivery of reliability projects and features that improve resiliency across Coinbase's service environment in partnership with other engineering teams.
Partner with critical T0/T1 services to understand architecture, improve scalability, and reduce operational toil.
Build and enhance systems that securely manage service configurations and secrets at scale.
Improve canary-based release systems and expand deployment capabilities to support thousands of services and hundreds of daily deployments with fewer incidents.
Drive reliability best practices and strengthen reliability culture across engineering teams at Coinbase.
Requirements
5+ years of software engineering experience designing, building, and maintaining production services in service-oriented architectures, including experience with Ruby, Go, Terraform, and cloud platforms (AWS, GCP or Azure).
Demonstrated ability to design and operate reliable, high-throughput, low-latency distributed systems at scale, with a track record of writing well-tested, production-quality code.
Proven experience with observability and monitoring tools (e.g., Kibana, Datadog) to debug complex production issues, tune system performance, and reduce incident frequency.
Experience writing and verbally communicating architecture decisions to cross-functional engineering stakeholders.
Ability to participate in on-call rotations and respond to issues outside normal business hours.
Utilizes generative AI responsibly, maintaining human oversight to deliver business-ready outputs and drive measurable improvements in workflow efficiency, cost, and quality.
Benefits
Total compensation may also include equity and bonus eligibility and benefits (including medical, dental, and vision)
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.
Team Leader guiding Remote’s global SRE platform for compliant international employment. Leading engineers and reliability across Kubernetes, AWS, observability, and infrastructure.
AWS and DevOps Engineer establishing secure, automated environments for a bilingual nonprofit digital platform. Managing deployment, monitoring, recovery, and operational handover.