Site Reliability Engineer managing Zora’s infrastructure to ensure availability, scalability, and efficiency. Collaborating with teams to automate workflows and enhance the developer experience.
Responsibilities
Design, build, and deliver software to enhance the availability, scalability, latency, and efficiency of Zora’s infrastructure platform
Provide technical and strategic input to shape the direction of the infrastructure platform
Operate and maintain core infrastructure systems in service of enhancing the developer experience
Automate key infrastructure workflows, including service lifecycle management and critical operational processes
Participate in the team’s on-call rotation and respond to production incidents as needed
Requirements
5+ years of experience in infrastructure and/or platform engineering, with a focus on software development
Proficiency in Go and Python
Hands-on experience with Kubernetes, Docker, and similar orchestration systems
Operational experience with distributed storage systems like Ceph
Familiarity with Python’s asyncio programming model
Experience with frontend serverless platforms like Vercel, Cloudflare Pages Functions, etc
Understanding of Ethereum and EVM-compatible chains, including L2 architectures like Polygon, Optimism, and Base
Knowledge of L2 scaling solutions such as optimistic rollups, zk-rollups, Plasma, sidechains, and bridges
Passion for learning and exploring new technologies and paradigms
Experience designing large-scale systems, breaking down complex problems, and aligning with cross-functional engineering teams
Familiarity with blockchain infrastructure, including RPC nodes and IPFS
Strong experience with database technologies such as MongoDB and Postgres
Solid understanding of monitoring and observability practices using tools like Datadog and/or OpenTelemetry
Strong debugging and troubleshooting skills, from orchestration layers down to system runtimes
Clear and effective communication skills, with a collaborative and service-oriented mindset
Benefits
Remote-First Culture: Work from anywhere in the world!
Competitive Compensation: Including salary, pre-IPO stock options, token compensation, and additional financial incentives
Comprehensive Benefits: Robust healthcare options, including fully covered medical, dental, and vision for employees
Retirement Contributions: Up to 4% employer match on your 401(k) contributions
Health & Wellness: Free memberships to One Medical, Teladoc, and Health Advocate
Unlimited Time Off: Flexible vacation policies, company holidays, and recharge weeks to prioritize wellness
Home Office Reimbursement: To cover home office items, monthly home internet, and monthly cell phone (if applicable)
Ease of Life Reimbursement: To cover everything from an Uber home in the rain, childcare, or meal delivery
Career Development: Access to mentorship, training, and opportunities to grow your career
Inclusive Environment: A culture dedicated to diversity, equity, inclusion, and belonging
DevOps Intern supporting CI/CD, cloud infrastructure, and automation for Ludia’s mobile game studio. Improving reliability and developer tools in production game environments.
Manager leading global SRE teams for Akamai's distributed Cloud IAM services. Improving reliability, scalability, security, and usability through cloud - native tooling and software.
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.