Staff Software Engineer contributing across Newrich's product, improving infrastructure and workflows while handling production code for creator platform.
Responsibilities
You will work across our infra stack (AWS, ECS, Fargate, RDS, S3, EC2, Lambda) contributing directly to production code while helping improve workflows, documentation, and processes.
This is a hands-on role for someone who thrives in a dynamic environment and takes ownership of getting things done.
Requirements
Must have a hands-on experience on setting up Kubernetes platform, deploying microservices and other web applications, and managing secure secrets.
Knowledge and practical experience with Docker is a must - including setting up and managing Docker registries as well as creating Dockerfiles to create custom images.
Should have working knowledge of container orchestration using Kubernetes.
Should have knowledge of overlay networking needed for inter-container communications from different nodes.
Experience building CI/CD pipelines with Github Actions.
Experience automating systems deployments and configuration management using tools like Ansible, Chef, Puppet, Terraform, Saltstack.
Working experience with source control systems like Git.
Ability to work well with people from many different disciplines with varying degrees of technical experience.
Ability to demonstrate a clear, energetic and excited interest in automating everything (build, test, release/deploy, monitoring, reporting), which includes "Infrastructure as Code ".
Able to build, deploy, and host a demo web app written in the code of your choice to prove minimum general IT/DevOps proficiency.
Experience working as DevOps, SysOps, SRE or similar for at least 2 years.
Familiarity with AWS services, including CloudTrail, AWS Organizations, EC2, ECS, EKS, Elasticache, IAM, Lambda, RDS/Aurora, Route53, and S3.
Experienced with Docker and deploying micro services.
Understands how to control Cloud costs.
Proficiency in Terraform for infrastructure provisioning and automation.
Strong working knowledge of Linux and shell scripting.
Ability to work remotely and effectively collaborate with distributed teams.
Proficiency in at least one programming language such as PHP, Javascript, or Python is a plus.
Understanding of incident management processes and service level objectives (SLOs).
Ability to perform root cause analysis in AWS environments and debug micro services using Cloudwatch or related monitoring products.
Comfortable being part of the On-Call rotation.
This position will be part-time/hourly to start- with the option to go full time after 3 months.
Benefits
**Paid Adventure Time** – Take an all-expenses-paid remote working trip for 3 weeks to a destination of your choice with one of our remote work-trip partners. On top of that, you’ll have “Me-Days” – flexible personal days you can take whenever you need a reset.
**Fast Growth, Big Upside** – We’re a small, ambitious team. That means more ownership, faster learning, and a real chance to shape the future of our company (and your career).
**Equity + Bonus - Take some early equity and grab a piece of our future success. Bonuses paid out based on company hitting specific milestones and KPIs.**
**Unlimited Learning** – You’ll get full access to every course and program on our NewRich platform. We invest in your growth because your growth fuels ours.
**Home Office Stipend **– Your setup matters. We’ll support you with a budget to create your ideal workspace and provide you with a new MacBook to power your productivity.
**Annual Retreat** – Work remote, but meet the team IRL. Every year we gather in amazing locations – next stop: Colombia.
Senior DevOps Engineer building secure, scalable Azure platforms for CARET’s legal and accounting practice - management software. Leading infrastructure, Kubernetes, CI/CD, security, observability, and reliability initiatives.
Senior SRE operating Kubernetes and cloud infrastructure for Penn Entertainment’s sports betting and media platforms. Leading migrations, automation, observability, and incident response across regulated production services.
DevOps Engineer operating multi - cloud Kubernetes infrastructure for InfluxData’s time - series platform. Automating operations and supporting highly available distributed services.
Senior Reliability Engineer improving embedded protection, control, and software products for utility grids. Leading reliability testing, failure analysis, and modernization initiatives for resilient energy systems.
Senior Reliability Engineer improving embedded grid automation reliability for utility - scale energy systems. Leading testing, failure analysis, KPIs, and modernization initiatives with utilities.
Site Reliability Expert leading observability and SRE for Valtech, an experience innovation company. Improving reliability across cloud - native, microservices - based environments.
Staff DevOps Engineer owning reliable, scalable infrastructure for Nexxa’s AI systems. Supporting machine learning workloads across heavy - industry operations.
Staff SRE securing and scaling IAM systems at RBC, a Canadian bank. Designing resilient infrastructure, automating operations, and leading incident response.