Site Reliability Engineer responsible for the installation, configuration, maintenance of middleware technologies at Hyve Solutions. Managing applications on container platforms and ensuring reliable operation of critical middleware components.
Responsibilities
Install, configure, and maintain middleware technologies such as Apache NiFi, Redis, Nginx, Hadoop Ecosystem, and related tools.
Deploy and manage applications on container platforms like Kubernetes.
Monitor performance, availability, and health of middleware services.
Perform routine patching, upgrades, backups, and recovery procedures.
Troubleshoot issues related to application connectivity, data flows, caching, and web serving.
Implement basic security configurations and access controls.
Automate repetitive tasks using scripting and configuration tools.
Provide support to development and operations teams.
Strengthen monitoring, alerting, performance degradation management, and emergency intervention for relevant applications.
Requirements
Bachelor's degree in Computer Science, IT, or related field (or equivalent experience).
3–5 years of experience in middleware or application administration.
Hands-on experience with Apache NiFi, Redis, Nginx, and Kubernetes.
Proficiency in Linux/Unix environments and shell scripting (Bash).
Familiarity with containerization and basic cloud platforms (AWS, Azure, GCP).
Understanding of networking, load balancing, TLS/SSL, and security best practices.
Experience with monitoring tools (Prometheus, Grafana, ELK stack).
Understand distributed system architecture and microservice design preferred.
Good problem-solving skills and ability to work in a team.
DevOps Intern supporting CI/CD, cloud infrastructure, and automation for Ludia’s mobile game studio. Improving reliability and developer tools in production game environments.
Manager leading global SRE teams for Akamai's distributed Cloud IAM services. Improving reliability, scalability, security, and usability through cloud - native tooling and software.
Cloud Engineer supporting Kinaxis’s AI - powered supply chain orchestration platform reliability. Automating cloud infrastructure, deployments, and production operations across Canadian locations.
Senior DevOps Engineer owning CI/CD, Kubernetes, and cloud infrastructure for Bounteous, a global AI services firm. Automating secure, reliable platforms across the DevOps lifecycle.
DevOps Engineer owning CI/CD and app releases for a gamified sports training platform. Maintaining React Native, Expo/EAS, Supabase, Next.js, and React delivery workflows.
Site Reliability Engineer managing AWS and Kubernetes reliability for Rentsync’s rental - property software products. Leading incident response, observability, automation, and infrastructure hardening.
Senior DevOps Engineer securing Boeing Canada’s Azure, Kubernetes, and on - premise platforms. Leading CI/CD, infrastructure automation, reliability, compliance, and technical mentorship.
Site Reliability Engineer automating enterprise release orchestration and delivery operations for Sun Life. Supporting platform reliability, Kubernetes automation, and transition to a future release management solution.