Site Reliability Engineer responsible for the installation, configuration, maintenance of middleware technologies at Hyve Solutions. Managing applications on container platforms and ensuring reliable operation of critical middleware components.
Responsibilities
Install, configure, and maintain middleware technologies such as Apache NiFi, Redis, Nginx, Hadoop Ecosystem, and related tools.
Deploy and manage applications on container platforms like Kubernetes.
Monitor performance, availability, and health of middleware services.
Perform routine patching, upgrades, backups, and recovery procedures.
Troubleshoot issues related to application connectivity, data flows, caching, and web serving.
Implement basic security configurations and access controls.
Automate repetitive tasks using scripting and configuration tools.
Provide support to development and operations teams.
Strengthen monitoring, alerting, performance degradation management, and emergency intervention for relevant applications.
Requirements
Bachelor's degree in Computer Science, IT, or related field (or equivalent experience).
3–5 years of experience in middleware or application administration.
Hands-on experience with Apache NiFi, Redis, Nginx, and Kubernetes.
Proficiency in Linux/Unix environments and shell scripting (Bash).
Familiarity with containerization and basic cloud platforms (AWS, Azure, GCP).
Understanding of networking, load balancing, TLS/SSL, and security best practices.
Experience with monitoring tools (Prometheus, Grafana, ELK stack).
Understand distributed system architecture and microservice design preferred.
Good problem-solving skills and ability to work in a team.
DevOps Manager overseeing releases, enterprise tooling, and incident response for Delta Controls, a building - automation solutions manufacturer. Establishing standards across global product teams and offices.
DevOps Engineer building AWS infrastructure and automated systems for S&P Global’s financial data and technology solutions. Supporting resilient applications through Terraform, CI/CD, containerization, monitoring, and cloud operations.
Staff SRE leading GCP reliability, observability, and infrastructure automation for Calix’s broadband communications platform. Building resilient GKE, Kafka, database, and networking systems.