Data Warehouse Administrator managing operations of data warehouse platforms at Hyve Solutions. Ensuring high availability and performance for business-critical analytics and reporting workloads.
Responsibilities
Operate and support data warehouse platforms such as Hive, Hadoop, ClickHouse, and Vertica in production environments
Own the stability, availability, and performance of data warehouse systems and supporting infrastructure
Monitor system health, query performance, and resource utilization, and proactively identify potential issues
Troubleshoot production incidents across data pipelines, ETL workflows, Spark jobs, and underlying infrastructure
Perform root cause analysis for system failures, data issues, and performance bottlenecks
Design and enhance monitoring, alerting, and observability for data platforms
Automate repetitive operational tasks to improve efficiency and reduce manual intervention
Manage deployment activities, configuration changes, patching, and system upgrades
Support data workflows including ETL pipelines (e.g., Azkaban), PySpark/SparkSQL processing, and ingestion processes
Support data platform integrations such as Azure Blob storage and Power BI Gateway connectivity
Collaborate with data engineering, analytics teams, and vendors to improve system robustness and scalability
Participate in on-call rotation and support incident response to ensure timely resolution
Develop and maintain operational documentation, SOPs, and runbooks
Requirements
Bachelor’s degree in Computer Science, Information Systems, or equivalent practical experience
3+ years of experience in data warehouse administration, data platform SRE, or production support roles
Proven experience supporting large-scale data platforms in production environments
Experience managing incidents, troubleshooting system failures, and driving resolution
Familiarity with data warehouse architecture and large-scale data processing concepts
Experience working with cross-functional teams including data engineering, platform, analytics teams, and external vendors
Experience supporting business-critical systems with high availability and performance requirements.
Senior Site Reliability Engineer deploying Kubernetes - based AI infrastructure on NVIDIA - certified hardware for Mirantis. Ensuring reliable, secure, scalable cloud operations and customer delivery.
DevOps Engineer building and maintaining cloud infrastructure, automation, and CI/CD pipelines for Calliere's software platform. Operating containers, observability tooling, and production workloads across public clouds.
Senior DevOps Engineer improving Sherweb’s cloud - based IT solution delivery through CI/CD, IaC, GitOps, and AI automation. Supporting secure platforms, operational transitions, and developer self - service.
Senior DevOps / Cloud Infrastructure Engineer needed for hybrid role in North York, ON. Requires 10+ years experience with GCP, AWS, Kubernetes, Terraform, and CI/CD.
Staff Site Reliability Engineer strengthening AWS and Kubernetes resilience for Caseware, a fintech company building audit and accounting software. Driving secure delivery, observability, and incident management.
Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.
Mozilla release engineer optimizing Firefox build, test, and deployment pipelines at global scale. Improving developer experience, maintaining automation, and responding to critical service outages without an on - call rotation.