Full-Time

Site Reliability Engineer

Data warehouse administration

Hyve Solutions

Hyve Solutions

1,001-5,000 employees

Designs and deploys hyperscale datacenter infrastructure

Compensation Overview

CA$60k - CA$95k/yr

Ontario, Canada + 2 more

More locations: Greenville, SC, USA | Clearwater, FL, USA

In Person

On-site in listed locations; no fully remote option.

Category
DevOps & Infrastructure (1)
Data & Analytics (1)
Required Skills
Bash
Python
Airflow
Apache Spark
SQL
ETL
ClickHouse
Observability
Hadoop
Linux/Unix

Get referred to Hyve Solutions

See people who can refer or advise you

Requirements
  • Strong SQL skills with experience working on large-scale data systems
  • Hands-on experience with data warehouse platforms (Hive, ClickHouse, Vertica, or similar)
  • Experience supporting production data platforms with SLA/SLO awareness (required)
  • Strong understanding of distributed systems, preferably within the Hadoop ecosystem
  • Experience supporting ETL pipelines, data ingestion, and workflow orchestration tools (e.g., Azkaban, Airflow)
  • Solid troubleshooting skills across data pipelines, query performance, and system/infrastructure layers
  • Experience with monitoring, alerting, and observability tools
  • Working knowledge of Linux systems in production environments
  • Experience with scripting (Shell, Python, or similar) for automation
  • Familiarity with deployment processes, environment management, patching, and upgrade activities
  • Experience with Spark, PySpark, or SparkSQL troubleshooting
  • Familiarity with cloud storage (e.g., Azure Blob) and BI integrations (e.g., Power BI Gateway) is a plus
  • Bachelor’s degree in Computer Science, Information Systems, or equivalent practical experience
  • 3+ years of experience in data warehouse administration, data platform SRE, or production support roles
  • Proven experience supporting large-scale data platforms in production environments
  • Experience managing incidents, troubleshooting system failures, and driving resolution
  • Familiarity with data warehouse architecture and large-scale data processing concepts
  • Experience working with cross-functional teams including data engineering, platform, analytics teams, and external vendors
  • Experience supporting business-critical systems with high availability and performance requirements
Responsibilities
  • Operate and support data warehouse platforms such as Hive, Hadoop, ClickHouse, and Vertica in production environments
  • Own the stability, availability, and performance of data warehouse systems and supporting infrastructure
  • Monitor system health, query performance, and resource utilization, and proactively identify potential issues
  • Troubleshoot production incidents across data pipelines, ETL workflows, Spark jobs, and underlying infrastructure
  • Perform root cause analysis for system failures, data issues, and performance bottlenecks
  • Design and enhance monitoring, alerting, and observability for data platforms
  • Automate repetitive operational tasks to improve efficiency and reduce manual intervention
  • Manage deployment activities, configuration changes, patching, and system upgrades
  • Support data workflows including ETL pipelines (e.g., Azkaban), PySpark/SparkSQL processing, and ingestion processes
  • Support data platform integrations such as Azure Blob storage and Power BI Gateway connectivity
  • Collaborate with data engineering, analytics teams, and vendors to improve system robustness and scalability
  • Participate in on-call rotation and support incident response to ensure timely resolution
  • Develop and maintain operational documentation, SOPs, and runbooks

Hyve Solutions designs and deploys hyperscale digital infrastructures. It collaborates with customers to create and deliver server, storage, and networking solutions tailored for data centers worldwide, guiding projects from initial design to full implementation. The company relies on deep industry experience and strong vendor partnerships to build purpose-built systems that meet current and future data-center needs. As a wholly owned subsidiary of TD SYNNEX, it leverages scale and procurement strength to support large deployments. Compared with competitors, Hyve differentiates itself through end-to-end, design-to-deployment capabilities, a focused hyperscale approach, and a broad partner ecosystem that enables customized, scalable infrastructure.

Company Size

1,001-5,000

Company Stage

N/A

Total Funding

N/A

Headquarters

Fremont, California

Founded

1980

Get referred to Hyve Solutions

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • AI training and inference demand favors Hyve's rack-scale platforms.
  • US manufacturing expansion supports hyperscaler ramps and shorter lead times.
  • NVIDIA GTC visibility and partner collaboration improve pipeline generation.

What critics are saying

  • Revenue depends heavily on hyperscaler AI capex cycles.
  • NVIDIA-centered strategy weakens if customers adopt ASICs or rival accelerators.
  • Expanded capacity becomes a cost burden if program ramps slip.

What makes Hyve Solutions unique

  • Rack-scale AI systems integrate servers, networking, cooling, and deployment.
  • Orion uses NVIDIA HGX for plug-and-play, high-density AI clusters.
  • Design-to-deployment services bundle engineering, manufacturing, testing, and rollout.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Health Insurance

Dental Insurance

Vision Insurance

401(k) Company Match

Paid Vacation

Paid Sick Leave

Paid Holidays

Profit Sharing

Employee Stock Purchase Plan

Tuition Reimbursement

Wellness Program

Company News

Commercial Times
Mar 18th, 2026
Lightmatter forms consortium to seize CPO discourse power.

Lightmatter forms consortium to seize CPO discourse power. Connects major manufacturers like Celestica, Dell, and Quanta to establish interoperability standards needed for next-generation AI systems. * 2026.03.18 * 03:00 * Industrial and Commercial Times, Li Juanping Silicon photonics startup Lightmatter, in collaboration with partners including Celestica, Corning, Dell, Flex, Foxconn Interconnect Technology, Hyve Solutions, Keysight, Qualcomm, and Quanta Cloud Technology, has launched the Co-Packaged Optics (CPO) Reference Architecture Initiative within the Open Compute Project (OCP). This initiative aims to promote open specifications and interoperability standards required for next-generation AI systems. With the rapid expansion of generative AI models, data centers face bottlenecks in bandwidth, power consumption, and spatial density in internal electrical interconnections, making CPO a crucial technological direction for next-generation AI infrastructure. Compared to traditional pluggable optical modules, CPO significantly enhances transmission bandwidth by tightly integrating photonic components into silicon chips and system ends, while also improving power efficiency and rack space utilization. It is regarded as a key solution for hyperscale cloud providers building AI clusters. The industry believes that moving CPO from technical validation to large-scale commercialization still faces several challenges. These include high complexity in component and system integration, insufficient cross-supply-chain interoperability, and incomplete reliability verification and mass production testing standards, all of which are critical factors affecting adoption speed. By leveraging the OCP platform to call for supply chain collaboration, Lightmatter aims to establish shared reference architectures, open standards at the component and system levels, and further build interoperability testing and certification frameworks to accelerate the deployment of CPO solutions. Nick Harris, founder and CEO of Lightmatter, stated that the AI industry is at a critical turning point, and the market requires solutions that can be manufactured, deployed on a large scale, and are interoperable, rather than single closed architectures. Open collaboration through the OCP community will help build a more comprehensive CPO ecosystem. Sameh Boujelbene, Vice President of Dell, also pointed out that the rapid growth of AI workloads has made data center interconnections a major bottleneck. If technological and supply chain gaps can be addressed through open collaboration, it will help expedite the adoption timeline for CPO. Lightmatter's collaboration with partners across servers, networking, fiber optics connectivity, measurement verification, and system manufacturing indicates that CPO development has shifted from isolated technological competition to ecosystem integration. If OCP-related specifications gradually take shape, they will facilitate the transition of hyperscale data centers from traditional pluggable optical modules to integrated optics. This will also bring a new wave of growth momentum to the supply chains of AI servers, silicon photonics, advanced packaging, and high-speed interconnections.

Business Wire
Jan 27th, 2026
Jerry Kagele named president of Hyve Solutions as Steve Ichinaga transitions to advisory role

Hyve Solutions, a TD SYNNEX subsidiary specialising in hyperscale digital infrastructure, has appointed Jerry Kagele as president. He succeeds Steve Ichinaga, who is transitioning to an advisory role after 40 years at TD SYNNEX, including 15 years founding and leading Hyve Solutions. Kagele joined Hyve in 2025 and brings extensive technology industry experience, including senior roles at Western Digital and Sandisk, where he served as chief revenue officer. Ichinaga will remain with the organisation for one year as senior advisor, focusing on customer and partner success. The planned leadership transition aims to position Hyve Solutions for continued growth whilst maintaining operational continuity. Hyve designs and deploys hyperscale data centre infrastructure solutions for customers worldwide.

Newsvidia
Mar 19th, 2025
Hyve Solutions Unveils Comprehensive AI Infrastructure Portfolio at NVIDIA GTC 2025 - Empowering Scalable AI Deployments for Data center, Hybrid Cloud, and Edge Environments

Hyve Solutions unveils comprehensive AI infrastructure portfolio at NVIDIA GTC 2025 - empowering Scalable AI Deployments for Data center, Hybrid Cloud, and Edge Environments.

Sapio Asia
Mar 18th, 2025
Hyve Solutions Unveils Comprehensive AI Infrastructure Portfolio at NVIDIA GTC 2025

Hyve Solutions unveils comprehensive AI infrastructure portfolio at NVIDIA GTC 2025.

HR Today
Dec 16th, 2024
Sharon Baronessa Appointed as Senior Director, Human Resources at Hyve Solutions

Fremont, California, United States, December 2024 - Sharon Baronessa has been named Senior Director, Human Resources at Hyve Solutions, marking a significant milestone in her professional journey.