JobRaahGet matched free

Jobs

Data Center Operations Engineer

Bitdeer Technologies Group · Needham, MA · United States · On-site

Posted Sep 15, 2026

Apply with JobRaah

Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.

Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure. Bitdeer is committed to providing comprehensive Bitcoin mining solutions for its customers and building AI computational infrastructure to support the AI revolution. Bitdeer handles complex processes involved in computing such as equipment procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer also offers advanced cloud capabilities to customers with high demand for artificial intelligence. Headquartered in Singapore, Bitdeer has deployed data centers across multiple countries, including the United States, Norway, Bhutan, and Ethiopia. To learn more, visit https://ir.bitdeer.com/ Key Responsibilities Responsible for the daily operation and maintenance of the Data Center infrastructure to ensure high availability and stable service operation. Perform installation, rack and stack, cabling, commissioning, maintenance, and troubleshooting of AI/HPC cluster infrastructure, including: NVIDIA B300 Cluster GPU Servers x86 Servers Storage Servers Ethernet and InfiniBand Switches DAC, AOC, Optical Fiber, and related cabling infrastructure Monitor and maintain the health status of cluster systems, including servers, GPUs, storage, networking devices, and associated infrastructure. Conduct hardware replacement and maintenance activities, including FRU replacement, BIOS/BMC/Firmware upgrades, and hardware diagnostics. Support server provisioning, operating system installation, cluster expansion, network validation, and burn-in testing. Troubleshoot hardware and infrastructure issues, including server failures, GPU errors, storage issues, network connectivity problems, switch failures, and cabling faults. Perform routine inspections, preventive maintenance, and maintain accurate operational records and maintenance logs. Execute incident response procedures and provide timely escalation and resolution according to operational standards. Prepare shift handover reports and maintain operation documents, SOPs, and incident reports. Work closely with engineering, network, and infrastructure teams to support new deployments and ongoing operation improvements. Participate in a three-shift rotation schedule, including night shifts, weekends, and holidays as required. Qualifications Education Bachelor's degree or above in Computer Science, Computer Engineering, Electrical Engineering, Electronics Engineering, Information Technology, or related disciplines. Technical Skills Basic understanding of Data Center infrastructure and server hardware architecture. Familiarity with one or more of the following systems: NVIDIA B300 Cluster GPU Servers x86 Servers Storage Servers Ethernet and InfiniBand Networks Knowledge of server hardware components, including CPU, memory, storage, GPU, BMC/IPMI, and firmware management. Familiarity with network concepts, including TCP/IP,…