Bitdeer (NASDAQ: BTDR) logo
Bitdeer (NASDAQ: BTDR)
Posted 38 days agoVerified live 2d ago

GPU Compute & Bare Metal / DPU Engineer

Brief overview

Remote
$145k–$260k/yrStated range
3+ yrsMinimum
2 H-1B approvalsDept. of Labor
Bare-Metal Server OperationsGPU Server OperationsLinuxCUDAFirmware ManagementPXEIPMIRedfishOS ImagingAutomated ProvisioningDPU/SmartNICBare-Metal NetworkingAnsibleTerraformPythonGoIncident Management

About the company

Bitdeer (NASDAQ: BTDR) logo
Bitdeer (NASDAQ: BTDR)bitdeer.com

Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure.

Visa sponsorship history

3 years sponsoring, last filed FY2025

Data powered by U.S. Department of Labor. This does not guarantee sponsorship for this specific role.
2H-1B approved
100%approval rate
1new H-1B hires
$132,000median wage / yr
H-1B Petition ApprovalsVisas USCIS actually granted: the strongest sign the company sponsors.
20231
20251
LCA Certified ApplicationsAn early filing step, not a visa approval: it signals intent, not confirmed sponsorship.
20242
20252
Top sponsored roles
Legal CounselSenior Legal CounselDirector of Government Relations (North America Region)

Job description

Summary

Bitdeer is a technology company building AI computational infrastructure and Bitcoin mining solutions. The GPU Compute & Bare Metal / DPU Engineer will own the end-to-end lifecycle of bare-metal GPU nodes across multiple regions, including provisioning, firmware and DPU management, break-fix, reliability, and decommissioning. The role will also automate node delivery, improve fleet operations, and collaborate with storage, imaging, and networking teams.

Responsibilities

  • Own the full lifecycle of bare-metal GPU nodes: provisioning, delivery / onboarding, in-service operation, break-fix, and decommissioning across multiple regions
  • Build and operate automated, repeatable node-delivery pipelines to eliminate the current delivery backlog and keep pace with fleet growth toward 10,000+ GPUs
  • Manage DPU / SmartNIC and server firmware (BMC / BIOS / NIC / GPU firmware): version baselines, upgrades, and validation
  • Drive fleet reliability: reduce MTTR, lead incident response and root-cause analysis, and improve hardware-health monitoring
  • Participate in a sustainable 7×24 multi-region on-call rotation; build runbooks and tooling that reduce manual toil
  • Partner with Storage / Image and Network teams to streamline the provisioning-to-handoff path
  • Define bring-up, rack, capacity, and acceptance standards for new GPU SKUs and data-center regions

Skills

  • 3+ years (Senior: 6+ years) in large-scale bare-metal / server-fleet operations, HPC, or cloud infrastructure
  • Hands-on experience operating GPU servers at scale (e.g. NVIDIA HGX / DGX-class), including driver / CUDA and firmware management
  • Strong Linux systems skills; experience with PXE / IPMI / Redfish, OS imaging, and automated provisioning
  • Familiarity with DPU / SmartNIC (e.g. NVIDIA BlueField) and bare-metal networking
  • Infrastructure automation skills (Ansible, Terraform, Python / Go)
  • Comfortable owning on-call, incident management, and operational runbooks
  • Multi-region / large-fleet operations experience is a strong plus

Qualifications

Must Haves

  • 3+ years (Senior: 6+ years) in large-scale bare-metal / server-fleet operations, HPC, or cloud infrastructure
  • Hands-on experience operating GPU servers at scale (e.g. NVIDIA HGX / DGX-class), including driver / CUDA and firmware management
  • Strong Linux systems skills; experience with PXE / IPMI / Redfish, OS imaging, and automated provisioning
  • Familiarity with DPU / SmartNIC (e.g. NVIDIA BlueField) and bare-metal networking
  • Infrastructure automation skills (Ansible, Terraform, Python / Go)
  • Comfortable owning on-call, incident management, and operational runbooks

Nice to Haves

  • Multi-region / large-fleet operations experience is a strong plus

Benefits

  • Remote (within locations)

More jobs like this