664,245open jobs
38,805companies
99,560added this week
Browse all
Salary
$95k – $189k per year (Estimated)
Location
In office (Needham)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Bitdeer Technologies Group is a global technology company specializing in Bitcoin mining, proprietary ASIC hardware manufacturing, and high-performance computing infrastructure. Headquartered in Singapore, the firm operates data center facilities across North America, Europe, and Asia to provide self-mining, cloud hash rate sharing, and colocation hosting services. By expanding into GPU-accelerated cloud platform capabilities, it delivers scalable computing solutions for both cryptocurrency networks and artificial intelligence workloads.

Bitdeer is a world-leading technology company for AI and Bitcoin mining infrastructure. Bitdeer is committed to providing comprehensive Bitcoin mining solutions for its customers and building AI computational infrastructure to support the AI revolution. Bitdeer handles complex processes involved in computing such as equipment procurement, transport logistics, data center design and construction, equipment management, and daily operations. Bitdeer also offers advanced cloud capabilities to customers with high demand for artificial intelligence.

Headquartered in Singapore, Bitdeer has deployed data centers across multiple countries, including the United States, Norway, Bhutan, and Ethiopia.

To learn more, visit https://ir.bitdeer.com/

Key Responsibilities

  • Lead and manage the daily operations of the Data Center site, ensuring the availability, reliability, and operational excellence of all infrastructure and systems.
  • Supervise and manage a team of Operations Engineers, including manpower planning, shift scheduling, task assignment, performance management, coaching, and professional development.
  • Ensure 24x7 operational coverage and maintain adequate staffing to support business and customer requirements.
  • Act as the primary escalation point for operational incidents and coordinate cross-functional teams to drive timely issue resolution and root cause analysis.
  • Oversee the operation, maintenance, and troubleshooting of AI/HPC infrastructure, including:
    • NVIDIA B300 Clusters
    • GPU Servers
    • x86 Servers
    • Storage Servers
    • Ethernet and InfiniBand Switches
    • Optical fiber, DAC, AOC, and associated cabling systems
  • Establish, maintain, and continuously improve operational procedures, Standard Operating Procedures (SOPs), Emergency Operating Procedures (EOPs), and preventive maintenance programs.
  • Monitor site health, operational KPIs, incident trends, and infrastructure performance to ensure service quality and operational efficiency.
  • Coordinate hardware installation, rack and stack activities, system commissioning, infrastructure expansion, and lifecycle management.
  • Review and approve maintenance activities, change requests, incident reports, and shift handover records.
  • Ensure compliance with company policies, operational standards, safety requirements, and security procedures within the Data Center.
  • Collaborate with engineering, network, facilities, and vendor teams to support new deployments and operational improvement initiatives.
  • Participate in on-call rotation and provide hands-on operational support when necessary, including covering shift duties during manpower shortages, emergencies, or critical incidents.
  • Drive a culture of operational excellence, teamwork, accountability, and continuous improvement within the site operation team.

Qualifications

  • Bachelor's degree or above in Computer Science, Computer Engineering, Electrical Engineering, Electronics Engineering, Information Technology, or related disciplines.
  • Minimum 5 years of experience in Data Center operations, IT infrastructure, or HPC/AI infrastructure management.
  • Minimum 2 years of experience in team leadership or people management.
  • Proven experience managing 24x7 shift operations in a mission-critical environment is preferred.
  • Experience with large-scale AI or HPC clusters is highly desirable.
  • Strong knowledge of Data Center operations and infrastructure management, including:
    • NVIDIA B300 Clusters
    • GPU Servers
    • x86 Servers
    • Storage Systems
    • Ethernet Networking
    • InfiniBand Networking
  • Familiarity with NVIDIA GPU architecture, NVLink, NVSwitch, and AI cluster deployment concepts.
  • Experience with server hardware troubleshooting, firmware management, and hardware lifecycle management.
  • Good understanding of structured cabling systems, including optical fiber, MPO/LC connectors, DAC, and AOC cabling.
  • Familiarity with infrastructure monitoring and management tools.

Linux and System Administration

  • Strong Linux administration and troubleshooting skills, including:
    • System and service management
    • Hardware and performance diagnostics
    • Log analysis and incident investigation
    • Network troubleshooting
    • Basic scripting and automation

Leadership and Management Skills

  • Demonstrated ability to lead, motivate, and develop a team of Operations Engineers.
  • Experience in:
    • Shift scheduling and workforce planning
    • Incident and escalation management
    • Performance management and coaching
    • SOP/EOP development and operational governance
    • Vendor and stakeholder coordination
  • Strong decision-making skills with the ability to manage priorities in a fast-paced and mission-critical environment.

Personal Attributes

  • Willingness to provide hands-on operational support and participate in on-call duties when required.
  • Strong ownership mindset and accountability for site operations.
  • Excellent communication and interpersonal skills.
  • Ability to remain calm and make sound decisions during critical incidents.
  • Highly organized, detail-oriented, and committed to operational excellence and continuous improvement.

Equal Opportunity Employer

Bitdeer is committed to providing equal employment opportunities in accordance with country, state, and local laws. Bitdeer does not discriminate against employees or applicants based on conditions such as race, color, gender identity and/or expression, sexual orientation, marital and/or parental status, religion, political opinion, nationality, ethnic background or social origin, social status, disability, age, indigenous status, and union.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
664,245 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Needham
$106k – $205k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Aurora
Python
AI/ML
NCCL
InfiniBand
Edge AI
DevOps
Terraform
Ansible
Zabbix
Prometheus
Docker
Kubernetes
Configuration Management
HPC
Web3
Bitcoin
Apply
$111k – $215k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Needham
Python
AI/ML
NCCL
InfiniBand
Edge AI
DevOps
Terraform
Ansible
Zabbix
Prometheus
Docker
Kubernetes
Configuration Management
HPC
Web3
Bitcoin
Apply
$91k – $180k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Aurora
AI/ML
InfiniBand
NVLink
DevOps
HPC
Web3
Bitcoin
Apply
$113k – $248k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Needham
AI/ML
InfiniBand
DevOps
Prometheus
SLURM
Kubernetes
Grafana
HPC
Web3
Bitcoin
Apply
$107k – $236k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Aurora
AI/ML
InfiniBand
DevOps
Prometheus
SLURM
Kubernetes
Grafana
HPC
Web3
Bitcoin
Apply
$172k – $297k per year (Estimated) • Remote • Full-Time • 10+ years exp • San Jose
AI/ML
CoreWeave
DevOps
GCP
Azure
AWS
AWS Lambda
SLI/SLO/SLA
Web3
Bitcoin
Apply
$106k – $205k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Aurora
Python
AI/ML
NCCL
InfiniBand
Edge AI
DevOps
Terraform
Ansible
Zabbix
Prometheus
Docker
Kubernetes
Configuration Management
HPC
Web3
Bitcoin
Apply
$111k – $215k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Needham
Python
AI/ML
NCCL
InfiniBand
Edge AI
DevOps
Terraform
Ansible
Zabbix
Prometheus
Docker
Kubernetes
Configuration Management
HPC
Web3
Bitcoin
Apply
$91k – $180k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Aurora
AI/ML
InfiniBand
NVLink
DevOps
HPC
Web3
Bitcoin
Apply
$113k – $248k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Needham
AI/ML
InfiniBand
DevOps
Prometheus
SLURM
Kubernetes
Grafana
HPC
Web3
Bitcoin
Apply
$111k – $215k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Needham
Python
AI/ML
NCCL
InfiniBand
Edge AI
DevOps
Terraform
Ansible
Zabbix
Prometheus
Docker
Kubernetes
Configuration Management
HPC
Web3
Bitcoin
Apply
$113k – $248k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Needham
AI/ML
InfiniBand
DevOps
Prometheus
SLURM
Kubernetes
Grafana
HPC
Web3
Bitcoin
Apply
$75k – $152k per year (Estimated) • In office • Full-Time • 1+ year exp • High School Diploma • Needham
Apply
$74k – $105k per year • Equity • In office • 5+ years exp • Bachelor's Degree • Needham
Design
SolidWorks
Apply
$94k – $150k per year • Equity • In office • 8+ years exp • Bachelor's Degree • Needham
Design
SolidWorks
Apply
See all jobs
This is one of many
664,245 more open roles from verified company boards, updated every day.