
201 - 500 employees
💼 Consulting
📦 Logistics
🏗️ Construction
💰 Post-IPO Equity on 2023-05
Consulting • Logistics • Construction
Bitdeer Group is a leader in the blockchain and high-performance computing industry. It is one of the world’s largest holders and suppliers of hash rate, offering specialized mining infrastructure and high-quality hash rate sharing products. Founded by cryptocurrency pioneer Jihan Wu and led by CEO Matt Linghui Kong, the company is headquartered in Singapore with mining datacenters in the United States, Norway, and Bhutan. Bitdeer is committed to providing comprehensive computing solutions, including cloud services and AI capabilities, while emphasizing dedication, authenticity, and trustworthiness in its mission to be the most reliable provider in the industry.
🔥 0 minutes ago
Improve your chances of getting an interview by checking your resume score before you apply.

201 - 500 employees
💼 Consulting
📦 Logistics
🏗️ Construction
💰 Post-IPO Equity on 2023-05
Consulting • Logistics • Construction
Bitdeer Group is a leader in the blockchain and high-performance computing industry. It is one of the world’s largest holders and suppliers of hash rate, offering specialized mining infrastructure and high-quality hash rate sharing products. Founded by cryptocurrency pioneer Jihan Wu and led by CEO Matt Linghui Kong, the company is headquartered in Singapore with mining datacenters in the United States, Norway, and Bhutan. Bitdeer is committed to providing comprehensive computing solutions, including cloud services and AI capabilities, while emphasizing dedication, authenticity, and trustworthiness in its mission to be the most reliable provider in the industry.
• Architect and scale high-cardinality telemetry infrastructure using highly available time-series databases such as VictoriaMetrics, Thanos, or Mimir • Integrate hardware-level exporters, including NVIDIA DCGM, network switch telemetry, and IPMI/Redfish, into the Kubernetes observability stack • Build eBPF-based diagnostic tools to trace network congestion, kernel-level I/O latency, and distributed training bottlenecks • Develop automated dashboards and alerting pipelines that proactively cordon degraded hardware before it impacts customer training jobs • Design metric pipelines for accurate, multi-tenant consumption billing based on real-time GPU and network utilization metrics • Collaborate with GPU Systems and Scheduling teams to create observability standards for AI-native workloads • Lead technical design reviews for observability architecture • Mentor team members on high-performance telemetry collection and analysis best practices
• Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, or a related field • 6+ years of software or site reliability engineering experience • Deep, hands-on expertise in the Prometheus/OpenTelemetry ecosystem • Advanced proficiency in Go • Extensive experience writing custom Kubernetes metric exporters and operators • Hands-on experience with kernel-level tracing tools, including eBPF and BCC • Deep performance tuning experience with Linux systems • Strong familiarity with AI hardware metrics, including GPU power states, SM utilization, and memory bandwidth • Strong familiarity with high-performance network telemetry • Proven track record of operating, debugging, and scaling large-scale telemetry stacks in high-performance computing or cloud environments • Strong technical leadership skills and ability to influence architectural decisions and align cross-functional teams around observability standards • Excellent communication skills and ability to translate complex system requirements into manageable engineering milestones • Experience working in high-velocity, high-growth engineering environments is strongly preferred
• Equal employment opportunities in accordance with country, state, and local laws • Non-discrimination protections based on race, color, gender identity and/or expression, sexual orientation, marital and/or parental status, religion, political opinion, nationality, ethnic background or social origin, social status, disability, age, indigenous status, and union
Apply Now🔥 21 hours ago
AI transformation leader making Net at Work, a technology consultancy, AI-native. Redesigning workflows, deploying virtual workers, and delivering measurable operating leverage.
🔥 21 hours ago
AI services director building Net at Work’s AI and data consulting practice for SMBs. Leading client AI deployments, teams, service offerings, and P&L growth.
🕒 Yesterday
Corporate AI Engineer building Laurel’s internal AI software factory and infrastructure. Enabling professional services firms through Laurel’s AI Time platform and enterprise AI governance.
🇺🇸 United States – Remote
💵 $191k - $285k / year
⏰ Full Time
🔴 Lead
🤖 Artificial Intelligence
🦅 H1B Visa Sponsor
🕒 2 days ago
AI Operations Manager ensuring reliable, observable, and secure traditional ML, GenAI, and agentic systems. Advancing ethical AI operations for Xsolis’s healthcare technology platform.
🇺🇸 United States – Remote
💰 $75M Private Equity Round - Xsolis on 2021-06
⏰ Full Time
🟠 Senior
🔴 Lead
🤖 Artificial Intelligence
🕒 2 days ago
Executive Director shaping frontier AI strategy and scalable capabilities at Amgen, a biotechnology company developing innovative human therapeutics. Leading advanced AI research, partnerships, governance, and high-performing technical teams.
🇺🇸 United States – Remote
💵 $303.5k - $410.7k / year
💰 $28.5G Post-IPO Debt on 2022-12
⏰ Full Time
🔴 Lead
🤖 Artificial Intelligence
🦅 H1B Visa Sponsor