Hiring.Camp

Hardware Engineer, AI Cluster Commissioning

Firmus Technologies

·

Today

Location
Batam, Riau Islands, Indonesia · Batam, Indonesia
Department
AI Cluster Commissioning
Experience
3+ years
Source
Greenhouse

Description

Firmus Technologies

Firmus Technologies is a global leader pioneering the development and operation of efficient AI infrastructure across Asia Pacific.  

Founded in Australia in 2019, our mission is to create the most efficient AI infrastructure by combining cutting-edge technology with a steadfast commitment to sustainability. 

At Firmus, we are unique in our approach. We design, build, and operate a new class of digital infrastructure – the AI Factory. Through our model-to-grid technology approach, we have pushed the boundaries of multi-generational liquid cooling systems, energy management, AI software orchestration, and construction. For our customers, this approach allows us to make every watt count and deliver low-cost AI tokens globally. 

 

Firmus AI Cloud

Our large-scale GPU cloud platform, Firmus AI Cloud, is purpose-built to deliver energy-efficient AI compute at scale to customers. 

It empowers developers, enterprises, educational institutions, and government users to train and deploy AI models with unmatched efficiency and cost savings. With an ever-growing suite of services and applications, we are committed to delivering a cloud experience that is market-leading, proprietary, and built to scale. 

 

Why Firmus?

As an NVIDIA Cloud and Engineering partner in Asia Pacific, you will gain skills, experience, and exposure across the AI industry and be part of shaping what this industry looks like for decades to come.

We are founder-led, not a big corporate. Decisions happen fast, our leaders are accessible, and there's minimum bureaucracy between you and the work. Ownership comes early. Whatever your role, you will have a direct line to outcomes, helping shape how the business grows as we scale nationally across a long-term, large-scale roadmap. 

Work alongside founders and experts in AI infrastructure, energy systems and next-generation compute.

What we build here has impact beyond the business. Our AI Factories are designed to operate as assets to the energy grid to actively strengthen the communities and regions they operate in rather than drawing from them.

Considering applying? You don't need a perfect background to join our team. If you're driven and curious, there's a path for you. We back our people to grow into new domains and take on challenges beyond their previous experience.

    

ROLE SUMMARY

The Hardware Engineer, AI Cluster Commissioning, is a hands-on team player who will play a crucial role in ensuring the efficient deployment, test, and operation of our HyperCube systems. The ideal candidate will be involved in the commissioning and remediation of the latest AI cluster platforms including compute, network, and storage components, provide high-quality technical support, troubleshoot issues, and work with vendors and partners to resolve them in a timely manner. This role requires hands-on work in demanding data center environments, including equipment installation, hardware replacement, structured cabling, and physical infrastructure remediation activities.

KEY RESPONSIBILITIES

  • Support the deployment, commissioning, and remediation of large-scale AI clusters comprising thousands of GPUs, high-speed networks, and high-performance storage systems.
  • Support site teams during commissioning, testing, and equipment setup.
  • Assist with system installations, modifications, and setup as part of the commissioning team.
  • Assist with network cabling and integration.
  • Diagnose server, GPU, storage, and networking failures.
  • Collect logs and diagnostic information, support technical teams during failure investigations, and assist engineering teams in identifying recurring failure patterns and root causes.
  • Perform component-level replacement and remediation.
  • Execute hardware validation procedures and acceptance tests.
  • Verify firmware, cabling, and system configurations.
  • Stakeholder Communication: Provide timely updates and technical support to project and operations teams, ensuring smooth collaboration.
  • Documentation & Reporting: Keep accurate service records, equipment logs, and submit required documentation promptly.
  • Safety & Compliance: Adhere to safety protocols.
  • Training & Development: Stay current on equipment, processes, and industry standards through ongoing learning.
  • Inventory & Quality Control: Maintain tools, spare parts, and equipment inventory.

SKILLS AND EXPERIENCE

Required

  • Diploma or equivalent, technical certification, trade qualification.
  • At least 3 years of relevant experience in a similar role, including hands-on experience with IT equipment maintenance and repair.
  • Good problem-solving and communication skills. Proven to be able to work effectively under direction.
  • Ability to work independently and as part of a team.
  • Ability to work flexible hours, including evenings and weekends, as required during system commissioning.
  • Must hold an unrestricted driver’s license.

Preferred

  • Familiarity with high-density GPU server and rack environments.
  • Experience with data center networking and structured cabling.
  • Experience performing hardware diagnostics and component replacement.
  • Familiarity with Linux operating systems and command-line troubleshooting.

Physical Requirements

  • Ability to safely lift and move IT equipment.
  • Ability to work in data center environments for extended periods.
  • Comfortable working in confined spaces and around high-density equipment racks.
  • Willingness to undertake domestic and international travel for deployment and commissioning activities.

 

Employment Basis

Full-time

 

 

 

Skills

LinuxCompliance

Similar Jobs

30

Senior Hardware Engineer - AI & Server Compute Design

Synnex·Fremont, California +1·Remote

1d ago

Software Engineer II: AI Agents Hardware Verification

Cadence Design·BELO HORIZONTE 02, Brazil

3d ago

Modem HW Design Engineer – AI Driven Next Gen Modem Hardware Development

Qualcomm·San Diego, CA

3d ago

Lead Software Engineer: AI for Hardware Verification

Cadence Design·SAN JOSE 11, US·Remote, Hybrid

4d ago

Cloud Hardware Dev Engineer, AWS AI/ML UltraServers

Amazon

6d ago

Sr Systems Development Engineer, AWS Hardware Engineering Services, AI UltraServers

Amazon·Onsite

6d ago

Principal Early Supplier Involvement (ESI) Engineer - AI / Hyperscale Hardware

Connect·USA - NJ - Secaucus - Bldg 200P, US +1

1w ago

Applied AI Engineer - Hardware

Etched·San Jose·Onsite

1w ago

Principal Early Supplier Involvement (ESI) Engineer - AI / Hyperscale Hardware

Connect·USA - NJ - Secaucus - Bldg 200P, US +1

1w ago

Field Application Engineer - AI Software & Hardware

Tenstorrent·Belgrade, Serbia +2·Remote, Hybrid

1w ago

Cloud Hardware Development Engineer, Cloud AI/ML server teams

Amazon

2w ago

Sr. Staff Hardware Machine Learning/AI Engineer

Rivian·London, UK

2w ago

Field Application Engineer - AI Software & Hardware

Tenstorrent·Gdańsk, Pomeranian Voivodeship +3

2w ago

Field Application Engineer - AI Software & Hardware

Tenstorrent·Munich, Germany +2

2w ago

AI Hardware Systems Engineer, Annapurna Labs, Trainium Machine Learning Fleet Operations

Amazon

2w ago

Sr Cloud Hardware Dev Engineer, AWS Generative AI & ML Servers

Amazon

2w ago

Sr Cloud Hardware Dev Engineer, AWS Generative AI & ML Servers

Amazon

2w ago

HPC & AI Hardware - Engineering Resolution Engineer

Hewlett Packard Enterprise (HP)·Bristol, Avon·Remote, Hybrid

4w ago

HPC & AI Hardware - Engineering Resolution Engineer

Hewlett Packard Enterprise (HP)·Bristol, Avon·Remote, Hybrid

4w ago

HPC & AI Hardware - Engineering Resolution Engineer

Hewlett Packard Enterprise (HP)·Bristol, Avon·Remote, Hybrid

4w ago

Modem Hardware Design Verification Engineer – AI Driven Next Gen Modem Hardware Development, Staff

Qualcomm·Boxborough, MA

1mo ago

Systems Development Engineer, GPU & AI Accelerator Servers, AWS Hardware Engineering

Amazon

1mo ago

Senior Hardware Development Engineer, Cloud AI/ML Server Team

Amazon

1mo ago

Staff Hardware Engineer, Power Delivery (AI Data Center)

Qualcomm·Raleigh, NC

1mo ago

Sr. Manufacturing Hardware Engineer, AI/ML Server Development

Amazon

2mo ago

Manufacturing hardware engineer, Cloud AI/ML/storage server teams

Amazon

2mo ago

Senior Cloud Hardware Development Engineer, Cloud AI/ML/storage server teams

Amazon

2mo ago

Senior Cloud Hardware Development Engineer, Cloud AI/ML/storage server teams

Amazon

2mo ago

Sr. Electrical Hardware Engineer, AI Satellites (Starmind)

Spacex·Bastrop, TX +1

2mo ago

Electrical Hardware Engineer, AI Satellites (Starmind)

Spacex·Bastrop, TX +1

2mo ago