Langsung ke konten utama
Kembali ke Lowongan

Compute Engineer, Deployment

Bangun infrastruktur AI skala sipil untuk Fluidstack

Seorang Compute Engineer Deployment di Fluidstack bertanggung jawab untuk mengelola proses turn-up compute dari ketersediaan fasilitas hingga siap layanan. Anda akan mengkualifikasi rack secara massal, menetapkan baseline firmware, dan memvalidasi kluster di berbagai situs. Kerja ini dilakukan dalam tim yang berfokus pada kecepatan dan skala, menggunakan pendekatan first principles.

Kenapa Menarik?

Bergabunglah untuk membangun infrastruktur AI yang mempengaruhi kebebasan manusia.

Tanggung Jawab Utama

  • Mengelola proses turn-up compute dari ketersediaan fasilitas hingga siap layanan
  • Mengkualifikasi rack secara massal dengan menetapkan baseline firmware
  • Memvalidasi kluster di berbagai situs dengan platform GPU dan akselerator khusus
  • Menggunakan pendekatan first principles untuk memecahkan masalah

Persyaratan

  • Memiliki pengalaman dalam pengelolaan infrastruktur compute
  • Mampu bekerja dengan autonomi penuh dan mengambil tanggung jawab end-to-end
  • Memiliki kemampuan untuk bekerja dengan kecepatan dan skala tinggi
  • Memiliki pengalaman dalam pengembangan perangkat lunak untuk otomatisasi

Skills Wajib

compute infrastructuredata center managementfirmwareautomationfirst principles
Lihat Deskripsi Asli dari Ashby Job Boards

Deskripsi asli dari Ashby Job Boards

ABOUT FLUIDSTACK We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it. We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI. We hire people who care deeply about this problem space. If that is you, please apply! HOW WE OPERATE - Extreme ownership. Full autonomy. Own things end to end often taking on scope outside your core role without being asked to get things done. - Velocity. We drive everything forward as fast as possible. - First principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins. - Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward. THE INFRASTRUCTURE TEAM Examples of key problems the team is working on - Bring gigawatts of accelerators from first power-on to production. Facility availability to ready-for-service across thousands of racks per site, with a new data hall landing every few weeks. - Make rack qualification faster than the fleet grows. Firmware baselines, burn-in, and cluster validation proven on every rack before a customer workload touches it, at a pace that never becomes the critical path. - Scale by tooling, not headcount. Deployed megawatts grow severalfold next year while the team stays near-flat, because anything done twice by hand becomes software. ROLE SCOPE - Own compute turn-up from facility availability to ready-for-service: the stretch after the network hands off and before customers run workloads. - Qualify racks at scale: establish firmware baselines, configure BMC and BIOS, run burn-in, and validate at node and cluster level across hundreds of racks per site on GPU and custom accelerator platforms. - Drive qualification through the base-management Kubernetes platform and provisioning stack (discovery, imaging, firmware updates, shared services), burning down qual queues with tooling rather than manual runs. - Triage hardware failures found in qualification: isolate to component, drive RMA and vendor escalation, and feed failure patterns back into the qual gates. - Run turn-up remotely by default, with on-site pulses of roughly a week per data hall as new halls reach facility availability, plus occasional overlapping-site weeks. - Partner with network deployment, ICT, data center operations, and hardware teams during turn-up windows, and support incident response on freshly-live capacity. - Ability to travel 30-40% of the time to our Data Centers and Labs, as needed. WHAT WE'RE LOOKING FOR The below is a starting point. We always make space for exceptional people, so if you don't fit this role exactly, tell us where you would https://jobs.ashbyhq.com/fluidstack/05c2e69c-42f9-4fcb-9cf0-a467aaf98f1c. - You've brought up server or GPU fleets at scale, hundreds of nodes or more, and taken them all the way to production. - You work deep in Linux and out-of-band management: BMC, IPMI, and Redfish are daily tools for you, not occasional lookups. - You've automated hardware workflows in Python or Go rather than clicking through them, and the second time you do anything by hand you turn it into software. - You've worked physically in data halls, racking, cabling, and swapping components, and you're just as effective acting as remote hands or directing them. - You triage failures methodically across hardware, firmware, and software, isolating the fault to a component before reaching for a fix. - You travel for turn-up windows when a new data hall comes online. - Bonus: Kubernetes-based bare-metal provisioning. Accelerator platform bringup (NVIDIA, AMD, or custom). Burn-in and stress harness design. DCIM and inventory tooling. We are committed to pay equity and transparency. Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law. You will receive a confirmation email once your application has successfully been accepted. If there is an error with your submission and you did not receive a confirmation email, please email careers@fluidstack.io with your resume/CV, the role you've applied for, and the date you submitted your application-- someone from our recruiting team will be in touch.

Konteks Gaji

Posisi Engineering serupa di LokerDollar dibayar sekitar $170k/yr (kisaran $855–1000k/yr, dari 877 listing aktif).

Perekrutan di Fluidstack

Fluidstack punya 16 lowongan aktif lain di LokerDollar dan telah merekrut di sini sejak 10 Jun 2026 — di kategori Engineering.

Lihat semua lowongan Fluidstack →

Pemberi kerja tidak menyatakan keterbukaan lokasi — cek langsung lowongannya

Perusahaan
Fluidstack
Gaji
$164k–206k/yr
Lihat selisih gaji remote (USD) vs lokal →
Tipe Lowongan
full time
Lokasi
Remote
Kategori
Level
mid
DipostingCek ulang sumbernya
11 Agu 2026

Bagikan lowongan ini

Bantu temanmu nemu kerja remote berikutnya.

Pertanyaan yang sering diajukan

Apakah Compute Engineer, Deployment di Fluidstack bisa dikerjakan remote?
Posisi ini berlokasi di Remote. Detail remote/onsite ada di deskripsi lowongan.
Berapa gaji untuk Compute Engineer, Deployment di Fluidstack?
Rentang gaji yang tercantum untuk posisi ini adalah $164k–206k/yr.
Jenis pekerjaan apa Compute Engineer, Deployment di Fluidstack?
Posisi ini adalah pekerjaan full time.
Bagaimana cara melamar?
Klik tombol "Lamar" pada halaman ini untuk menuju halaman aplikasi resmi Fluidstack.

Jelajahi lebih lanjut

Data & laporan pasar

Riset gaji & permintaan skill dari data lowongan kami sendiri.

Dari blog kami