Staff Software Engineer - Compute
Bangun kontrol plane untuk sistem komputasi GPU dan CPU di Lambda
Sebagai Staff Software Engineer di Lambda, Anda akan mendefinisikan visi teknis untuk lifecycle dan kontrol plane komputasi di Lambda. Anda akan bekerja pada sistem yang memungkinkan provisioning dan manajemen lifecycle platform komputasi heterogen secara massal. Anda akan memberikan kepemimpinan teknis dan mentoring untuk tim yang bekerja pada infrastruktur cloud untuk peneliti AI terkemuka.
Kenapa Menarik?
Anda akan bekerja pada infrastruktur cloud AI yang mempengaruhi peneliti dan perusahaan besar.
Tanggung Jawab Utama
- Membangun dan mengoptimalkan sistem komputasi GPU-first
- Mendefinisikan kontrol plane lifecycle untuk GPU dan CPU
- Mengarahkan keputusan teknis terkait arsitektur semiconductor dan BIOS/firmware
- Membangun model keamanan multi-tenant untuk platform komputasi
- Mengarahkan tim untuk mengembangkan infrastruktur yang dapat diukur
Persyaratan
- Pengalaman luas dalam infrastruktur cloud
- Paham tentang BIOS/firmware, Linux kernel, dan DPU
- Kemampuan memimpin tim dan mentoring engineer senior
- Paham tentang sistem terdistribusi dan operasi CSP
- Kemampuan mengubah kebutuhan pelanggan menjadi infrastruktur yang dapat diukur
Skills Wajib
Konteks Indonesia
- Overlap Jam Kerja:
- Fleksibel — atur jam kerjamu sendiri
Lihat Deskripsi Asli dari Ashby Job Boards
Deskripsi asli dari Ashby Job Boards
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our Bellevue, San Francisco, or San Jose office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. About the Role As a Staff Software Engineer for the Compute pillar, you will play a critical role in defining the technical vision for Lambda's next-generation GPU and CPU host instance lifecycle and compute control plane. This role bridges the gap between high-level distributed systems and low-level semiconductor architecture to enable seamless, reliable cloud provisioning and lifecycle management of a heterogeneous compute platform at a massive scale. You will provide hands-on technical leadership that will guide development of a resilient compute control plane utilizing durable execution concepts and deep/unique hardware integration. The position requires a deep understanding of the entire stack, from BIOS/firmware (UEFI), Linux kernel internals, modern DPU capabilities, distributed systems, cradle-to-grave system lifecycle management, to large-scale cloud-service provider (CSP) operations. You will drive high-impact, cross-functional initiatives, leading the work of multiple engineers to deliver enterprise-grade SLAs for the world's leading AI researchers. What You'll Do We are seeking an engineer with extensive experience in cloud infrastructure to build and optimize GPU-first compute systems. In this role, you will be responsible for: - Designing and implementing a highly available and reliable GPU and CPU “host and instance lifecycle” control plane. - Guide technical decisions involving semiconductor architecture, BIOS/Firmware settings, system boot methodologies, and DPU utilization to optimize host capabilities, performance and reliability. - Guide design of compute platform multi-tenant security model - Provide technical leadership and mentorship for senior engineers across several teams to execute on complex infrastructure roadmaps and technical strategy. - Collaborate with product and data center organizations to translate customer requirements into scalable infrastructure capabilities. - Work with customers on translating vague customer technical requirements into concrete engineering deliverables. - Set engineering standards and lead design reviews for mission-critical cloud software at scale. Who You are - 10+ years of experience working on compute control plane distributed systems used for deploying and lifecycle managing heterogeneous compute platforms into data-centers, built for resilience at scale. - Deep expertise in durable execution models and distributed systems used in cloud-service provisioning. - Basic knowledge of software defined networking fundamentals that informs secure, multi-tenant distributed systems. - Proven track record of leading large-scale semi-conductor hardware enablement and deployment initiatives. - Proven experience in deploying net-new data-centers into a global compute platform (not just working in existing data-centers). - Proficiency in one of more of the following programming languages: C/C++, Rust, Python, Go. Nice to Have - Knowledge of Nvidia’s AI Factory architectural components (including GPU hosts, CPU hosts, SuperNICs (ConnectX and Bluefield DPUs , and switches). - Knowledge of Nvidia’s AI Factory software offerings (like DOCA, DOCA SNAP, CUDA, et al.) - Knowledge of Linux kernel internals, device drivers, and virtualization technologies (KVM, QEMU), kernel bypass technologies (like SR-IOV, DPDK, SPDK). - Experience with Cloud Service Provider Kubernetes offerings. - Knowledge of high-performance networking (InfiniBand, RoCE) and storage protocols (NVMe-oF). Salary Range Information The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description. About Lambda - Founded in 2012, with 500+ employees, and growing fast - Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove - We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG - Our values are publicly available: https://lambda.ai/careers - We offer generous cash & equity compensation - Health, dental, and vision coverage for you and your dependents - Wellness and commuter stipends for select roles - 401k Plan with 2% company match (USA employees) - Flexible paid time off plan that we all actually use Equal Opportunity Employer Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
Konteks Gaji
Posisi Engineering serupa di LokerDollar dibayar sekitar $205.1k/yr (kisaran $19.575k–2846.004k/yr, dari 468 listing aktif).
Perekrutan di Lambda
Lambda punya 8 lowongan aktif lain di LokerDollar dan telah merekrut di sini sejak 23 Jun 2026 — di kategori Engineering.
- Senior Site Reliability Engineer - Fleet
- Senior Software Engineer - Managed Kubernetes
- Senior Site Reliability Engineer - Managed Kubernetes
Pemberi kerja tidak menyatakan keterbukaan lokasi — cek langsung lowongannya
Pertanyaan yang sering diajukan
- Apakah Staff Software Engineer - Compute di Lambda bisa dikerjakan remote?
- Posisi ini berlokasi di Remote. Detail remote/onsite ada di deskripsi lowongan.
- Berapa gaji untuk Staff Software Engineer - Compute di Lambda?
- Rentang gaji yang tercantum untuk posisi ini adalah $314k–465k/yr.
- Jenis pekerjaan apa Staff Software Engineer - Compute di Lambda?
- Posisi ini adalah pekerjaan full time.
- Bagaimana cara melamar?
- Klik tombol "Lamar" pada halaman ini untuk menuju halaman aplikasi resmi Lambda.
Jelajahi lebih lanjut
Data & laporan pasar
Riset gaji & permintaan skill dari data lowongan kami sendiri.
- Lowongan IT Indonesia vs Remote Global (2026)Analisis data primer 2.049 lowongan: metodologi, klasifikasi, dataset bisa diunduh.
- Permintaan Skill AI: Indonesia vs Global (2026)10.000+ lowongan, classifier taxonomy-first, Wilson CI, pra-registrasi sebelum analisis.
- Remote ≠ Remote: Skill yang Membuka Kerja Global untuk Indonesia (2026)12.891 lowongan remote: skill coding bergaji tertinggi justru paling terkunci untuk pelamar Indonesia. Dataset agregat CC BY 4.0.
- Laporan Hiring Indonesia: Tech vs Non-TechPermintaan lowongan per bidang dari hitungan agregat — bukan listing per-listing.
- Benchmark Gaji IndonesiaKisaran gaji agregat lintas peran, dengan metodologi dan dataset terbuka.
- Indeks Gaji & Permintaan Kerja Remote untuk IndonesiaBerapa banyak lowongan remote global yang terbuka untuk Indonesia, dan gajinya (USD) per bidang.
- Laporan Kuartalan Pasar Kerja IndonesiaPHK, pendanaan, gaji & skill per kuartal — agregat terbuka.
- Laporan Pasar Remote per PeranLaporan otomatis per kelompok peran — skill, senioritas, perusahaan, gaji.
- Benchmark Gaji Remote GlobalGaji tahunan per bidang & mata uang, plus porsi lowongan terbuka untuk seluruh dunia.
Dari blog kami
- Lowongan Radiologi Remote: Update Ags 2026Analisis lowongan radiologi remote terbaru di Agustus 2026: gaji, tren, dan tips apply. Peluang kerja dokter spesialis radiologi remote.
- 3 Secret WebsitesTahu 3 situs rahasia untuk menghasilkan dolar dengan bekerja online.
- Funding Turun 43%, Malah Buka Lowongan?Pendanaan startup Indonesia turun 43% di H1 2026. Tapi perusahaan global justru buka lowongan remote untuk talenta Indonesia.