Skip to main content
Back to Jobs

Platform Engineer

Build multi-cloud platform infrastructure for AI agent workloads from scratch

Design and build the cloud compute layer for AI agents as first-class users, including multi-cloud substrate, compute fleet management, networking, and tooling. Work with physical infrastructure and GPUs to support bursty, ephemeral, untrusted workloads with millisecond latency and hard tenant isolation. Collaborate directly with founders on a greenfield project to create production infrastructure that is deployable, diagnosable, and recoverable.

Why This Role?

Collaborate directly with the founding team on a foundational, greenfield project

Key Responsibilities

  • Design and implement multi-cloud networking, IAM, DNS/TLS, and compute fleets across AWS, GCP, and Azure
  • Provision and manage physical infrastructure and GPU resources for AI workloads
  • Build and maintain IaC using Terraform/Pulumi, CI/CD pipelines, and release processes
  • Implement observability including metrics, logs, traces, alerts, dashboards, and SLO tracking
  • Manage incident response, on-call rotation, postmortems, runbooks, capacity planning, and cost optimization

Requirements

  • Ability to write code and build tools/automation that other engineers depend on
  • Experience working across multiple cloud providers with understanding of networking, IAM, and compute differences
  • Experience debugging cloud networking, DNS/TLS, IAM, and host-level failures without a separate ops team
  • Experience with physical infrastructure, bare-metal, or GPU environments beyond managed cloud services
  • Focus on reliability as a product feature and clear communication during incidents

Required Skills

Terraform/PulumiMulti-cloud (AWS, GCP, Azure)ObservabilityPhysical InfrastructureGPU Provisioningcloud infrastructuremulti-cloudIaCincident response

Keywords

platform engineermulti-cloudAI infrastructureGPU computingTerraformSRE
View Original Description from WeWorkRemotely

Original description from WeWorkRemotely

Headquarters: URL: https://catamaran-research.ai/ About Us We are building the cloud compute layer for the agent age: think AWS or GCP, but designed around AI agents as first-class cloud users. Existing cloud primitives were designed for humans clicking consoles or writing static IaC, not for autonomous software that needs to spin up environments, run untrusted code, and manage its own resources at runtime. We are venture-backed and led by three technical co-founders with backgrounds in high-frequency trading, ML engineering, and quantitative research. The Role This is a foundational, greenfield project. You will design and build the platform infrastructure from scratch, collaborating directly with the founding team. You will build the infrastructure that other infrastructure runs on: the multi-cloud substrate, compute fleet management, networking, and the tooling that lets the team ship safely. The workloads are bursty, ephemeral, and untrusted, with millisecond-latency requirements and hard isolation between tenants. You will work with physical infrastructure and GPUs, not just managed cloud services. The goal is production that is boring in the useful sense: deployable, diagnosable, recoverable. What You'll Own Multi-cloud platform infrastructure (AWS, GCP, Azure): networking, IAM, DNS/TLS, and compute fleets. Physical infrastructure and GPU provisioning. IaC (Terraform/Pulumi), CI/CD pipelines, rollback workflows, and release processes. Observability: metrics, logs, traces, alerts, and dashboards. SLOs, incident response, on-call rotation, postmortems, and runbooks. Capacity planning, cost modelling, and cloud spend optimization. You Are A Fit If You can write code, not just glue scripts. You build tools and automation that other engineers depend on. You have worked across multiple cloud providers and understand their networking, IAM, and compute differences. You can debug cloud networking, DNS/TLS, IAM, and host-level failures without a separate ops team. You have experience with physical infrastructure, bare-metal, or GPU environments, not just managed cloud services. You care about reliability as a product feature, not just infrastructure hygiene. You communicate clearly and keep the team informed. When something breaks, people hear it from you first. You stay constructive when requirements change. Uncertainty doesn't block you. You think about the business, not just the code. You spot opportunities, seek out the actual workflow, and prototype solutions within constraints. You're excited to work closely with founders and shape a product from the ground up. Useful Experience Building platform tooling or internal developer platforms. Bare-metal provisioning, GPU fleet management, or physical infrastructure at scale. Infrastructure for multi-tenant isolation workloads (VMs, containers, or sandboxes). Low-latency or high-throughput infrastructure where milliseconds matter. Cost modelling and unit-economics visibility for cloud compute. Cross-cloud networking: VPCs, load balancers, DNS, CDN, and connectivity between providers. Compensation And Logistics Competitive salary and early-stage equity. Singapore preferred; fully remote with meaningful Singapore-time overlap. Full-time. Contact If this excites you, we'd love to hear from you at careers@catamaran-research.ai To apply: https://weworkremotely.com/remote-jobs/catamaran-research-platform-engineer

Openness not stated by employer — check the listing

Source
WeWorkRemotely
Salary
Job Type
full time
Location
Remote · Open worldwide
Category
Seniority
senior
Posted
May 13, 2026

Share this job

Help a friend find their next remote role.

Explore related

Market data & reports

Salary & skill-demand research built from our own listings data.