Skip to main content
Back to Jobs

Director of Cloud Operations

Lead cloud infrastructure reliability across multi-region AWS platform

As Director of Cloud Operations, you will own the availability, performance, and resilience of Firstup's multi-region AWS platform serving 40 Fortune 100 companies. You will drive improvements in system reliability through SLIs/SLOs, error budgets, and proactive engineering practices while reducing MTTR and enhancing incident response effectiveness. You will guide architecture decisions for microservices, Kubernetes (EKS), and serverless workl...

Why This Role?

Direct impact on systems serving 17 million employees daily at a company trusted by 40 Fortune 100 firms

Key Responsibilities

  • Own availability, performance, and resilience of multi-region AWS platform
  • Drive system reliability improvements using SLIs/SLOs and error budgets
  • Reduce MTTR and improve incident response effectiveness organization-wide
  • Guide architecture decisions for microservices, EKS, and serverless workloads
  • Advance observability strategy using Datadog for infrastructure and applications
  • Establish and refine incident management practices including on-call processes

Requirements

  • Deep technical expertise in AWS cloud infrastructure
  • Experience leading distributed engineering teams across US and UK
  • Strong background in system reliability, observability, and incident management
  • Hands-on leadership approach to improving systems and processes
  • Ability to partner with Engineering, Security, and Product teams

Required Skills

cloud operationsawsmicroserviceskubernetesserverlessobservabilityleadershipIncident ManagementSystem ReliabilityTeam Leadership

Indonesia Context

Working Hours Overlap:
Flexible — work your own hours
See remote (USD) vs local pay →

Keywords

Director Cloud OperationsAWS InfrastructureSite Reliability EngineeringMulti-region ArchitectureObservability StrategyIncident Response
View Original Description from WeWorkRemotely

Original description from WeWorkRemotely

Headquarters: Remote - US URL: http://firstup.io Who We Are At Firstup, our mission is to improve the employee experience at every moment that matters, large and small. As the communication pipeline for the world's workforce, we now serve 40 of the Fortune 100 companies, reaching and connecting more than 17 million employees daily. Our employees are experts in the employee experience, workforce communications and technology. Joining Firstup means joining a movement to make work better for every worker. As the world’s first intelligent communication platform, Firstup meaningfully engages employees at every moment from hire to retire, and delivers engagement insights to help companies support, promote and retain their talent. Our movement has taken root and is evident in our world-class customer base. Now we need your help. Ready to make a difference in the world? Job Summary: We are seeking a Director of Cloud Operations (CloudOps) to lead and evolve our cloud infrastructure and operational practices across a globally distributed SaaS platform. This is a hands-on leadership role responsible for ensuring the reliability, scalability, and efficiency of our systems running across multiple AWS regions in the United States and Europe. As part of the senior leadership team, you will partner closely with Engineering, Security, and Product to strengthen operational excellence, enhance system observability, and drive continuous improvement in how we build and run services. You will lead a distributed team of engineers across the US and UK, fostering a high-performing, collaborative, and growth-oriented environment. This role is ideal for a leader who combines deep technical expertise with a pragmatic approach to improving systems, processes, and team capabilities. What You’ll Do Cloud Platform & Reliability Own the availability, performance, and resilience of our multi-region AWS platform. Drive improvements in system reliability through well-defined SLIs/SLOs , error budgets, and proactive engineering practices. Lead efforts to reduce MTTR and improve incident response effectiveness across the organization. Guide architecture decisions for microservices, Kubernetes (EKS), and serverless workloads to ensure scalability and fault tolerance. Observability & Incident Management Advance our observability strategy using Datadog , ensuring actionable insights across infrastructure and applications. Establish and refine incident management practices, including on-call processes, escalation paths, and post-incident reviews. Act as an incident commander for critical events and contribute to the on-call rotation. Operational Excellence & Efficiency Elevate operational standards through automation, standardization, and adoption of modern best practices. Drive cost optimization initiatives across AWS environments without compromising performance or reliability. Leverage AI and automation to improve operational efficiency, accelerate root cause analysis, and enhance system insights. Continuously improve CI/CD pipelines (CircleCI) and infrastructure-as-code practices (Terraform). Team Leadership & Development Lead, mentor, and support a distributed team of CloudOps engineers across the US and UK. Foster a culture of accountability, learning, and continuous improvement. Provide technical guidance while enabling the team to grow in ownership and capability. Hybrid & Legacy Environment Support Ensure stability and support for existing customers while maintaining clear operational boundaries with the cloud platform. What We’re Looking For Experience 10+ years in cloud infrastructure, SRE, or DevOps roles, with 3+ years experience leading CloudOps/SRE teams . Proven track record of leading operational or platform transformations in a SaaS environment. Experience operating multi-region, customer-facing systems at scale . Technical Expertise Strong hands-on experience with: AWS (multi-region architectures) Kubernetes (EKS) and containerized environments Inf

Salary Context

Similar Engineering roles on LokerDollar pay around $215k/yr (range $19.575k–2745.996k/yr, n=426 active listings).

Track this application + get a follow-up reminder

Free account · no credit card · Log in

Pro $9/mo · unlimited applies + AI resume

Remote-friendly · fits your timezone
Company
Firstup
Source
WeWorkRemotely
Job Type
full time
Location
Remote · Open worldwide
Category
Seniority
lead
PostedFresh
Jul 21, 2026

Share this job

Help a friend find their next remote role.

Frequently asked questions

Is Director of Cloud Operations at Firstup a remote job?
Yes. Director of Cloud Operations at Firstup is a fully remote role open to candidates worldwide.
What type of employment is Director of Cloud Operations at Firstup?
This is a full time position.
How do I apply?
Click the "Apply" button on this page to go to the official application at Firstup.

Explore related

Market data & reports

Salary & skill-demand research built from our own listings data.

From the blog

Track this application + get a follow-up reminder

Free account · no credit card · Log in

Pro $9/mo · unlimited applies + AI resume