Senior Database Reliability Engineer
Ensure PostgreSQL reliability and automate DBA workflows
As a Senior Database Reliability Engineer, you will own production PostgreSQL reliability, improve disaster recovery, and automate DBA workflows. You will support the wider database estate, including ClickHouse, MongoDB, and Redis, and help build DBaaS-style self-service capabilities.
Why This Role?
Opportunity to learn and operate ClickHouse environment safely
Key Responsibilities
- Own production PostgreSQL reliability, including HA design and replication
- Improve disaster recovery and operational evidence, such as tested restores and documented recovery paths
- Automate DBA workflows with Ansible, Terraform, and GitLab CI/CD
- Support the wider database estate, including ClickHouse, MongoDB, and Redis
- Help build DBaaS-style self-service capabilities for engineering teams
Requirements
- PostgreSQL experience, including HA design, replication, and query tuning
- Linux and automation experience, including Ansible and Terraform
- Incident-response and database operations experience
- ClickHouse experience is a strong plus
Required Skills
Keywords
View Original Description from RemoteOK
Original description from RemoteOK
CloudLinux / TuxCare is a remote-first infrastructure and security company. More than 300 engineers build and operate products used by hosting providers, enterprises, and internal service teams worldwide. Our Infrastructure Department runs the platforms behind CloudLinux OS, Imunify, KernelCare, TuxCare ELS, and our engineering systems. We are hiring a Senior Database Reliability Engineer to join the Infrastructure DBA cell. This is a hands-on production ownership role, not a narrow ticket-processing DBA position. You will keep critical database services reliable, automate repeated work, support engineering teams, and reduce single-person dependency in our PostgreSQL, ClickHouse, MongoDB, and Redis operations. PostgreSQL is the main requirement. ClickHouse experience is a strong plus, but it is not a day-one blocker. We need a senior engineer with enough database, Linux, automation, and incident-response depth to learn our ClickHouse environment quickly and operate it safely. Your Responsibilities: Own production PostgreSQL reliability: HA design, Patroni, PgBouncer, replication, failover, upgrades, vacuum/bloat control, query tuning, locks, indexes, capacity, backups, PITR, and restore validation. Improve disaster recovery and operational evidence: tested restores, documented recovery paths, measurable RTO/RPO targets, runbooks, and safe maintenance plans. Support the wider database estate: ClickHouse, MongoDB, and Redis. You will troubleshoot incidents, review access and data-safety changes, improve monitoring, and learn the production ClickHouse patterns already in use. Automate DBA workflows with Ansible, Terraform/OpenTofu, GitLab CI/CD, scripts, and reproducible runbooks for provisioning, grants, backups, restores, health checks, and ownership metadata. Help build DBaaS-style self-service capabilities so engineering teams can request databases, access, credentials, and operational checks with less manual DBA intervention. Improve observability and incident response through Grafana, metrics, logs, SLOs, alert rules, Opsgenie routing, and clear communication during production issues. What Success Looks Like: PostgreSQL clusters have tested backup and restore paths, useful dashboards, clear ownership, and documented failover procedures. Repeated DBA tickets become automation or self-service workflows. ClickHouse operational knowledge is no longer a single-person dependency. Database incidents have owners, runbooks, evidence, and measurable recovery paths. Product and engineering teams get database help faster without sacrificing safety, auditability, or reliability. Why CloudLinux? You will work on real production infrastructure used across CloudLinux and TuxCare products. You will have a direct impact on reliability, incident response, developer experience, and operational resilience. You will also work in an AI-assisted engineering culture where automation, documentation, Claude, Codex, and careful human verification are part of the daily operating model. What We Expect From You: Deep hands-on PostgreSQL experience in business-critical production environments, typically 5+ years or equivalent depth. Strong understanding of PostgreSQL internals and operations: MVCC, WAL, transactions, locks, indexes, query planning, replication, autovacuum, bloat, major upgrades, backups, PITR, and restore testing. Proven experience with highly available databases and the ability to reason about quorum, split-brain risk, failover, rollback, and recovery. Strong Linux and infrastructure fundamentals: systemd, networking, storage, filesystems, CPU/memory/disk bottlenecks, TLS, DNS, firewalls, and root-cause troubleshooting. Automation skills with Ansible and scripting. Terraform/OpenTofu, GitLab CI/CD, and merge-request based delivery are strong advantages. Ability to support more than one database engine. You do not need to be a ClickHouse expert on day one, but you must be ready to learn it quickly and take responsibility for it. Practical
Free account · no credit card · Log in
Pro $9/mo · unlimited applies + AI resume
Source site may be blocked by Indonesian ISPs
Some Indonesian ISPs (Telkomsel, Indihome) block RemoteOK. If the Apply button doesn't open, try mobile data or a VPN.
Tip: switch network or enable a VPN, then click Apply again.
Frequently asked questions
- Is Senior Database Reliability Engineer at Cloudlinux a remote job?
- Yes. Senior Database Reliability Engineer at Cloudlinux is a fully remote role open to candidates worldwide.
- What type of employment is Senior Database Reliability Engineer at Cloudlinux?
- This is a full time position.
- How do I apply?
- Click the "Apply" button on this page to go to the official application at Cloudlinux.
Explore related
Market data & reports
Salary & skill-demand research built from our own listings data.
- Indonesia IT Jobs vs Global Remote (2026)Primary analysis of 2,049 listings: methodology, classification rules, downloadable datasets.
- AI-Skill Demand: Indonesia vs Global Remote (2026)10,000+ postings, taxonomy-first classifier, Wilson CIs, pre-registered before analysis.
- Indonesia Hiring Report: Tech vs Non-TechJob demand by field from aggregate open-job counts — never individual listings.
- Indonesia Salary BenchmarkAggregate salary ranges across roles, with open methodology and dataset.
- Indonesia Quarterly Labor Market ReportLayoffs, funding, salaries & skills per quarter — open aggregates.
- Remote Market Reports by RoleAuto-generated per role family — skills, seniority, companies, salary.
- Global Remote Salary BenchmarkAnnual salary by role & currency, plus the share of listings open worldwide.
From the blog
- Senior Tech Roles Remote Salaries June 2026In-depth analysis of 7 senior tech role remote salaries from Vercel, Airbnb, Stripe, to Notion. Compare with local market and negotiation strategies.
- Sourcing Specialist: The Complete GlobalCurious about becoming a remote Sourcing Specialist? Learn the essential skills, tools, and how to land a USD-paying job—no matter where you live.
- Remote Mobile Developer Jobs July 2026A roundup of USD-paying remote mobile developer jobs from Loker Dollar. Analyze trends and get application tips.
Free account · no credit card · Log in
Pro $9/mo · unlimited applies + AI resume