{"id":10191,"url":"https://alion.io/job/cloudlinux-senior-database-reliability-engineer-dbre-worldwide-remote","title":"Senior Database Reliability Engineer (DBRE) (remote work)","company":{"id":3432,"name":"Cloudlinux","domain":"cloudlinux.com","url":"https://alion.io/company/cloudlinux","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workable","truth_index":{"grade":"A","score":88,"open_postings":4,"ghost_share":0,"stale_share":0.5,"repost_share":0,"time_to_fill_p50_days":30,"computed_at":"2026-10-01T05:45:00Z"}},"role":"Data Science","role_family":"Data Science","seniority":"senior","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"inferred","locations":["Warsaw, Poland"],"countries":["PL"],"hiring_countries":["PL","ES","AM","GE","SK"],"hiring_countries_total":5,"salary":null,"salary_estimate":{"min_usd":67000,"max_usd":99000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":61},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Ansible","optional":false},{"name":"CI/CD","optional":false},{"name":"Claude","optional":false},{"name":"ClickHouse","optional":false},{"name":"DNS","optional":false},{"name":"GitLab CI","optional":false},{"name":"Grafana","optional":false},{"name":"Linux","optional":false},{"name":"OpenTofu","optional":false},{"name":"Opsgenie","optional":false},{"name":"PostgreSQL","optional":false},{"name":"Redis","optional":false},{"name":"SQL","optional":false},{"name":"Terraform","optional":false},{"name":"ZooKeeper","optional":true}],"status":"live","first_seen_at":"2026-07-13T00:00:00Z","employer_posted_date":"2026-07-13","last_verified_at":"2026-10-01T23:01:09Z","board_verified":true,"closed_at":null,"days_open":81,"trust":{"level":"ok","repost_count":0,"flags":[],"days_open":80},"description":"CloudLinux / TuxCare is a remote-first infrastructure and security company. More than 300 engineers build and operate products used by hosting providers, enterprises, and internal service teams worldwide. Our Infrastructure Department runs the platforms behind CloudLinux OS, Imunify, KernelCare, TuxCare ELS, and our engineering systems.\nWe are hiring a Senior Database Reliability Engineer to join the Infrastructure DBA cell. This is a hands-on production ownership role, not a narrow ticket-processing DBA position. You will keep critical database services reliable, automate repeated work, support engineering teams, and reduce single-person dependency in our PostgreSQL, ClickHouse, MongoDB, and Redis operations.\nPostgreSQL is the main requirement. ClickHouse experience is a strong plus, but it is not a day-one blocker. We need a senior engineer with enough database, Linux, automation, and incident-response depth to learn our ClickHouse environment quickly and operate it safely.\nYour Responsibilities:\nOwn production PostgreSQL reliability: HA design, Patroni, PgBouncer, replication, failover, upgrades, vacuum/bloat control, query tuning, locks, indexes, capacity, backups, PITR, and restore validation.\nImprove disaster recovery and operational evidence: tested restores, documented recovery paths, measurable RTO/RPO targets, runbooks, and safe maintenance plans.\nSupport the wider database estate: ClickHouse, MongoDB, and Redis. You will troubleshoot incidents, review access and data-safety changes, improve monitoring, and learn the production ClickHouse patterns already in use.\nAutomate DBA workflows with Ansible, Terraform/OpenTofu, GitLab CI/CD, scripts, and reproducible runbooks for provisioning, grants, backups, restores, health checks, and ownership metadata.\nHelp build DBaaS-style self-service capabilities so engineering teams can request databases, access, credentials, and operational checks with less manual DBA intervention.\nImprove observability and incident response through Grafana, metrics, logs, SLOs, alert rules, Opsgenie routing, and clear communication during production issues.\nWhat Success Looks Like:\nPostgreSQL clusters have tested backup and restore paths, useful dashboards, clear ownership, and documented failover procedures.\nRepeated DBA tickets become automation or self-service workflows.\nClickHouse operational knowledge is no longer a single-person dependency.\nDatabase incidents have owners, runbooks, evidence, and measurable recovery paths.\nProduct and engineering teams get database help faster without sacrificing safety, auditability, or reliability.\nWhy CloudLinux?\nYou will work on real production infrastructure used across CloudLinux and TuxCare products. \nYou will have a direct impact on reliability, incident response, developer experience, and operational resilience. \nYou will also work in an AI-assisted engineering culture where automation, documentation, Claude, Codex, and careful human verification are part of the daily operating model.\nRequirements\nWhat We Expect From You:\nDeep hands-on PostgreSQL experience in business-critical production environments, typically 5+ years or equivalent depth.\nStrong understanding of PostgreSQL internals and operations: MVCC, WAL, transactions, locks, indexes, query planning, replication, autovacuum, bloat, major upgrades, backups, PITR, and restore testing.\nProven experience with highly available databases and the ability to reason about quorum, split-brain risk, failover, rollback, and recovery.\nMongoDB replica sets and Percona Backup for MongoDB.\nStrong Linux and infrastructure fundamentals: systemd, networking, storage, filesystems, CPU/memory/disk bottlenecks, TLS, DNS, firewalls, and root-cause troubleshooting.\nAutomation skills with Ansible and scripting. Terraform/OpenTofu, GitLab CI/CD, and merge-request based delivery are strong advantages.\nAbility to support more than one database engine. You do not need to be a ClickHouse expert on day one, but you must be ready to learn it quickly and take responsibility for it.\nPractical use of AI engineering assistants such as Claude and Codex. We expect you to use them to improve speed and quality, while personally verifying generated SQL, commands, scripts, and operational conclusions.\nEnglish - upper-intermediate or higher - to ensure clear communication of progress within the teams.\nNice to Have:\nClickHouse operations: replication, Keeper/ZooKeeper, MergeTree engines, distributed DDL, grants, row policies, backups, query troubleshooting, and cluster recovery.\nRedis/Sentinel and broker/cache failure modes.\nDatabase observability, SLOs, golden signals, alert tuning, and executable incident runbooks.\nBuilding internal platforms, self-service portals, or DBaaS workflows for engineering teams.\nBenefits\nWhat's in it for you?\nA focus on professional development.\nInteresting and challenging projects.\nFully remote work with flexible working hours, which allows you to schedule your day and work from any location worldwide.\nPaid 24 days of vacation per year, 10 days of national holidays, and unlimited sick leaves.\nCompensation for private medical insurance.\nCo-working and gym/sports reimbursement.\nBudget for education.\nThe opportunity to receive a reward for the most innovative idea that the company can patent.\nBy applying for this position, you consent to the processing of your personal data as described in our Privacy Policy (https://cloudlinux.com/candidate-privacy-notice), which provides detailed information on how we maintain and handle your data.","description_format":"text","description_chars":5556,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[{"language":"English","level":"Upper-Intermediate (B2)","optional":false}]},"benefits":["Flexible schedule","Health insurance","Professional development"],"hiring_locations":[{"name":"Poland","iso":"PL","kind":"country"},{"name":"Spain","iso":"ES","kind":"country"},{"name":"Armenia","iso":"AM","kind":"country"},{"name":"Georgia","iso":"GE","kind":"country"},{"name":"Slovakia","iso":"SK","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["IT Service & Asset Management","Operating Systems","Vulnerability Management"],"lifecycle":[{"event":"open","at":"2026-05-15T00:00:00Z"}],"liveness":{"score":11,"band":"cold","label":"Long shot","p_open":1,"p_active":0.385,"p_room":0.28,"age_days":80,"expected_fill_days":30,"reasons":["conf:11","velocity","win:tail","crowd:"],"computed_at":"2026-10-01T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/cloudlinux-senior-database-reliability-engineer-dbre-worldwide-remote","json_url":"https://alion.io/job/cloudlinux-senior-database-reliability-engineer-dbre-worldwide-remote.json","meta":{"generated_at":"2026-10-02T01:13:56Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1384,"day_limit":5000,"remaining_today":3616,"minute_limit":60,"resets_at":"2026-10-03T00:00:00Z"}}}