← Back to list
Job

Database Reliability Engineer

Other • Remote • Full-time • Poland Poland

The company is hiring a senior database reliability engineer for its infrastructure DBA cell. You take hands-on ownership of production PostgreSQL, ClickHouse, MongoDB and Redis services, with automation and self-service work.

Responsibilities

  • ▹Own production PostgreSQL reliability: HA design, Patroni, PgBouncer, replication, failover, upgrades, vacuum/bloat control, query tuning, locks, indexes, capacity, backups, PITR and restore validation
  • ▹Improve disaster recovery and operational evidence: tested restores, documented recovery paths, measurable RTO/RPO targets, runbooks
  • ▹Support ClickHouse, MongoDB and Redis: troubleshoot incidents, review access and data-safety changes, improve monitoring
  • ▹Automate DBA workflows with Ansible, Terraform/OpenTofu, GitLab CI/CD and scripts
  • ▹Build DBaaS-style self-service capabilities for requesting databases, access and credentials
  • ▹Improve observability and incident response with Grafana, metrics, logs, SLOs, alert rules and Opsgenie routing

Requirements

  • ▹Deep hands-on PostgreSQL experience in business-critical production, typically 5+ years
  • ▹Understanding of PostgreSQL internals and operations: MVCC, WAL, transactions, locks, indexes, query planning, replication, autovacuum, bloat, major upgrades, backups, PITR, restore testing
  • ▹Experience with highly available databases, including quorum, split-brain, failover and recovery
  • ▹Strong Linux and infrastructure fundamentals: systemd, networking, storage, filesystems, bottlenecks, TLS, DNS, firewalls
  • ▹Automation skills with Ansible and scripting
  • ▹Ability to support more than one database engine and learn ClickHouse quickly
  • ▹Practical use of AI assistants such as Claude and Codex, personally verifying generated SQL, commands and scripts
  • ▹English at upper-intermediate level or higher

Nice to have

  • ▹ClickHouse operations: replication, Keeper/ZooKeeper, MergeTree engines, distributed DDL, grants, row policies, backups, troubleshooting, cluster recovery
  • ▹MongoDB replica sets and Percona Backup for MongoDB
  • ▹Redis/Sentinel and broker/cache failure modes
  • ▹Database observability, SLOs, golden signals, alert tuning and executable incident runbooks
  • ▹Building internal platforms, self-service portals or DBaaS workflows

Soft skills

Clear communication during incidentsSupportive approach toward engineering teamsOwnership

What we offer

  • ▹A focus on professional development
  • ▹Interesting and challenging projects
  • ▹Fully remote work with flexible hours, from anywhere in the world
  • ▹24 days of paid vacation, 10 days of national holidays, unlimited sick leave
  • ▹Compensation for private medical insurance
  • ▹Co-working and gym/sports reimbursement
  • ▹Education budget
  • ▹Reward for the most innovative idea the company can patent
Languages: Angol: felső-középfok vagy magasabb

Similar jobs