Senior Database Reliability Engineer (DBRE)
CloudLinux · Erevan
Job description
About the role
We are looking for a Senior Database Reliability Engineer to join our Infrastructure DBA cell. This remote‑first position is hands‑on production ownership, focused on keeping critical database services reliable and reducing single‑person dependencies across PostgreSQL, ClickHouse, MongoDB, and Redis.
Key responsibilities
- Own production PostgreSQL reliability: HA design, Patroni, PgBouncer, replication, failover, upgrades, vacuum/bloat control, query tuning, backups, PITR and restore validation.
- Improve disaster recovery: test restores, document recovery paths, set measurable RTO/RPO targets, create runbooks and safe maintenance plans.
- Support the broader database estate (ClickHouse, MongoDB, Redis): troubleshoot incidents, review access changes, enhance monitoring, and learn existing ClickHouse patterns.
- Automate DBA workflows with Ansible, Terraform/OpenTofu, GitLab CI/CD, scripts and reproducible runbooks for provisioning, grants, backups, restores and health checks.
- Build DBaaS‑style self‑service capabilities so engineering teams can request databases, credentials and operational checks with minimal manual intervention.
- Enhance observability and incident response using Grafana, metrics, logs, SLOs, alert rules, Opsgenie routing and clear communication during production issues.
Required profile
- Senior‑level engineer with deep PostgreSQL expertise and strong Linux background.
- Proven experience in automation, incident response and capacity planning.
- Ability to quickly acquire operational knowledge of ClickHouse and other databases.
- Collaborative mindset with a focus on reducing manual DBA effort.
Required skills
- PostgreSQL (Patroni, PgBouncer, replication, backup/restore).
- ClickHouse (experience is a strong plus).
- MongoDB and Redis.
- Linux system administration.
- Ansible, Terraform/OpenTofu.
- GitLab CI/CD pipelines.
- Scripting (bash, Python or similar).
- Grafana, monitoring, metrics and alerting (Opsgenie).
What we offer
- Work on real production infrastructure used by CloudLinux and TuxCare products.
- Direct impact on reliability, incident response and developer experience.
- Remote‑first culture with AI‑assisted engineering tools.
- Opportunity to shape DBaaS services and automation frameworks.
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches for Հայաստան.
Salaries by job title
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
Published 3 շաբաթ առաջ
Expires 1 ամիսից
26 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
CloudLinux
Erevan