Sr Database Administrator - Production Reliability
- Full-time
- Quick Apply
Must Have:
8+ years of experience in database engineering, administration, and production support.
Hands-on expertise in Oracle, MongoDB, and PostgreSQL, including high availability, replication,
clustering, patching, and upgrades.
Strong experience in performance tuning, query optimization, indexing, capacity planning, backup,
PITR, and disaster recovery.
Ability to monitor and troubleshoot complex production issues, perform root-cause analysis, and
support on-call operations.
Working knowledge of Linux/Unix and scripting or automation using Python, Shell, or equivalent
tools.
Good To Have
Experience with AWS, Azure, or Google Cloud managed database services.
Knowledge of infrastructure as code, CI/CD, and database observability tools such as
Prometheus, Grafana, Splunk, or ELK/OpenSearch.
Understanding ITIL processes and experience preparing runbooks, operational dashboards, and
DR test reports.
Relevant database, cloud, or site reliability engineering certifications.
Responsibility / Expectation
Administer and maintain highly available Oracle, MongoDB, and PostgreSQL production
environments.
Perform database performance tuning, query optimization, indexing, and capacity planning.
Monitor database health, replication, storage, connections, locks, and resource consumption, and
resolve issues proactively.
Lead database incident triage, service restoration, root-cause analysis, and preventive actions.
Automate provisioning, monitoring, maintenance, backup validation, recovery testing, and
operational reporting.
Partner with application and platform teams on schema design, migrations, releases, security, and
non-functional requirements.
Maintain database documentation, runbooks, DR evidence, and operational status updates for
stakeholders.
Design and manage robust backup, point-in-time recovery (PITR), failover and DR strategies to
guarantee strict Recovery Point Objective (RPO) and Recovery Time Objective (RTO) targets