Cluster Administration & HA
- Install, configure, and administer PostgreSQL clusters across production and non-production environments.
- Configure and troubleshoot streaming replication (sync/async), replication slots, WAL archiving, standby servers, failover and switchover.
- Build and maintain HA infrastructure using distributed consensus, and HAProxy/Keepalived (or equivalent) for traffic management.
- Design and troubleshoot failover paths — not just operate an existing cluster, but reason about how it fails and recovers.
Connection Management
- Configure and manage PgBouncer: pool modes, connection limits, saturation behavior, and connection-related incident troubleshooting.
Upgrades & Migrations
- Plan and execute PostgreSQL major-version upgrades and migrations off legacy environments, including compatibility assessment, rollback strategy, downtime planning, and post-upgrade validation.
Backup & Recovery
- Implement and troubleshoot physical and logical backup/recovery, WAL archiving, and Point-in-Time Recovery (PITR).
Performance & Troubleshooting
- Diagnose locks, blocking, deadlocks, long-running transactions, connection exhaustion, replication lag, WAL accumulation, checkpoint pressure, autovacuum stalls, table/index bloat, and disk/I/O constraints.
- Read EXPLAIN/EXPLAIN ANALYZE plans and drive query, indexing, partitioning, and configuration tuning.
- Tune core parameters (buffers,mems, WAL/checkpoint settings, connection limits, autovacuum) against actual workload characteristics.
- Determine whether a given performance problem originates in SQL, the planner, PostgreSQL config, locking, replication, OS resources, or storage I/O.
Monitoring & Automation
- Install and maintain PMM, Grafana, Prometheus (or equivalent) for database observability.
- Build alerting for replication, WAL, connections, locks, transactions, checkpoints, autovacuum, database growth, CPU, memory, disk, and I/O.
- Automate DBA operations with Bash/Shell, Python, Ansible (or equivalent).
SQL & Documentation
- Write, review, and optimize complex SQL — CTEs, window functions, joins, subqueries, aggregations, data manipulation.
- Maintain configuration, architecture, backup/recovery, and operational documentation.
- Perform root-cause analysis on production incidents and drive prevention, not just remediation.
Experience: 3+ Years
Please send your CV at vacancy@khalti.com
Khalti खाता छैन?
Download now For more updates about Khalti’s campaign, events, services, and offer, you can also follow us on our official Facebook page, Youtube, Twitter, Viber, Linkedin, and Instagram.
