We are looking for a motivated Junior Database Reliability Engineer (DBRE) to join our expanding IT Operations team. The successful candidate will be part of a large general multi-international organization working with an array of NoSQL database platforms and applications. You will own the end-to-end lifecycle of NoSQL database instances.
What You Will Do
- Lifecycle Management: Own the full lifecycle of NoSQL database instances, including provisioning, configuration, upgrades, monitoring, performance tuning, and decommissioning.
- Provisioning & Deployment: Use Infrastructure-as-Code tools (Terraform, Ansible, Puppet) to provision and manage NoSQL clusters in a mixture of containerized (Docker/Kubernetes) and non-containerized environments.
- Monitoring & Observability: Implement and maintain monitoring dashboards, alerting, and health checks for distributed NoSQL clusters to enable proactive incident detection and resolution.
- Performance Engineering: Profile server resource usage, analyze query and cluster-level performance, identify bottlenecks,
and implement data models, indexing strategies, and configuration changes to optimize throughput and latency.
- Capacity Planning: Plan and model resource requirements based on usage trends, and recommend right-sizing or scaling actions to prevent capacity-related incidents.
- Incident & Change Management: Follow established change management processes, participate in on-call rotations, and contribute to post-incident reviews to drive long-term reliability improvements.
- Cross-Data Centre Replication: Support replication and failover configurations across data centres to ensure high availability and disaster recovery.
- Automation: Develop and maintain automation scripts (Shell, Python) and tooling to reduce manual operational overhead and improve consistency.
- Documentation: Prepare clear, well-structured documentation covering database architecture, runbooks, procedures, and operationa