Database Administrator
Indexed description
We are looking for a motivated Database Administrator (Database Reliability Engineer) to join our Clients expanding IT Operations team. The successful candidate will be part of a large global multi-international organization working with an array of NoSQL database platforms and applications. You will own the end-to-end lifecycle of NoSQL database instances.
Responsibilities:
- Lifecycle Management: Own the full lifecycle of NoSQL database instances, including provisioning, configuration, upgrades, monitoring, performance tuning, and decommissioning.
- Provisioning & Deployment: Use Infrastructure-as-Code tools (Terraform, Ansible, Puppet) to provision and manage NoSQL clusters in a mixture of containerized (Docker/Kubernetes) and non-containerized environments.
- Monitoring & Observability: Implement and maintain monitoring dashboards, alerting, and health checks for distributed NoSQL clusters to enable proactive incident detection and resolution.
- Performance Engineering: Profile server resource usage, analyze query and cluster-level performance, identify bottlenecks, and implement data models, indexing strategies, and configuration changes to optimize throughput and latency.
- Capacity Planning: Plan and model resource requirements based on usage trends, and recommend right-sizing or scaling actions to prevent capacity-related incidents.
- Incident & Change Management: Follow established change management processes, participate in on-call rotations, and contribute to post-incident reviews to drive long-term reliability improvements.
- Cross-Data Centre Replication: Support replication and failover configurations across data centres to ensure high availability and disaster recovery.
- Automation: Develop and maintain automation scripts (Shell, Python) and tooling to reduce manual operational overhead and improve consistency.
- Documentation: Prepare clear, well-structured documentation covering database architecture, runbooks, procedures, and operational standards.
- Collaboration: Work closely with SREs, platform engineers, application teams, and business stakeholders to understand requirements and deliver reliable database services.
Requirements:
- 2+ years of experience in IT operations, site reliability, database administration, or a related engineering role.
- Solid understanding of NoSQL databases (e.g., Cassandra, Elasticsearch, MongoDB) — hands-on experience with at least one of these is expected.
- Strong Linux administration skills and comfort working in shell-based environments.
- Experience with Infrastructure-as-Code and configuration management tools (Terraform, Ansible, Puppet).
- Proficiency in scripting languages (Shell, Python, or Ruby).
- Familiarity with Docker and/or Kubernetes for containerized database deployments.
- Basic knowledge of SQL databases (MySQL, PostgreSQL, Oracle, etc.) is a plus.
- Experience with Git and working in an open-source version control workflow.
- Working knowledge of *nix system tuning (kernel parameters, filesystem, memory, CPU) is a plus.
- Strong interpersonal skills, the ability to work in a team, and a proactive problem-solving mindset.
- Up-to-date knowledge of database reliability best practices and industry standards.
- Ability to plan and model resource requirements from high-level specifications.
- Training or certifications in cloud platforms (AWS/Azure/GCP) or NoSQL databases is a plus.
- Able to communicate fluently in written and spoken English.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search