Now hiring

Database Administrator - Bengaluru Only (m/f/d) @ Cuculus

BengaluruOnsiteFull-time
Apply with ResuMinder

Opens on the employer's site

About this role

Shape the utilities market of the future with us! At <strong>CUCULUS</strong>, we build intelligent digital solutions that help utilities become more efficient, sustainable, and future-ready. By joining our technology team, you’ll work on business-critical systems that power large-scale utility operations, contribute to high-impact projects, and collaborate with global teams across cloud and hybrid environments. What is the role about? As a Senior Database Administrator, you will be responsible for highly available, mission-critical database environments and ensure their availability, performance, security, resilience and recoverability.<br><strong>You will:</strong><br><ul><li>Administer and support<strong> Oracle </strong>19c databases, including RAC and Data Guard.</li><li>Manage <strong>PostgreSQL</strong> databases across on-premise and cloud environments.</li><li>Administer and support <strong>ClickHouse</strong> database environments.</li><li>Support business-critical 24/7/365 production operations within defined SLAs.</li><li>Perform database performance tuning, SQL optimization and capacity planning.</li><li>Implement proactive database monitoring and health checks.</li><li>Plan and execute database migrations, upgrades and patching.</li><li>Maintain backup, restore and recovery strategies.</li><li>Support High Availability and Disaster Recovery environments.</li><li>Deploy and support database workloads on <strong>Kubernetes.</strong></li><li>Troubleshoot database, <strong>Linux</strong>, storage, connectivity and performance issues.</li><li>Develop scripts and automation for recurring DBA activities.</li><li>Perform deep root-cause analysis for database-related incidents.</li><li>Support critical production incidents and customer escalations.</li><li>Maintain DBA procedures, runbooks and technical documentation.</li><li>Work closely with System Engineering, Support and Development teams.</li></ul> We are always looking for support in the following areas: <strong>Your work will contribute directly to:</strong><br><ul><li>Production database availability</li><li>Database reliability and resilience</li><li>Performance and scalability</li><li>Faster incident resolution</li><li>SLA fulfilment</li><li>Backup and recovery readiness</li><li>HA/DR readiness</li><li>Proactive monitoring</li><li>Automation and operational efficiency</li><li>Continuous improvement of database operations</li></ul> Your Mission & Impact Your mission is to ensure that our database platforms remain stable, performant, secure, scalable and recoverable.<br>You will take ownership of complex database issues from investigation through recovery and root-cause analysis while proactively identifying risks before they affect production systems or customers Your Role at a Glance We are looking for a <strong>Senior Database Administrator</strong> with strong expertise in <strong>Oracle, PostgreSQL and ClickHouse, backed by solid Linux, Kubernetes</strong> and cloud knowledge.<br>You will manage and support highly available, mission-critical production database environments across on-premises, cloud and hybrid platforms.<br>The role requires strong hands-on technical capability, independent troubleshooting, production ownership and the ability to support critical database incidents within strict SLAs. Must-haves (to thrive in this role) <strong>Strong hands-on experience with Oracle 19c, including:</strong><br><ul><li>Oracle RAC</li><li>Data Guard</li><li>Database administration</li><li>Performance tuning</li><li>SQL optimization</li><li>Execution analysis</li><li>Backup and recovery</li><li>Capacity management</li><li>Patching</li><li>Upgrades</li><li>Migration</li><li>Security</li><li>High Availability</li><li>Disaster Recovery</li></ul><strong>PostgreSQL Administration</strong><br>Strong production PostgreSQL administration, including:<br><ul><li>Installation and configuration</li><li>Database administration</li><li>Backup and restore</li><li>Replication</li><li>PITR</li><li>Recovery</li><li>Performance analysis</li><li>Query troubleshooting</li><li>Migration</li><li>Upgrades</li><li>Capacity management</li><li>Security</li><li>High Availability</li></ul><strong>ClickHouse Administration</strong><br>Hands-on ClickHouse administration and operational support, including:<br><ul><li>Installation and configuration</li><li>Cluster operations</li><li>Query troubleshooting</li><li>Performance monitoring and tuning</li><li>Storage and capacity management</li><li>Backup and recovery</li><li>Replication</li><li>High Availability</li></ul><strong>Performance &amp; Capacity Management</strong><br>Monitor, analyse and optimize database performance, including:<br><ul><li>SQL/query optimization</li><li>Execution analysis</li><li>Database performance tuning</li><li>CPU and memory utilization</li><li>Storage utilization</li><li>Capacity planning</li><li>Connection/session analysis</li><li>Bottleneck identification</li><li>Storage growth</li><li>Database scalability</li></ul>Proactively identify capacity and performance risks before they become production incidents.<br><strong>Monitoring &amp; Observability</strong><br>Implement, maintain and continuously improve database monitoring and observability.<br>Experience with MONIT, Grafana, Prometheus or equivalent tools is expected.<br><strong>Monitor:</strong><br><ul><li>Database availability</li><li>Database health</li><li>Performance</li><li>Capacity</li><li>Storage</li><li>Connections and sessions</li><li>Replication</li><li>Backup status</li><li>HA/DR status</li><li>Logs</li><li>Alerts</li></ul><strong>Backup, Recovery &amp; Disaster Recovery</strong><br>Design, maintain and support reliable database backup and recovery capabilities, including:<br><ul><li>Backup strategies</li><li>Restore procedures</li><li>PITR</li><li>Replication</li><li>Recovery procedures</li><li>Recovery validation</li><li>Disaster Recovery</li><li>Failover/failback</li><li>Recovery testing</li><li>Production recovery</li></ul><strong>Linux Administration</strong><br>Strong hands-on Linux administration and troubleshooting skills, including:<br><ul><li>Processes and services</li><li>CPU and memory</li><li>Disk and storage</li><li>File systems</li><li>Permissions</li><li>Networking</li><li>Logs</li><li>Security</li><li>Performance troubleshooting</li></ul>The DBA should be able to determine whether a database issue originates from the database itself or from the underlying operating system, storage, infrastructure or network.<br><strong>Kubernetes</strong><br>Experience supporting databases in production Kubernetes environments, including:<br><ul><li>Pods</li><li>Services</li><li>Networking</li><li>Storage</li><li>Persistent volumes</li><li>Configuration</li><li>Monitoring</li><li>Logs</li><li>Database workload troubleshooting</li></ul>Ability to deploy and support database workloads on Kubernetes.<br><strong>Cloud &amp; Infrastructure</strong><br>Experience with:<br><ul><li>Oracle Cloud Infrastructure (OCI)</li><li>AWS and/or Azure</li><li>Cloud database services</li><li>RDS/PaaS database services</li><li>Cloud networking</li><li>Storage</li><li>Security</li><li>Access control</li><li>HA and multi-AZ concepts</li><li>Backup and recovery</li><li>Hybrid environments</li></ul><strong>Scripting &amp; Automation</strong><br>Experience with Shell/Bash, Python or equivalent scripting technologies.<br><strong>Use automation to improve:</strong><br><ul><li>Database health checks</li><li>Monitoring</li><li>Backup validation</li><li>Capacity checks</li><li>Diagnostics</li><li>Log collection</li><li>Reporting</li><li>Repetitive DBA activities</li><li>Operational efficiency</li></ul> <br><strong>Database Security &amp; Reliability</strong><br><strong>Support and maintain:</strong><br><ul><li>Database access controls</li><li>Users and roles</li><li>Permissions</li><li>Security patches</li><li>Database hardening</li><li>Linux security</li><li>Operational security controls</li><li>Database reliability standards</li></ul> <br><strong>Incident Management &amp; Root-Cause Analysis</strong><br>Independently handle complex and critical database incidents.<br><strong>Responsibilities include:</strong><br><ul><li>Incident investigation</li><li>Database troubleshooting</li><li>Service recovery</li><li>Root-cause analysis</li><li>Corrective actions</li><li>Preventive actions</li><li>Technical escalation</li><li>Post-incident review</li><li>Documentation of findings and solutions</li></ul>The objective is not only to restore the database service but also to identify the underlying root cause and prevent recurrence.<br> <br><strong>SLA &amp; Production Support</strong><br>Operate within SLA-driven production environments.<br><strong>You will:</strong><br><ul><li>Prioritize incidents according to severity and customer impact.</li><li>Support critical production incidents.</li><li>Identify potential SLA risks early.</li><li>Escalate when specialist or additional technical support is required.</li><li>Maintain clear technical documentation and ticket updates.</li><li>Drive database issues through sustainable resolution.</li></ul>Experience with structured ITSM/ticketing tools, preferably Jira / Jira Service Management, is an advantage.<br><strong>Documentation &amp; Knowledge Management</strong><br>Create, maintain and continuously improve:<br><ul><li>DBA procedures</li><li>SOPs</li><li>Runbooks</li><li>Troubleshooting guides</li><li>Backup/recovery procedures</li><li>Monitoring procedures</li><li>Known-error documentation</li><li>Knowledge-base articles</li></ul> <br><strong>Cross-Skilling – ZONOS Knowledge</strong><br>The primary responsibility of this role remains Database Administration.<br>As part of the Support cross-skilling approach, the DBA will progressively acquire operational knowledge of the ZONOS platform to better understand how databases interact with the wider application environment.<br>This includes basic operational understanding of:<br><ul><li>ZONOS architecture and components</li><li>Application/database dependencies</li><li>Application health</li><li>Application logs</li><li>Linux/Kubernetes dependencies</li><li>Database/application connectivity</li><li>Monitoring</li><li>Standard documented ZONOS troubleshooting procedures</li><li>This knowledge enables the DBA to contribute more effectively to L3 incident investigation and work closely with the E2E Support team.</li><li>The DBA remains the database specialist. Complex ZONOS application troubleshooting, architecture, product defects and code-level investigation remain with the respective E2E/System Engineering/Development specialists.</li><li>Previous ZONOS knowledge is not required at recruitment and will be developed through structured knowledge transfer and practical experience.</li></ul><strong>24/7 Production Operations</strong><br>Willingness and capability to participate in 24/7/365 production Support operations according to the defined shift/on-call model.<br>This may include scheduled:<br><ul><li>Day/night coverage</li><li>Weekend coverage</li><li>Public-holiday coverage</li><li>Critical incident support</li></ul>Structured technical handover is required to ensure service continuity.<br> <br><strong>Required Skills &amp; Qualifications</strong><br><ul><li>Minimum 5 years of relevant DBA experience, including at least 3 years supporting business-critical production database environments.</li><li>Strong Oracle 19c administration.</li><li>Strong Oracle RAC and Data Guard expertise.</li><li>Strong Oracle performance tuning and SQL optimization.</li><li>Strong PostgreSQL administration.</li><li>PostgreSQL backup, replication, PITR, recovery and migration.</li><li>Hands-on ClickHouse administration.</li><li>Strong Linux administration and troubleshooting.</li><li>Kubernetes production knowledge.</li><li>Experience supporting database workloads on Kubernetes.</li><li>Shell/Bash/Python or equivalent scripting skills.</li><li>OCI, AWS, Azure or equivalent cloud experience.</li><li>Strong monitoring and observability capabilities.</li><li>Strong backup, restore and recovery expertise.</li><li>Experience with HA and DR environments.</li><li>Strong production troubleshooting and RCA skills.</li><li>Experience operating within SLA-driven production environments.</li><li>Ability to handle critical production incidents independently.</li><li>Strong documentation and knowledge-sharing skills.</li></ul> Nice-to-haves (great if you have them, but not a dealbreaker) <ul><li>Experience with large-scale enterprise production systems.</li><li>Hybrid and multi-cloud experience.</li><li>Jira / Jira Service Management.</li><li>Grafana, Prometheus, MONIT or equivalent monitoring tools.</li><li>Terraform.</li><li>Ansible.</li><li>Docker.</li><li>Kubernetes Operators.</li><li>OKE.</li><li>Jenkins.</li><li>GitLab.</li><li>GitHub Actions.</li></ul> Your profile <strong>You bring:</strong><br><ul><li>Strong analytical and problem-solving skills.</li><li>Deep database troubleshooting capability.</li><li>Strong ownership and accountability.</li><li>Ability to work independently on critical production systems.</li><li>Ability to work effectively under pressure.</li><li>A proactive approach to performance, security and reliability.</li><li>Clear technical communication.</li><li>Strong collaboration across technical teams.</li><li>Willingness to continuously develop your technical knowledge.</li><li>Willingness to share knowledge and support cross-skilling.</li><li>Willingness to support 24/7 production operations.</li></ul> What You Bring to the Team <ul><li><p>Strong problem-solving skills and ability to handle high-pressure production issues</p></li><li><p>Experience working in <strong>high-availability and disaster recovery environments</strong></p></li><li><p>Clear communication and collaboration skills for cross-team coordination</p></li><li><p>A proactive mindset with attention to performance, security, and reliability</p></li><li><p>Willingness to support <strong>24×7 operations</strong> and participate in on-call rotations</p></li></ul><h3><strong>Good to Have</strong></h3><ul><li><p>Infrastructure as Code (Terraform, Ansible)</p></li><li><p>Docker, Kubernetes Operators, OKE</p></li><li><p>CI/CD tools (Jenkins, GitLab, GitHub Actions)</p></li></ul>

Ready to apply?

Install the ResuMinder extension and we'll auto-fill the application in seconds — no rewriting.

See how your CV scores