1 day ago
Remote, IndiaStaff+
Responsibilities
- Serve as the global Tier-3 escalation authority and database subject matter expert for functionality, concurrency, performance, and complex troubleshooting.
- Monitor and resolve advanced database alerts and lead mitigation of high-severity production incidents across AWS and GCP environments.
- Perform execution-plan analysis, DMV/DMF audits, index optimization, query recompilation, wait-event analysis, and live concurrency troubleshooting.
- Execute database maintenance, integrity checks, backup and snapshot validation, replication monitoring, point-in-time restores, failovers, schema deployments, storage expansions, and configuration changes.
- Support database engine upgrades, security patch cycles, parameter modifications, and managed or self-managed cloud database operations.
- Create automation scripts, diagnostic runbooks, SOPs, self-service troubleshooting tools, and self-healing responses to reduce operational toil.
- Lead post-incident reviews and root-cause analysis while driving permanent remediation and alert-fatigue reduction.
- Partner with SRE, Platform Engineering, Support, Implementation, and Product Engineering teams on database releases, migrations, CI/CD tooling, and operational improvements.
- Conduct global shift handovers, provide technical leadership during IST daytime coverage, and mentor Tier-1 and Tier-2 support teams.
Requirements
- Bachelor’s degree in Computer Science, Information Technology, Computer Applications, or a related technical discipline, or equivalent education, certifications, and extensive professional experience.
- 10+ years of progressive experience in enterprise relational database administration and engineering for mission-critical, high-availability 24/7 cloud production systems.
- Principal- or lead-level database engineering or SME experience with autonomous off-hours incident response.
- Extensive hands-on experience with AWS or GCP database environments, including RDS, Aurora, Cloud SQL, or EC2/Compute Engine deployments.
- Deep expertise in Microsoft SQL Server internals, AlwaysOn Availability Groups, query optimization, locking and latching, transaction processing, and storage-engine architecture.
- Strong production experience with PostgreSQL or an equivalent open-source relational database engine.
- Expertise in query performance tuning, index optimization, wait-event analysis, dynamic tracing, backups, replication, failover, disaster recovery, and database maintenance.
- Experience with Dynatrace, Datadog, Prometheus, or Grafana for database observability and alerting.
- High proficiency in PowerShell, Python, T-SQL, or Bash for automation and diagnostics.
- Practical familiarity with Terraform and Ansible.
- Required certifications include Microsoft Certified: Azure Database Administrator Associate or legacy MCSE: Data Management and Analytics; AWS Certified Database – Specialty or AWS Certified Solutions Architect – Professional; and Google Cloud Professional Cloud Database Engineer.
- Prior Healthcare IT, HIPAA/HITRUST, or EHR/clinical data experience is expected or preferred, along with ITIL Foundation or SRE/DevOps certification experience.
- Strong written and verbal communication, technical mentorship, autonomous judgment, and cross-continental operational leadership skills.
Benefits
- IST daytime work schedule with structured overlap and handoffs supporting US overnight operations.
- Global 24/7 operational environment with collaboration across Cloud Support, Database Engineering, SRE, and Platform Engineering teams.
- Equal opportunity employer committed to an inclusive workplace.
Tech Stack
AnsibleAWSAzureBashDatadogGoogle Cloud PlatformGrafanaMicrosoft SQL ServerPostgreSQLPowerShellPrometheusPythonSQLTerraform
Categories
BackendSite Reliability
About NextGen
NEXTGEN is an Australia-based technology services and value-added distribution company that helps vendors and channel partners sell cybersecurity, cloud, enterprise software, and data management solutions. It offers software licensing, compliance and audit services, data centre and storage solutions, and go-to-market and digital marketing support. Founded in 2011 and headquartered in North Sydney, it is now part of Exclusive Networks, expanding its reach across international markets.
