Database Site Reliability Engineer (SRE) for Database Operations l
ocation
onsite
with our client. Kindly review the job description below and
Job
.
Job Description –
Job Title Database Site Reliability Engineer (SRE) for Database Operations
Location Alpharetta, GA near Atlanta, GA
Interview Process: 1 video and 1 in-person interview
Position Summary
We are seeking an experienced
Database Site Reliability Engineer (SRE)
to support and operate mission-critical database platforms within a fast-paced enterprise environment. This role is focused on operational excellence, reliability, resiliency, automation, and continuous improvement across multiple database technologies.
Key Responsibilities
Database Operations & Reliability
Partner with application teams to provide database guidance and operational support.
Platform Engineering
Deploy, maintain, and optimize database infrastructure across physical, virtual, and cloud environments.
Implement scalable, resilient database solutions.
Evaluate and recommend improvements to architecture, monitoring, automation, and operational processes.
Support capacity planning, performance tuning, and platform lifecycle management.
Automation & Continuous Improvement
Develop and maintain automation solutions using Python, , Ansible, or similar technologies.
Help eliminate manual operational activities through engineering and automation.
Improve monitoring, alerting, reporting, and operational workflows.
Drive incremental improvements that reduce risk, improve reliability, and increase operational efficiency.
Performance & Incident Management
Analyze and resolve database performance issues.
Troubleshoot replication, backup/recovery, storage, network, and infrastructure-related incidents.
Participate in root cause analysis and drive permanent corrective actions.
Review operational metrics and trends to identify opportunities for improvement.
Operational Excellence
Maintain accurate operational documentation, standards, and procedures.
Generate and present operational metrics, service health indicators, and reliability reporting.
Participate in incident response activities.
Demonstrate strong ownership from issue identification through resolution.
Required Qualifications
Strong experience administering enterprise database platforms, including:
o Sybase ASE
Oracle RAC
Additional database technologies such as MongoDB, Cassandra, Redis, PostgreSQL, MySQL, or similar platforms are a plus.
Experience performing:
o Installation
Configuration
Upgrades
Patching
Performance tuning
Backup and recovery
High availability and disaster recovery
Experience with database replication technologies including:
o SAP Replication Server
Data Guard
HVR (preferred)
Strong Linux administration skills.
Experience with automation and scripting:
o Python
Ansible
scripting
Understanding of storage, networking, operating systems, and infrastructure services.
Experience with Veritas Cluster Server, ASM, LVM, and SAN technologies.
Familiarity with
enterprise operational tooling such as Jira, Service Now and Confluence.
Strong analytical, troubleshooting, and problem-solving skills.