Database SRE
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+19 more
Job description
- We are seeking an experienced Database Site Reliability Engineer (SRE) to support and operate mission-critical database platforms within a fast-paced enterprise environment. This role is focused on operational excellence, reliability, resiliency, automation, and continuous improvement across multiple database technologies.
- This is not a development role. We are looking for hands-on database professionals who thrive in production operations, take ownership of issues, understand urgency, and are passionate about improving systems, processes, and themselves.
- The ideal candidate views reliability as a product, proactively identifies risks before they become incidents, and continuously seeks opportunities to automate repetitive tasks and improve platform stability., Database Operations & Reliability
- Install, configure, upgrade, patch, and maintain enterprise database platforms.
- Ensure availability, performance, recoverability, and security of production database environments.
- Monitor database platforms and associated infrastructure, responding rapidly to incidents and service degradations.
- Lead troubleshooting efforts for database, operating system, storage, replication, and application connectivity issues.
- Execute failovers, disaster recovery testing, and recovery procedures.
- Partner with application teams to provide database guidance and operational support.
Platform Engineering
- Deploy, maintain, and optimize database infrastructure across physical, virtual, and cloud environments.
- Implement scalable, resilient database solutions.
- Evaluate and recommend improvements to architecture, monitoring, automation, and operational processes.
- Support capacity planning, performance tuning, and platform lifecycle management.
Automation & Continuous Improvement
- Develop and maintain automation solutions using Python, Shell, Ansible, or similar technologies.
- Help eliminate manual operational activities through engineering and automation.
- Improve monitoring, alerting, reporting, and operational workflows.
- Drive incremental improvements that reduce risk, improve reliability, and increase operational efficiency.
Performance & Incident Management
- Analyze and resolve database performance issues.
- Troubleshoot replication, backup/recovery, storage, network, and infrastructure-related incidents.
- Participate in root cause analysis and drive permanent corrective actions.
- Review operational metrics and trends to identify opportunities for improvement.
Operational Excellence
- Maintain accurate operational documentation, standards, and procedures.
- Generate and present operational metrics, service health indicators, and reliability reporting.
- Participate in incident response activities.
- Demonstrate strong ownership from issue identification through resolution.
Requirements
Must have STRONG LINUX exp 90% of what this role is, is LINUX
- Experience performing:
-Installation -Configuration -Upgrades -Patching -Performance tuning -Backup and recovery -High availability and disaster recovery -Good Coding Exp -Python, Ansible and/or shell
- Relational Database Experience
-Sybase or Microsoft SQL Server
- Understanding of storage, networking, operating systems, and infrastructure services.
- Experience with Veritas Cluster Server, ASM, LVM, and SAN technologies.
- Familiarity with enterprise operational tooling such as Jira, Service Now and Confluence.
- Strong analytical, troubleshooting, and problem-solving skills.
- Great communication (written and especially verbal)
- Service Now, Jira, Confluence exp
- Good Sense of urgency
- Strong experience administering enterprise database platforms, including:
-Sybase ASE -Oracle RAC
- Additional database technologies such as MongoDB, Cassandra, Redis, PostgreSQL, MySQL, or similar platforms are a plus.
- Experience with database replication technologies including:
- SAP Replication Server
- Data Guard
- HVR (preferred)
Plus: Sybase + Linux = UNICORN Multiple Database Platforms, * Customer obsessed
- Accountable and dependable
- Urgent without being reckless
- Continuously learning
- Driven to automate repetitive work
- Detail-oriented
- Collaborative but willing to lead
- Focused on long-term platform reliability rather than short-term fixes
About the company
PrideGlobal and it’’s affiliates offers eligible employee’’s comprehensive healthcare coverage (medical, dental, and vision plans), supplemental coverage (accident insurance, critical illness insurance and hospital indemnity), 401(k)-retirement savings, life & disability insurance, an employee assistance program, legal support, pet insurance and employee discounts with preferred vendors.
Equal Employment Opportunity PrideGlobal and it’’s affiliates are an equal opportunity employer. We do not discriminate on the basis of the race, religious creed, color, national origin, ancestry, physical disability, mental disability, reproductive health decision making, medical condition, genetic information, marital status, sex, gender, gender identity, gender expression, age, sexual orientation, veteran or military status, or any other characteristic protected by applicable federal, state, or local law.
Fair Chance Employment PrideGlobal and it’’s affiliates are a Fair Chance employer. We consider all qualified applicants, including those with criminal histories, in a manner consistent with applicable state and local Fair Chance laws and ordinances, including, the California Fair Chance Act and all applicable local Fair Chance ordinances.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Making Data Warehouses Fast: A Developer’s Story
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Top Big Data Technologies That You Need to Know
What Are The Top Skills Required For Azure Developers?