> Markdown version of [/jobs/ext/2026565-senior-application-support-engineer-sre](https://www.wearedevelopers.com/jobs/ext/2026565-senior-application-support-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Application Support Engineer (SRE) - **Company:** The Depository Trust & Clearing Corporation - **Location:** Coppell, TX, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, CA Workload Automation Ae, Batch Processing, Software as a Service, Computer Programming, IBM DB2, Relational Databases, Linux, Disaster Recovery, Distributed Systems, Middleware, Monitoring of Systems, Job Control Language (JCL), Job Scheduling, Python (Programming Language), Log Analysis, Enterprise Messaging Systems, OpenShift, Oracle (Applications), Password Management, Release Management, Reliability Engineering, Standard Sql, Software Engineering, System Availability, Snowflake, Grafana, Reliability of Systems, Production Code, Splunk, Servicenow - **Published:** August 11, 2026 - **Apply:** https://ebxr.fa.us2.oraclecloud.com/hcmUI/CandidateExperience/en/sites/CX_1/requisitions/preview/214244 ## About the Role * Bachelor's degree preferred or equivalent practical experience * 6-8 years of experience in Application Support, Production Support, or SRE roles Talent Needed for Success * Strong understanding of application support and SRE principles including reliability engineering, observability and incident prevention. * Experience working in Linux and Windows environments including process inspection, log analysis and extensive troubleshooting. * Familiarity with monitoring and observability tools such as Splunk and Grafana and the ability to interpret system behavior and alerts. * Knowledge on application support - basic programming skills, log reading and analysis * Distributed Application troubleshooting and Support skills. Strong problem-solving skills with the ability to think creatively. * Familiarity working with relational databases (DB2, Oracle, Snowflake) * Working knowledge of SQL and ability to execute queries for analysis and troubleshooting. * Experience with ITSM and operational tooling (e.g., ServiceNow) for incident, problem, and change management. * Familiarity with job scheduling, containerized platforms and modern application environments (e.g., Autosys, OpenShift). * Understanding of security fundamentals including certificate and password management. * Exposure to capital markets and financial industry is required. * Exposure to messaging, networking or mainframe concepts is a plus. * Experience in handling issues related AWS and associated services . * Exposure to artificial intelligence concepts and their usage in production support * Demonstrates clear written and verbal communication, ownership and leadership in fast-paced production environments. * Comfortable operating with urgency and collaborating with global , distributed teams. * Proactive mindset with a focus on continuous improvement, automation and operational excellence Preferred Skills * Exposure to mainframe environments (job monitoring, batch processing, JCL) * Experience with AWS or cloud-based applications * Basic scripting or programming experience (Python or similar) * Understanding of messaging systems, networking, or middleware * Exposure to AI/automation concepts in production support ## Description As a Senior Application Support Engineer, you will play a critical role in supporting and improving the reliability of DTCC's risk management applications. This role goes beyond traditional support-you will apply Site Reliability Engineering (SRE) principles to improve system stability, reduce incidents, and drive proactive improvements across a complex environment spanning mainframe, distributed systems, and cloud platforms. You will support a large-scale portfolio of over 100 applications and work as part of a globally distributed team, partnering across regions to ensure seamless production operations and high system availability. You will partner closely with Application Development, Infrastructure, and Operations teams to ensure production stability for mission-critical systems supporting financial risk and settlement processes. Your Primary Responsibilities: * Act as a Lead Application Support Engineer with SRE responsibilities, partnering with Application Development, Infrastructure and Operations teams to improve system reliability, resilience and observability. * Lead the resolution of critical production incidents, providing clear impact analysis, root cause identification and preventative actions. * Drive incident, problem and major incident management maintaining ownership through resolution and post-incident review. * Proactively identify reliability risks and implement improvements to prevent recurrence and reduce operational toil. * Review and maintain runbooks, knowledge articles and operational documentation to ensure production readiness and consistency. * Execute change, release and deployment activities, including production code releases and vendor application upgrades. * Perform and support Disaster Recovery activities, including testing, execution, and audit/BCM evidence collection. * Identify and implement automation and alert rationalization opportunities to improve operational efficiency and service stability. * Embed reliability, risk and control considerations into day-to-day operations, escalating issues appropriately. ## Related Videos - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers)