> Markdown version of [/jobs/ext/2373024-site-reliability-engineer-spring](https://www.wearedevelopers.com/jobs/ext/2373024-site-reliability-engineer-spring). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer Spring - **Company:** IBM - **Location:** Winters, United States - **Contract:** Temporary contract - **Skills:** Amazon Web Services, Microsoft Azure, Bash Shell, System Configuration, Continuous Delivery, Continuous Integration, CouchDB, Linux, Monitoring of Systems, IBM Cloud Computing, Python (Programming Language), PostgreSQL, NoSQL, OpenShift, Redis, Reliability Engineering, Ansible, Prometheus, SQL Databases, Data Streaming, Scripting, Google Cloud, Grafana, Apache Spark, Break Fix, Kubernetes, Apache Kafka, Terraform, New Relic (SaaS), Jenkins - **Published:** August 14, 2026 - **Apply:** https://www.dice.com/job-detail/4537b589-31a6-4796-81a8-59d2d4c64cd1 ## About the Role This is a 16-week Co-Op position designed for actively enrolled undergraduate students who are interested in gaining hands-on, practical experience in their field while contributing to real IBM projects. To be eligible, candidates must be currently enrolled at an accredited university and able to pause their coursework for the duration of the Co-Op, as students may not be enrolled in classes while participating. This role requires full-time availability (40 hours per week) during the Spring 2027 semester between Jan 19 - May 7. Eligible candidates should also be available to join IBM in a full-time role between December 2027 and August 2028 upon successful completion of their degree., High School Diploma/GED Preferred education Bachelor's Degree Required technical and professional expertise, System Monitoring and Troubleshooting: knowledge in monitoring/observability, issue response, and troubleshooting for optimal system performance. Automation: knowledge in automation for production environment changes, streamlining processes for efficiency, and reducing toil. Linux: Knowledge of Linux operating systems. Operation and Support Experience: Understanding in handling day-to-day operations, alert management, incident support, migration tasks, and break-fix support. Scripting: knowledge or experience of Python, go or bash. Familiar with cloud providers like IBM Cloud, AWS, Azure or Google Cloud Platform Preferred technical and professional experience Kubernetes/OpenShift: knowledge or experience of Kubernetes/OpenShift environments. Automation/Scripting: knowledge or experience of Ansible, Python, Terraform, and CI/CD tools such as Jenkins, IBM Continuous Delivery, ArgoCD. Monitoring/Observability: knowledge or experience crafting alerts and dashboards using tools such as Instana, New Relic, Grafana/Prometheus., Being an IBMer means you'll be able to learn and develop yourself and your career, you'll be encouraged to be courageous and experiment everyday, all whilst having continuous trust and support in an environment where everyone can thrive whatever their personal or professional background. Our IBMers are growth minded, always staying curious, open to feedback and learning new information and skills to constantly transform themselves and our company. They are trusted to provide on-going feedback to help other IBMers grow, as well as collaborate with colleagues keeping in mind a team focused approach to include different perspectives to drive exceptional outcomes for our customers. The courage our IBMers have to make critical decisions everyday is essential to IBM becoming the catalyst for progress, always embracing challenges with resources they have to hand, a can-do attitude and always striving for an outcome focused approach within everything that they do. Are you ready to be an IBMer? ## Description As a Site Reliability Engineer, you will work in an agile, collaborative environment to build, deploy, configure, and maintain systems for the IBM client business. In this role, you will lead the problem resolution process for our clients, from analysis and troubleshooting, to deploying the latest software updates & fixes. Your primary responsibilities include: 24x7 Observability: Be part of a worldwide team that monitors the health of production systems and services around the clock, ensuring continuous reliability and optimal customer experience. Cross-Functional Troubleshooting: Collaborate with engineering teams to provide initial assessments and possible workarounds for production issues. Troubleshoot and resolve production issues effectively. Deployment and Configuration: Leverage Continuous Delivery (CI/CD) tools to deploy services and configuration changes at enterprise scale. Maintenance and Support: Tasks related to applying security patches and upgrades, and collaborating with Product support for issue resolution., DBA: Interest or experience configuring and maintaining SQL, NoSQL, and data streaming technologies (e.g. PostgreSQL, CouchDB, Redis, Kafka, Spark, etc.), Supplemental 1 employees may be eligible for up to 8 paid holidays, minimum of 56 hours paid sick time and the IBM Employee Stock Purchase Plan. IBM offers paid family medical leave and disability benefits to eligible employees where required by applicable law. This position was posted on the date cited in the key job details section and is anticipated to remain posted for 15 days from this date or less if not needed to fill the role. We consider qualified applicants with criminal histories, consistent with applicable law. The compensation range and benefits for this position are based on a full-time schedule for a full calendar year. The salary will vary depending on your job-related skills, experience and location. Pay increment and frequency of pay will be in accordance with employment classification and applicable laws. For part time roles, your compensation and benefits will be adjusted to reflect your hours. Benefits may be pro-rated for those who start working during the calendar year. ## Related Videos - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Building with IBM Bob](https://www.wearedevelopers.com/videos/100275-building-with-ibm-bob) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Can This Elephant Dance? IBM Bob and the Future of AI-First Software Development](https://www.wearedevelopers.com/videos/100256-can-this-elephant-dance-ibm-bob-and-the-future-of-ai-first-software-development) ## Related Articles - [Steps to Get a Software Engineer Internship](https://www.wearedevelopers.com/magazine/160-steps-to-get-a-software-engineer-internship) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [System change: restart as developer?](https://www.wearedevelopers.com/magazine/39-system-change-restart-as-developer) - [Top 6 Hackathons for Developers in 2023](https://www.wearedevelopers.com/magazine/263-top-6-hackathons-for-developers-in-2023) - [Best Coding Boot Camps in Germany](https://www.wearedevelopers.com/magazine/237-best-coding-boot-camps-in-germany)