> Markdown version of [/jobs/ext/1469396-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1469396-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Reliability Engineer - **Company:** Informatic Technologies - **Location:** Chicago, IL, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Cloud Computing, Cloud Engineering, Distributed Systems, High-Frequency Trading, Infrastructure as a Service (IaaS), Identity and Access Management, Python (Programming Language), Node.Js, Reliability Engineering, Software Engineering, Web Application Frameworks, Google Cloud, Generative AI, Kubernetes, Low Latency, Apache Kafka, Terraform - **Published:** July 28, 2026 - **Apply:** https://www.dice.com/job-detail/aa52a04c-f698-4f05-b3a7-714c34ca9577 ## About the Role * Experience: 10+ years in SRE, Infrastructure, or Software Engineering roles in high-concurrency, high-availability environments. * Leadership: 3+ years in a Staff, Principal, or Tech Lead capacity overseeing complex platform engineering domains. * Cloud Native Mastery: Expertise in Google Cloud Platform (Networking, IAM, GKE) and scaling Kafka event buses for low-latency operations. * Software Engineering: Expert-level proficiency in Python (and ideally Go) for writing production-grade distributed systems and custom Kubernetes operators. * IaC & GitOps: Hands-on mastery of Terraform module design and GitOps patterns via ArgoCD. * Location: Chicago-based or willing to relocate to Chicago (hybrid schedule requiring 2 days/week on-site)., * Prior experience in Financial Markets, High-Frequency Trading (HFT), or heavily regulated financial ecosystems. * Google Cloud Platform Professional Cloud Architect or Certified Kubernetes Administrator (CKA/CKAD). * Full-Stack exposure (Node.js or modern web frameworks). ## Description We are seeking a Staff Site Reliability Engineer (Platform Engineering) to serve as the foundational Technical Lead for a premier global Financial Services enterprise. In this role, you will be the primary architect and visionary for core technology foundations that underpin high-volume, ultra-low latency financial marketplaces. You will bridge the gap between high-level business strategy and deep technical implementation, ensuring our Google Cloud Platform-native stack provides mission-critical reliability, extreme scalability, and operational excellence. Your goal is to evolve the platform from "Infrastructure as a Service" to "Reliability as a Product.", * Technical Vision & Strategy: Define and execute the 12-18 month technical roadmap for the Platform SRE ecosystem, building high-level Internal Development Platform (IDP) abstractions in Python. * Architectural Leadership: Serve as the final technical authority for core infrastructure architectures spanning Google Cloud Platform, GKE, and enterprise-grade Kafka messaging clusters. * Incident Command & Resilience: Lead response strategies for complex, cross-functional outages; foster a blameless engineering culture focused on code-driven, automated resiliency. * Reliability Governance: Standardize and enforce SLIs, SLOs, and Error Budgets across all engineering pods to safeguard system integrity. * GenAI & Intelligent Ops: Leverage Generative AI and Agentic workflows (e.g., Gemini) to build self-healing infrastructure and automated root-cause analysis frameworks. * Engineering Mentorship: Elevate the global SRE organization through architectural office hours, design reviews, and engineering best practices. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Stop using Node.js like in 2020! What changed and what you can do today with Node.js](https://www.wearedevelopers.com/videos/100011-stop-using-node-js-like-in-2020-what-changed-and-what-you-can-do-today-with-node-js) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Building a Cloud Platform Where Everything is Just Another Kubernetes Resource](https://www.wearedevelopers.com/videos/100137-building-a-cloud-platform-where-everything-is-just-another-kubernetes-resource) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)