> Markdown version of [/jobs/ext/2976503-calypso-site-reliability-engineer-sre](https://www.wearedevelopers.com/jobs/ext/2976503-calypso-site-reliability-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Calypso Site Reliability Engineer (SRE) - **Company:** Quinnox Inc - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Application Release Automation, Bash Shell, Cloud Engineering, Cluster Analysis, Continuous Integration, DevOps, Disaster Recovery, Monitoring of Systems, Identity and Access Management, Java Virtual Machine (JVM), Python (Programming Language), Key Management, Networking Basics, Oracle (Applications), Performance Tuning, Windows PowerShell, Release Management, Reliability Engineering, Site Reliability Engineering Practices, Prometheus, Calypso Programming Language, SQL Databases, Data Logging, Scripting, Enterprise Software Applications, System Availability, Grafana, Reliability of Systems, Database Performance, Infrastructure as Code (IaC), Amazon Virtual Private Cloud (VPC), Gitlab, Cloudformation, Low Latency, Deployment Automation, Terraform, Splunk, Dynatrace - **Published:** September 18, 2026 - **Apply:** https://www.dice.com/job-detail/b49dfa38-fa56-4930-9c19-82f45b37fe12 ## About the Role * 6-10+ years of experience in Site Reliability Engineering / DevOps / Production Support in capital markets platforms/enterprise applications [experience with Calypso V17/V18 is added advantage] * Strong hands-on experience with: * Amazon Web Services (EC2, S3, networking, IAM, VPC) * GitLab CI/CD pipelines * Scripting: PowerShell, Bash/Shell [Python is added advantage] * Experience with: * Monitoring tools (e.g., ELK, Prometheus, Grafana, Splunk) * CI/CD and release automation * Infrastructure as Code (Terraform, CloudFormation - preferred) * Strong understanding of: * Linux/Unix systems * Networking fundamentals and cloud architecture * Basic Database concepts (Oracle/SQL) * Experience supporting high-availability, low-latency enterprise systems Roles & Responsibilities * Own reliability, availability, and performance of Calypso across production and non-production environments * Design, implement, and operate end-to-end SRE practices, including monitoring, alerting, incident management, and capacity planning * Build and manage CI/CD pipelines using GitLab, enabling automated build, deployment, and release of Calypso components * Automate deployment and environment provisioning on Amazon Web Services (AWS) using Infrastructure as Code (IaC) principles ## Description * Develop and maintain automation scripts using PowerShell, Shell (Bash), and Python for operational tasks, deployments, and monitoring * Ensure high availability and resiliency of Calypso services through failover strategies, clustering, and disaster recovery planning * Implement observability frameworks, including logging, metrics, and distributed tracing for proactive issue detection * Define and monitor SLOs/SLIs/SLAs, ensuring system performance meets business expectations in a trading environment * Lead incident management and root cause analysis (RCA), ensuring quick resolution of production issues and prevention of recurrence * Optimize system performance, including JVM tuning, database performance, and application-level optimizations for high-volume trade processing * Manage environment stability, including handling batch jobs, EOD processing, and trade lifecycle events in Calypso * Collaborate with development, QA, and infrastructure teams to ensure smooth releases and production readiness * Implement security best practices, including access controls, secrets management, and compliance with regulatory requirements * Support release management and deployment strategies, including blue-green deployments, canary releases, and rollback mechanisms * Drive continuous improvement and automation, reducing manual intervention and improving system reliability * Maintain runbooks, playbooks, and operational documentation for support and incident handling * Support production releases and provide hypercare support, ensuring system stability during critical business cycles ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [GitLab CI pipelines for a whole company](https://www.wearedevelopers.com/videos/143-gitlab-ci-pipelines-for-a-whole-company) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Shifting Stress to Progress— Understanding DevOps to do DevOps Better](https://www.wearedevelopers.com/videos/268-shifting-stress-to-progress-understanding-devops-to-do-devops-better) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Best Companies to work for in London: Top 25 Companies in 2023](https://www.wearedevelopers.com/magazine/187-best-companies-to-work-for-in-london-top-25-companies-in-2023) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023)