Site Reliability Engineer

TXSE Group Inc.
Dallas, United States
3 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services User Authentication Microsoft Azure Cloud Computing Continuous Integration Linux Distributed Systems Domain Name System (DNS) Github Monitoring of Systems Python (Programming Language) Linux Commands
+19 more
Octopus Deploy OpenShift Dell PowerEdge Reliability Engineering Ansible Prometheus Software Engineering SonarQube Grafana Multi-Cloud Kubernetes Information Technology Low Latency Github Enterprise Performance Monitor Data Management Terraform Jenkins Artifactory

Job description

TXSE is building the next-generation exchange infrastructure to support transparent, efficient, and resilient capital markets. With SEC approval and $275MM in funding, we are currently hiring a Site Reliability Engineer to help with a greenfield infrastructure build out., * Proficiency supporting Linux Systems, general Linux Sys Administration, and Linux command line expertise.

  • Monitoring & Incident Response: Monitor platform and infrastructure health, respond to incidents, troubleshoot service issues, and assist with root cause analysis to reduce downtime and operational risk. Exposure to Grafana, Prometheus or similar monitoring systems.
  • Cloud Operations: Support workloads in AWS and Azure, including troubleshooting issues related to compute, networking, storage, access, and security in hybrid and multi-cloud environments.
  • Enterprise Hardware & Storage Support: Support infrastructure running on Dell PowerEdge servers and assist in maintaining enterprise storage platforms, including performance monitoring, replication support, and capacity tracking.
  • Experience with both on-prem and cloud infrastructure
  • Experience with CI/CD and developer platform tools such as Jenkins, ArgoCD, GitHub Enterprise, SonarQube, Artifactory, or GitHub Actions runners.

Requirements

  • Proficiency in automation and scripting tools (Ansible and Terraform are required as they are a large part of our environment)., * Red Hat OpenShift and/or Kubernetes in production or non-production environments.
  • Multi-tenant cloud experience
  • Systems/Software Development experience in Python
  • Authentication services, DNS, and time synchronization concepts in distributed environments.
  • Experience with enterprise server hardware and storage platforms is a plus
  • Experience working in low latency environments, start-ups, prior experience within exchanges/trading firms
  • Ability to context switch, and work highly effective in autonomous environments.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

Videos

See all

Related articles

See all