> Markdown version of [/jobs/ext/2710422-infrastructure-support-engineer](https://www.wearedevelopers.com/jobs/ext/2710422-infrastructure-support-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Infrastructure Support Engineer - **Company:** ThoughtWorks - **Location:** Dallas, United States (Remote available) - **Experience:** Expert - **Salary:** $91,000.0 - $137,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, User Authentication, Microsoft Azure, Bash Shell, Continuous Integration, Debian Linux, Linux, Disaster Recovery, Failover, JSON, PostgreSQL, Memcached, MongoDB, Network Protocols, NoSQL, OpenShift, Performance Tuning, Red Hat Enterprise Linux, Redis, Reliability Engineering, Ansible, SQL Databases, Datadog, Circleci, SSL Certificate Management, Data Logging, Transport Layer Security, Caching, Infrastructure as Code (IaC), Backend, Gitlab, Cloudformation, Kubernetes, Terraform, Splunk, Docker, Jenkins - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/senior-infrastructure-support-engineer-referrals-only-8986728 ## About the Role * You have hands-on experience in using CI/CD tools such as Jenkins, CircleCI or Gitlab for executing deployments. * You have knowledge of Infrastructure as Code (IAC) tech stacks such as Terraform, Ansible, ARM or Cloudformation to provision and manage infrastructure. * You have working experience in using observability tools for logging, monitoring, tracing and alerting, e.g.: Datadog/PrometheusGrafana, ELK/EFK/Splunk. * You have experience in supporting at least one public cloud, e.g.: AWS, Azure or GCP. * You have hands-on experience executing most common operations in managing workloads on any container ecosystem tech stacks. e.g.: Docker, Kubernetes, Openshift, etc. * You understand system performance tuning and scaling to handle common heavy load scenarios along with concepts of highly available systems and basics of disaster recovery solutions, and are familiar with failover, backup and recovery concepts. * You have experience operating a Linux OS such as RHEL or a Debian-Based OS and are familiar with most common Linux OS operations and commands, reading and tweaking Bash scripts and managing runtime environment configurations such as Env Vars, Logs, etc. * You have experience supporting backend storage solutions such as SQL and NoSQL databases, e.g.: Postgres and MongoDB, and caching solutions such as Redis and Memcached. * You have experience in networking configuration and security, and are familiar with common networking setup and security practices, e.g.: loading, balancing, proxies, transport layer security (TLS) and certificate management, and an understanding of standard network protocols and configurations. * You have a good understanding of fundamental concepts of APIs such as request, response, headers, authentication, JSON payloads, etc. Professional Skills * You have strong communication and articulation skills, are proficient in English and able to confidently hold a Q&A discussion with senior stakeholders. * You have people skills with an emphasis on close collaboration with multiple, cross-functional teams from the client side or Thoughtworks. * You have the ability to work under pressure and with composure during production incidents. * You have strong analysis, deduction and reasoning skills, with the ability to identify patterns in data and draw conclusions. * You have strong drive and ownership to sign up and deliver work when called upon without being too concerned with role boundaries. * You are willing to be part of a rotation- and need-based 24x7 available team. ## Description * You will keep a vigilant eye on the operations of shipped products and services following the agreed upon "Eyes on glass/Follow the sun" engagement models. * You will monitor product/service operations against key performance indicators defined by the business and take necessary actions in response to detected deviations. * You will define and document the appropriate responses to various kinds of incident scenarios in collaboration with the Service Reliability Engineering (SRE) team and client stakeholders, and prepare runbooks. * You will reduce the human effort in day-to-day operations by automating operations, using the latest tech stacks befitting the task and improving the overall efficiency of the entire team as time progresses. * You will be the first responder to incidents in production/other high-value environments and execute the appropriate response as established by runbooks or based on your judgment of incidents. * You will initiate or establish communication to the support teams across all service functions, setting up war rooms for incident response, coordinating with tech leads, SRE leads and development teams to resolve incidents, as necessary. * You will prepare incident root cause analysis (RCA) and postmortem reports, explaining analyses and outlining preventive measures to clients; Collaborating with SRE, development teams or independently, your role is to ensure clear communication and proactive steps for future incident prevention. * You will implement service/product reliability improvement in collaboration with service reliability engineers by writing infrastructure/observability configuration code., There is no one-size-fits-all career path at Thoughtworks: however you want to develop your career is entirely up to you. But we also balance autonomy with the strength of our cultivation culture. This means your career is supported by interactive tools, numerous development programs and teammates who want to help you grow. We see value in helping each other be our best and that extends to empowering our employees in their career journeys. Responsible Use of AI in Recruitment At Thoughtworks, we use AI tools to support our recruitment team with administrative tasks such as drafting communications, scheduling interviews and writing job descriptions. Crucially, our AI tools do not screen, assess, rank or make hiring decisions. Every application is reviewed by our team and all selection decisions are made exclusively by our interviewers and hiring managers. We are committed to fairness and responsible AI. We actively manage our AI systems by testing, monitoring for biased outcomes and implementing mitigation measures. We hold our third-party vendors to these same high standards through a rigorous governance process. For additional information, please see our full Thoughtworks AI Policy for Recruitment., Thoughtworks is committed to providing reasonable accommodations to qualified applicants with disabilities or sincerely held religious beliefs, practices, or observances, in accordance with applicable law. If you need a reasonable accommodation to complete any part of the application process, participate in interviews, or otherwise engage in the hiring process, you may request an accommodation by completing this form or speaking with your recruiter. Requests may be made at any stage of the application or interview process. Once a request is received, Thoughtworks will engage in an interactive process with the applicant to determine an appropriate accommodation. Applicants are not required to disclose medical diagnoses or detailed personal information in order to request an accommodation. All accommodation requests will be handled in a timely, confidential, and respectful manner, consistent with applicable legal requirements. Requesting an accommodation will not negatively affect your consideration for employment. Company prohibits retaliation against any applicant for requesting an accommodation or participating in the accommodation process. Accommodations made during the recruitment process are not a guarantee of future or continued accommodations once hired. If you are hired by Thoughtworks, and require an accommodation to perform the essential functions of your role, you may be asked to engage in our reasonable accommodation process. ## Related Videos - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) - [Introducing JSON Structure](https://www.wearedevelopers.com/videos/100219-introducing-json-structure) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How to Answer the Interview Question: “Why Do You Want to Be a Software Engineer?”](https://www.wearedevelopers.com/magazine/392-how-to-answer-the-interview-question-why-do-you-want-to-be-a-software-engineer) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [7 Important Tips That Every Software Developer Should Know](https://www.wearedevelopers.com/magazine/101-7-important-tips-that-every-software-developer-should-know)