> Markdown version of [/jobs/ext/2104378-observability-sre-sme](https://www.wearedevelopers.com/jobs/ext/2104378-observability-sre-sme). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Observability SRE/SME - **Company:** Argyll Info Tech Inc., - **Location:** Dallas, TX, United States - **Salary:** $122,000.0 - $166,000.0 - **Contract:** Temporary to permanent - **Skills:** Amazon Web Services, Application Performance Management, Continuous Integration, Distributed Systems, Python (Programming Language), Reliability Engineering, Prometheus, Systems Integration, Datadog, Scripting, Cloud Platform System, Performance Testing, Delivery Pipeline, Grafana, Terraform, Splunk, Appdynamics, Dynatrace - **Published:** August 18, 2026 - **Apply:** https://www.careerjet.com/jobad/use662419f1c8a3050bc36951599864981 ## About the Role Has been done/been a part of a Grafana/Observability implementation SRE experience (understand how operation works), well versed with Splunk, Grafana and implement telemetry with Grafana Versed in scripting and automation Dashboarding with Grafana Good understanding of Open Telemetry, Extensive experience with Grafana dashboard design and implementation Proven track record in application performance monitoring (APM) and observability tooling Hands-on expertise in log tracing, root cause analysis, and troubleshooting distributed applications Strong knowledge of cloud platform integrations, especially with AWS environments Proficiency in using tools such as Prometheus, OpenTelemetry, Dynatrace, AppDynamics, Datadog, and Splunk Experience with infrastructure as code (Terraform, scripting with Python and shell) for deploying observability solutions Ability to configure SLO-based alerting and optimize observability stacks Nice to Have Skills: Experience with performance testing and load injection Familiarity with automation pipelines and CI/CD integrations for monitoring tools Knowledge of complex, regulated environments such as financial or government sectors Preferred Education and Experience: Bachelor's degree in a technical field or related discipline Prior hands-on roles focused on observability, site reliability engineering, or performance engineering in similar environments. Other Requirements: Ability to meet onsite requirements (3 days per week) at designated locations (Jersey City, NJ or Dallas) Availability to start immediately or as soon as possible Must have valid, redacted photo ID attached to profile for submission Candidates should be authorized to work in the US (USC sponsorship is necessary) ## Description Join a leading organization as a Grafana Technical SME, where your expertise will be pivotal in designing, implementing, and optimizing observability solutions. Bringing deep knowledge in Grafana, log tracing, and application performance monitoring, you will help elevate the company's monitoring infrastructure, ensuring robust system performance and reliability in a complex environment. This role offers the chance to work on impactful projects in a collaborative, innovative setting, blending onsite engagement with flexible work arrangements. ## Related Videos - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Software Engineering Social Connection: Yubo’s lean approach to scaling an 80M-user infrastructure](https://www.wearedevelopers.com/videos/1583-software-engineering-social-connection-yubo-s-lean-approach-to-scaling-an-80m-user-infrastructure) - [Keycloak case study: Making users happy with service level indicators and observability](https://www.wearedevelopers.com/videos/1599-keycloak-case-study-making-users-happy-with-service-level-indicators-and-observability) ## Related Articles - [Effortlessly Scale Prometheus With The Telemetry Data Platform – And Keep your Grafana Dashboards, Too!](https://www.wearedevelopers.com/magazine/3-effortlessly-scale-prometheus-with-the-telemetry-data-platform-and-keep-your-grafana-dashboards-too) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023)