Observability SRE/SME
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+6 more
Job description
Join a leading organization as a Grafana Technical SME, where your expertise will be pivotal in designing, implementing, and optimizing observability solutions. Bringing deep knowledge in Grafana, log tracing, and application performance monitoring, you will help elevate the company’s monitoring infrastructure, ensuring robust system performance and reliability in a complex environment. This role offers the chance to work on impactful projects in a collaborative, innovative setting, blending onsite engagement with flexible work arrangements.
Requirements
Has been done/been a part of a Grafana/Observability implementation SRE experience (understand how operation works), well versed with Splunk, Grafana and implement telemetry with Grafana Versed in scripting and automation Dashboarding with Grafana Good understanding of Open Telemetry, Extensive experience with Grafana dashboard design and implementation Proven track record in application performance monitoring (APM) and observability tooling Hands-on expertise in log tracing, root cause analysis, and troubleshooting distributed applications Strong knowledge of cloud platform integrations, especially with AWS environments Proficiency in using tools such as Prometheus, OpenTelemetry, Dynatrace, AppDynamics, Datadog, and Splunk Experience with infrastructure as code (Terraform, scripting with Python and shell) for deploying observability solutions Ability to configure SLO-based alerting and optimize observability stacks Nice to Have Skills: Experience with performance testing and load injection Familiarity with automation pipelines and CI/CD integrations for monitoring tools Knowledge of complex, regulated environments such as financial or government sectors Preferred Education and Experience: Bachelor’s degree in a technical field or related discipline Prior hands-on roles focused on observability, site reliability engineering, or performance engineering in similar environments. Other Requirements: Ability to meet onsite requirements (3 days per week) at designated locations (Jersey City, NJ or Dallas) Availability to start immediately or as soon as possible Must have valid, redacted photo ID attached to profile for submission Candidates should be authorized to work in the US (USC sponsorship is necessary)
Benefits & conditions
-
$122,000-166,000 per year Do you want your voice heard and your actions to count? Discover your opportunity with Mitsubishi UFJ Financial Group (MUFG), one of the world’s leading financial groups. Across …
-
22 days ago +
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Is Software Engineering Over-Saturated?
Fully Remote Software Engineer Jobs
Dev Digest 120 - Apple and peers