Observability Cloud) Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+2 more
Job description
We are seeking an experienced Observability Engineer to support a strategic
observability platform migration initiative.
This role will be responsible for playing a core part in executing the migration of observability artifacts: SLOs, dashboards, alerts, and
monitoring workflows from Honeycomb to Splunk Observability Cloud., * Assist in the migration of observability assets and telemetry workflows from
Honeycomb to Splunk Observability Cloud.
- Assess, document, and translate existing dashboards, alerts, SLOs, and other
monitoring configurations.
- Design and implement observability solutions utilizing OpenTelemetry standards
and best practices.
- Partner with enterprise engineering, SRE, and platform teams to validate
telemetry quality, monitoring coverage, and operational readiness.
- Identify opportunities to improve observability maturity, reliability, and operational
efficiency during the migration.
- Develop documentation, knowledge transfer materials, and migration playbooks.
- Communicate progress, risks, dependencies, and recommendations to technical
and non-technical stakeholders., * The successful candidate will help ensure a seamless migration from Honeycomb ton Splunk Observability Cloud while maintaining observability coverage, improving telemetry quality, and enabling engineering teams to effectively monitor and support production systems.
- This role offers an opportunity to drive a highly visible observability transformation initiative and help establish the foundation for future reliability and operational excellence efforts.
Requirements
The ideal candidate combines deep technical expertise with robust communication and
collaboration skills and has hands-on experience designing, implementing, and
operating modern observability solutions using OpenTelemetry, Honeycomb, and
Splunk Observability Cloud., * Robust technical, analytical, and communication skills.
- Experience working in Site Reliability Engineering (SRE), Observability
Engineering, Platform Engineering, or a related discipline.
- Hands-on experience with Splunk Observability Cloud, including dashboards,
detectors, APM, infrastructure monitoring, and related capabilities.
- Hands-on experience with Honeycomb, including telemetry analysis,
dashboards, and operational workflows.
- Experience implementing and supporting OpenTelemetry instrumentation,
collection, and telemetry pipelines.
- Experience working with distributed systems, cloud-native applications, APIs,
and microservices environments.
- Ability to collaborate effectively across engineering, operations, and leadership
teams.
- Experience documenting technical solutions and communicating complex
- concepts to diverse audiences.
Preferred Qualifications
- Experience leading or participating in enterprise-scale observability platform
migrations.
- Experience with cloud platforms such as AWS, Azure, or Google Cloud Platform.
- Familiarity with CI/CD pipelines and automated deployment practices.
- Knowledge of modern monitoring, tracing, logging, and metrics best practices.
- Experience supporting production environments with high availability and
reliability requirements.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Find a Developer Job: 12 Best Job Sites For Developers
Fully Remote Software Engineer Jobs
Where To Find Software Engineering Jobs
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again