Grafana enginner

Diverse Lynx LLC
Morristown, NJ, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$90,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Application Performance Management Microsoft Azure Bash Shell Cloud Computing Program Optimization Data Security DevOps Elasticsearch Monitoring of Systems Python (Programming Language) Lynx
+25 more
Windows PowerShell Reliability Engineering Ansible Prometheus SQL Databases Systems Integration Virtual Machines Datadog Data Logging Scripting Google Cloud Enterprise Software Applications Cloud Monitoring System Availability Grafana Kubernetes Infrastructure Automation Frameworks Influxdb Performance Monitor Cloudwatch Terraform Splunk New Relic (SaaS) Appdynamics Dynatrace

Job description

  • Design, develop, and maintain Grafana dashboards, visualizations, and reports for infrastructure, applications, and business metrics.
  • Configure and manage Grafana Alerting for proactive monitoring and incident response.
  • Integrate Grafana with various data sources such as:

o Prometheus o Loki o Elasticsearch/OpenSearch o InfluxDB o SQL databases o Cloud monitoring platforms (AWS CloudWatch, Azure Monitor, Google Cloud Operations)

  • Develop observability solutions covering metrics, logs, and traces.
  • Implement monitoring strategies for Kubernetes, containers, virtual machines, cloud infrastructure, and enterprise applications.
  • Collaborate with SRE, DevOps, Application Support, and Platform Engineering teams to define monitoring requirements.
  • Automate dashboard deployment and configuration using Infrastructure as Code (IaC) tools.
  • Tune monitoring systems to minimize alert fatigue and improve operational efficiency.
  • Perform root cause analysis using monitoring and logging data.
  • Create and maintain technical documentation, monitoring standards, and operational runbooks.
  • Support capacity planning, performance analysis, and system optimization initiatives.
  • Ensure security and governance for monitoring platforms and data access.

Requirements

We are seeking a skilled Grafana Engineer to design, implement, and maintain enterprise monitoring and observability solutions using Grafana and related technologies. The ideal candidate will have hands-on experience in building dashboards, configuring alerts, integrating multiple data sources, and supporting cloud-native environments. The role involves collaborating with DevOps, SRE, Infrastructure, and Application teams to ensure high availability, performance, and reliability of business-critical applications and platforms., * Strong experience with Grafana dashboard development and administration.

  • Expertise in Grafana Alerting, notification channels, and alert rule management.
  • Experience with Grafana Enterprise is a plus.
  • Experience with:

Prometheus, Loki, Tempo, OpenTelemetry, Elasticsearch/OpenSearch, Splunk (preferred)

  • Understanding of Metrics, Logs, and Distributed Tracing concepts.
  • Experience with one or more cloud platforms: AWS, Azure and GCP
  • Familiarity with Kubernetes and container orchestration.
  • Knowl edge of CI/CD pipelines and DevOps practices.
  • Proficiency in scripting languages such as:

o Python o Bash/Shell o PowerShell

  • Experience with Terraform, Ansible, or similar automation tools.
  • Experience querying and analyzing monitoring data.
  • Strong analytical and troubleshooting skills.
  • Excellent communication and stakeholder management.
  • Ability to work independently and collaboratively.
  • Problem-solving mindset with attention to detail.
  • Strong documentation and knowledge-sharing capabilities.
  • Grafana Enterprise deployment experience.
  • OpenTelemetry implementation experience.
  • Experience with AIOps and observability platforms.
  • Exposure to application performance monitoring (APM) tools such as Dynatrace, AppDynamics, Datadog, or New Relic.

Keywords: Grafana, Prometheus, Loki, Tempo, OpenTelemetry, Kubernetes, Cloud Monitoring, Observability, SRE, DevOps, Terraform, AWS, Azure, Monitoring, Alerting, PromQL, LogQL. Diverse Lynx LLC is an Equal Employment Opportunity employer. All qualified applicants will receive due consideration for employment without any discrimination. All applicants will be evaluated solely on the basis of their ability, competence and their proven capability to perform the functions outlined in the corresponding role. We promote and support a diverse workforce across all levels in the company.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role β€” technically off-topic, practically not.

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet Β· LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 Β· LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum Β· WWC Europe 2026

3:57 min

Introduction to the Lynx cross-platform UI framework

Xuan Huang Xuan Huang Β· WWC 2025

28:32 min

Configuring Prometheus remote write and connecting Grafana dashboards

Liam Hurrell Β· LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 Β· LIVE

Videos

See all

Related articles

See all