- Workplace
- Remote, Hybrid, Onsite
- Type
- Contract
- Source
- RecruiterFlow
Description
NJ. Details are below
The key here is we need people who aren’t just users but someone who knows how to troubleshoot, implement and understand how to fix issues in production. We need candidates focused on the infrastructure side of things not the developer side. Looking for strong SREs using open telemetry to implement Grafana.
Overview:
Join a leading organization as a Grafana Technical SME, where your expertise will be pivotal in designing, implementing, and optimizing observability solutions. Bringing deep knowledge in Grafana, log tracing, and application performance monitoring, you will help elevate the company's monitoring infrastructure, ensuring robust system performance and reliability in a complex environment. This role offers the chance to work on impactful projects in a collaborative, innovative setting, blending onsite engagement with flexible work arrangements.
Required Skills:
- Extensive experience with Grafana dashboard design and implementation. Strong experience with Grafana Cloud, migrating to the cloud.
- Strong Python scripting and automation experience
- Proven track record in application performance monitoring (APM) and observability tooling
- Hands-on expertise in log tracing, root cause analysis, and troubleshooting distributed applications
- Strong knowledge of cloud platform integrations, especially with AWS environments
- Proficiency in using tools such as Prometheus, OpenTelemetry, Dynatrace, AppDynamics, Datadog, and Splunk
- Experience with infrastructure as code (Terraform, scripting with Python and shell) for deploying observability solutions
- Ability to configure SLO-based alerting and optimize observability stacks