Grafana L3 Engineer / Administrator
Noida, Uttar Pradesh, India · Full Time
Be the first to apply
- Experience
- 6+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 6 days ago
- Work mode
- In office
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
Overview
We are seeking an experienced Grafana Level 3 Engineer / Administrator to oversee and maintain enterprise-grade Grafana environments. The successful candidate will play a critical role in configuring, optimizing, and supporting Grafana for robust observability, monitoring, and reporting solutions.
Key Responsibilities
- Manage, configure, and sustain both Grafana Enterprise and Open Source platforms.
- Create sophisticated dashboards for infrastructure, applications, cloud services, and business metrics monitoring.
- Integrate Grafana with diverse data sources including Prometheus, Loki, Tempo, Mimir, Elasticsearch, InfluxDB, MySQL, PostgreSQL, Azure Monitor, CloudWatch, and Splunk.
- Set up alert rules, notification pipelines, reporting mechanisms, and escalation procedures.
- Diagnose and resolve issues related to dashboards, alerts, integrations, and system performance providing root cause analyses.
- Improve platform scalability, availability, reliability, and overall performance.
- Manage security configurations including RBAC, LDAP, SSO, and OAuth authentication and governance.
- Automate administrative functions through Grafana APIs, Python or Bash scripting, and Infrastructure as Code tools.
- Handle software upgrades, migrations, backup/restoration, and disaster recovery processes.
- Collaborate closely with Site Reliability Engineering (SRE), DevOps, cloud teams, and application owners to strengthen observability and monitoring frameworks.
- Provide tool design, development, and deployment support aligned with project requirements and timelines.
- Engage with internal teams or clients to understand tool needs, plan solutions considering licensing and existing tools, and negotiate with third-party vendors.
- Conduct rigorous testing to guarantee error-free deployment, ensuring projects are delivered within budget and on schedule.
Required Qualifications and Skills
- Minimum 6 years’ IT experience with practical expertise in Grafana administration and monitoring platforms.
- Proven skills in installation, configuration, upgrade, and troubleshooting of Grafana Enterprise and Open Source editions.
- Deep knowledge of the LGTM stack components: Loki, Grafana, Tempo, Mimir, and Prometheus.
- Hands-on experience implementing observability solutions via OpenTelemetry.
- Advanced capabilities in dashboard development, alerting, and adopting monitoring best practices.
- Competence in integrating Grafana with multiple database systems and monitoring tools.
- Strong Linux system administration skills.
- Proficient in SQL and experience with database connectivity and integrations.
- Familiarity with monitoring services on AWS and Azure cloud platforms.
- Understanding of incident, event, capacity, and performance monitoring methodologies.
- Experience supporting large enterprise-scale monitoring systems.
- Scripting proficiency in Python or Bash for automation tasks.
Preferred Skills
- Experience working with Kubernetes, Docker, and cloud-native monitoring solutions.
- Exposure to CI/CD pipelines and Infrastructure as Code approaches.
- Certifications related to Grafana or observability platforms are advantageous.
Company Vision
We are a digitally transformative organization, committed to ongoing reinvention and innovation. Join us to contribute your expertise and accelerate your career by being part of a purpose-driven company that empowers its people and evolves with changing industry landscapes.