TymblHub

© 2026 TymblHub

Observability Lead

Tranzeal
Posted on
Tranzeal logo

Experience
10 - 15 yrs
Salary (CTC)
₹25L - ₹40L
Job Location
Hyderabad, India
Vacancy
1
Designation
Implementation Lead
Job Type
Not specified

Job Description

Senior Observability/AIOps Automation Engineer or Architect

The client is looking for a Senior Observability/AIOps Engineer (10+ years) who can design and automate enterprise monitoring using New Relic, LogicMonitor, Terraform, and AIOps, replacing manual alert configuration with intelligent, automated observability solutions.

  • Experience: 10+ years
  • Location: Pune / PAN India (Hybrid)
  • Duration: 6+ months
  • Role Level: Senior Engineer / Lead / Architect

JD:

We have below requirement in Observability \ AIOps. Customer's primary need is automation in the Observability space. they are looking for someone who can lead automation of observability configuration and help eliminate manual alert configuration and automate alert aggregation instead of completing manual tasks.

We're looking for a candidate with ~ 10 + years of Experience and a strong background in Observability \ AIOps. He should have a proven track record of managing and optimizing observability tools, a deep understanding of standardization & automation, and the ability to work effectively across different teams to drive business value & outcome.

Skills -
Observability Platform Management: Provide expert operational support for a range of enterprise monitoring tools, including New Relic, Logic Monitor with a specific emphasis on Digital Employee Experience (DEX) platforms like Nexthink. Drive new features and capabilities by leading Proofs of Concept (POCs)
Observability-as-Code & Automation: Champion the 'Observability-as-Code' paradigm using terraform / equivalent by integrating monitoring configuration directly into CI/CD pipelines & automate the action
AIOps for Proactive Insights: Utilize AIOps to integrate machine learning and AI into our monitoring systems, automating the analysis of data to predict issues, automate incident detection and event correlation with focus how to reduce MTTR, increase SLA & shift mindset of observability
Incident & Alert Management: Lead incident management by creating, refining, and automating monitoring alerts to ensure proactive issues detection and minimize downtime.
Proactive Problem-Solving: Use Observability platform to proactively identify and resolve employee-impacting issues like slow logins, application crashes, and network latency before they escalate.
Strategic Collaboration & Enablement: Partner with business and development teams to understand their requirements and define comprehensive monitoring .Act as a key collaborator by providing compelling demos and presentations to internal teams, showcasing the value and insights gained from our observability & AIops.
Data-Driven Insights: Work with various teams to deliver best-in-class tools for visualizing Real User Monitoring (RUM), Synthetics, log data and performance data.
Analytics & Reporting: Develop comprehensive health and performance reports, create AIOps rules, design custom dashboards, and create business values out of Observability.



No Referrers Available

There are currently no referrers available for this job. You can still apply, will let you know once there is any referrer available.