Job Description
Role Purpose
The purpose of this role is to support delivery through development and deployment of tools.
͏
Elastic AIOps Engineer – L3 Job Description
Job Title: Elastic AIOps Engineer – L3
Experience: 6–10 Years
Practice: Observability / BSM Tools / SRE
Primary Skills
- Elastic Observability
- Elastic AIOps
- Elasticsearch
- Kibana
- Log Analytics
- APM (Application Performance Monitoring)
- Machine Learning & Anomaly Detection
- Incident Management
Key Responsibilities
- Manage and support Elastic AIOps and Observability platforms in a 24x7 enterprise environment.
- Build and maintain dashboards, visualizations, alerts, and reports using Kibana.
- Configure and optimize Elasticsearch clusters, indices, shards, and data retention policies.
- Implement and manage AIOps capabilities including anomaly detection, root cause analysis, event correlation, and predictive analytics.
- Analyze logs, metrics, traces, and events to identify performance bottlenecks and operational issues.
- Troubleshoot Elasticsearch cluster health, performance, indexing, and query-related issues.
- Integrate Elastic with enterprise monitoring, ITSM, and automation platforms.
- Support incident, problem, and change management activities.
- Create operational runbooks, SOPs, and knowledge articles.
- Collaborate with application, infrastructure, cloud, and security teams for issue resolution.
Technical Skills
- Strong hands-on experience with Elasticsearch, Kibana, Logstash, and Beats.
- Experience with Elastic Observability and AIOps features.
- Knowledge of Machine Learning-based anomaly detection and event correlation.
- Experience with Linux administration and shell scripting.
- Understanding of distributed systems and cluster management.
- Exposure to cloud platforms (AWS, Azure, GCP) is preferred.
- Knowledge of APIs, JSON, REST services, and automation tools.
Secondary Skills
- ITIL Foundation
- DevOps/SRE practices
- Splunk, Dynatrace, Datadog, Grafana, or other observability tools
- Python/Shell scripting
Roles & Responsibilities (L3 Level)
- Handle complex incidents and perform advanced troubleshooting.
- Lead root cause analysis (RCA) and problem management activities.
- Drive platform optimization and automation initiatives.
- Mentor L1/L2 engineers and provide technical guidance.
- Ensure platform availability, performance, and compliance with SLAs.
͏
Deliver
| No. | Performance Parameter | Measure |
| 1. | Tool Development and deployment | Quality of solution Timely development and within budget Timely deployment of tool Error free deployment |
͏
͏
Experience: 5-8 Years .
Reinvent your world. We are building a modern Wipro. We are an end-to-end digital transformation partner with the boldest ambitions. To realize them, we need people inspired by reinvention. Of yourself, your career, and your skills. We want to see the constant evolution of our business and our industry. It has always been in our DNA - as the world around us changes, so do we. Join a business powered by purpose and a place that empowers you to design your own reinvention.