Ydhya
All roles

Engineering

Senior AIOps Engineer — Incident Intelligence

Build AI systems that reduce operational noise: event correlation, anomaly detection, runbooks, and incident triage.

Location

India / Global Remote

Workplace

Remote

Type

Full-time

Team

Engineering

Experience

4+ years

Compensation

Compensation discussed during process

Operational intelligence

Help teams see the signal before incidents spread.

AIOps work is valuable only when it fits the tools operators already use and reduces noise instead of adding another dashboard.

You will turn logs, alerts, incidents, tickets, and runbooks into workflows that help teams recover faster.

What you will do

01

Design pipelines for logs, metrics, events, alerts, tickets, and incident timelines

02

Build correlation, summarization, anomaly detection, and suggested-runbook workflows

03

Integrate with observability, ITSM, chat, and on-call systems

04

Define reliability checks that measure signal quality, false positives, and recovery impact

What we are looking for

01

4+ years in backend, platform, SRE, data engineering, observability, or AIOps work

02

Strong Python, data processing, APIs, and operational debugging skills

03

Experience with logs, metrics, tracing, incident management, or runbook automation

04

Practical understanding of LLM summarization, classification, and workflow orchestration

Apply

Send a clear note about what you have shipped.

Include the role, your resume, and links to work that show your judgment. Short, concrete notes are better than generic cover letters.

Apply now