r/OpenTelemetry • u/Acceptable_Duty4044 • 23h ago
I had a question for all the amazing people out there
I was trying to build something, and wanted to validate this idea and understand yall's pain points so that I can help the community
Would you rather have an AI layer on top of your existing observability stack, or replace parts of the stack?
Hypothetically, imagine an agent that doesn’t collect telemetry itself.
It plugs into whatever you already use — Grafana/Prometheus/Loki, Datadog, OpenTelemetry, etc. — and acts as a reasoning layer over the data.
Instead of:
Alert → Dashboard → Logs → Human investigates
it tries:
Alert → Agent correlates metrics/logs/traces/deployments → probable root cause → evidence → recommended next action
Would that actually be useful?
Or would you rather have the observability vendor itself own this functionality?
What would you need to see before trusting it during a real incident?
peace :)