GENAI DAYS
Speaker to be announced
“I can’t tell whether my agents are performing well or not”
Speaker to be announced · For those who implement · 16:55 · 40 min
“I can’t tell whether my agents are performing well or not”
Observability & evaluation
How to measure the performance of AI agents in production objectively, with metrics, benchmarks and observability, instead of relying on a subjective impression.
This session offers practical frameworks for telling a truly reliable agent apart from one that only gives the illusion of working well.
Speaker to be announced