Enterprise AI in Production: Mumbai Call for Proposals
Share what it takes to run AI inside an enterprise: architectures, trade-offs, and lessons from production.
Oct 2026
26 Mon
27 Tue
28 Wed
29 Thu
30 Fri
31 Sat 11:00 AM – 04:00 PM IST
1 Sun
Submitted Oct 2, 2026
While building and operating LLM-powered systems in production, one lesson keeps repeating itself which is that the hardest failures are often invisible.
Unlike traditional software, agentic systems can fail in countless ways while every dashboard remains green. An agent may misunderstand user intent, choose the wrong tool, retrieve stale information, drift from its objective, or silently regress after a model upgrade. In many cases, users discover the problem before the engineering team does.
In this talk, we will explore what it takes to truly understand, evaluate, and operate AI agents in production. We will examine why testing and monitoring agents differ fundamentally from traditional software and even classical ML systems, and introduce practical frameworks for identifying failures across comprehension, specification, and generalization.
Drawing from experiences building production agentic systems, we will cover the full evaluation lifecycle: component-level evaluations, workflow and end-to-end testing, synthetic data generation, observability and tracing, safety and alignment checks, human review processes, LLM-as-a-Judge systems, Agent-as-a-Judge architectures, and feedback loops that continuously improve agent behavior.
Along the way, we will discuss the strengths, limitations, and trade-offs of each approach, including how to handle subjective evaluations and measure quality at scale.
Anyone moving agentic systems from prototype to production and dealing with real-world failures
I’m a Solution Consultant at Sahaj Software. My work spans multi agentic systems and the practical applications of GenAI in engineering.
I enjoy exploring how AI technologies can augment human creativity and decision-making. My research has been presented at the International Conference on Data Analytics and Management, and I’ve spoken at multiple DevDays events and Fifth Elephant conferences, sharing insights on AI agents, AI-assisted software development, and emerging agentic architectures.
https://www.linkedin.com/in/swetha0302
{{ gettext('Login to leave a comment') }}
{{ gettext('Post a comment…') }}{{ errorMsg }}
{{ gettext('No comments posted yet') }}