Sign In
Register

Partnership opportunities

Secure your pass

Call to action
Your text goes here. Insert your content, thoughts, or information in this space.
Button

Back to speakers

Vanchhit
Khare
Solutions Developer
M&T Bank
Vanchhit Khare is a Solutions Developer. He has written 8+ research papers and reviewed 40+ more. He won Best Paper at IEMTRONICS 2025. He teaches kids at TechBridge AI4ALL, who usually ask better questions than he does. All day he thinks about Claude Code. At night he thinks about it again. Then he sleeps, and his agents keep going without him, which is either very helpful or mildly creepy. He wakes up, reads what they found, and writes it in his diary. When he's not working he's on a plane somewhere with bad WiFi.
Button
26 August 2026 13:30 - 14:00
Simulating human emotion at scale: How swarm AI predicts what polls and sentiment tools cannot
Agentic AI introduces a new set of evaluation challenges that go well beyond traditional LLM benchmarking. When a model plans, calls tools, recovers from errors, and operates over long horizons, the question shifts from "is this answer correct?" to "did this trajectory accomplish the goal - safely, efficiently, and for the right reasons?" This talk surveys the practical challenges of evaluating agents in production: why static benchmarks saturate and mispredict real-world behavior, how path-dependence and multiple valid solutions complicate scoring, and why trajectory-level metrics (steps, tool calls, cost, latency) often matter as much as final-task success.