Sveriges mest populära poddar
MLOps.community

Are Evals Dead?

25 min26 september 2025

AI Conversations Powered by Prosus Group 


Your AI agent isn’t failing because it’s dumb—it’s failing because you refuse to test it. Chiara Caratelli cuts through the hype to show why evaluations—not bigger models or fancier prompts—decide whether agents succeed in the real world. If you’re not stress-testing, simulating, and iterating on failures, you’re not building AI—you’re shipping experiments disguised as products.


Guest speaker: Chiara Caratelli - Data Scientist @ Prosus Group

Host: Demetrios Brinkmann - Founder of MLOps Community


~~~~~~~~ ✌️Connect With Us ✌️ ~~~~~~~

Catch all episodes, blogs, newsletters, and more: https://go.mlops.community/TYExplore

Join our Slack community [https://go.mlops.community/slack]

Follow us on X/Twitter [@mlopscommunity](https://x.com/mlopscommunity) or [LinkedIn](https://go.mlops.community/linkedin)]

Sign up for the next meetup: [https://go.mlops.community/register]

MLOps Swag/Merch: [https://shop.mlops.community/]

MLOps.community med Demetrios finns tillgänglig på flera plattformar. Informationen på denna sida kommer från offentliga podd-flöden.