Sveriges mest populära poddar

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Coercing LLMs to Do and Reveal (Almost) Anything with Jonas Geiping - #678

48 min•1 april 2024

Today we're joined by Jonas Geiping, a research group leader at the ELLIS Institute, to explore his paper: "Coercing LLMs to Do and Reveal (Almost) Anything". Jonas explains how neural networks can be exploited, highlighting the risk of deploying LLM agents that interact with the real world. We discuss the role of open models in enabling security research, the challenges of optimizing over certain constraints, and the ongoing difficulties in achieving robustness in neural networks. Finally, we delve into the future of AI security, and the need for a better approach to mitigate the risks posed by optimized adversarial attacks.

The complete show notes for this episode can be found at twimlai.com/go/678.

Fler avsnitt av The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Why AI Agents Break the GenAI Security Model with Devvret Rishi - #770

16 juni•56 min

Is RAG Dead? Lessons from Building AI for Tax Law with Alex Bowcut - #769

9 juni•52 min

Relational Foundation Models for Enterprise Data with Jure Leskovec - #768

21 maj•1 tim 6 min

How to Find the Agent Failures Your Evals Miss with Scott Clark - #767

7 maj•53 min

How to Engineer AI Inference Systems with Philip Kiely - #766

30 apr.•55 min

How Capital One Delivers Multi-Agent Systems with Rashmi Shetty - #765

16 apr.•54 min

The Race to Production-Grade Diffusion LLMs with Stefano Ermon - #764

26 mars•1 tim 3 min

Agent Swarms and Knowledge Graphs for Autonomous Software Development with Siddhant Pardeshi - #763

10 mars•1 tim 16 min

AI Trends 2026: OpenClaw Agents, Reasoning LLMs, and More with Sebastian Raschka - #762

26 feb.•1 tim 19 min

The Evolution of Reasoning in Small Language Models with Yejin Choi - #761

29 jan.•1 tim 6 min

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) med Sam Charrington finns tillgänglig på flera plattformar. Informationen på denna sida kommer från offentliga podd-flöden.