Coercing LLMs to Do and Reveal (Almost) Anything with Jonas Geiping - #678

FromThe TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Start listening View podcast show

Coercing LLMs to Do and Reveal (Almost) Anything with Jonas Geiping - #678

FromThe TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

ratings:

Length:

48 minutes

Released:

Apr 1, 2024

Format:

Podcast episode

Description

Today we're joined by Jonas Geiping, a research group leader at the ELLIS Institute, to explore his paper: "Coercing LLMs to Do and Reveal (Almost) Anything". Jonas explains how neural networks can be exploited, highlighting the risk of deploying LLM agents that interact with the real world. We discuss the role of open models in enabling security research, the challenges of optimizing over certain constraints, and the ongoing difficulties in achieving robustness in neural networks. Finally, we delve into the future of AI security, and the need for a better approach to mitigate the risks posed by optimized adversarial attacks.

The complete show notes for this episode can be found at twimlai.com/go/678.

Released:

Apr 1, 2024

Format:

Podcast episode

Titles in the series (100)

This Week in Machine Learning & AI is the most popular podcast of its kind. TWiML & AI caters to a highly-targeted audience of machine learning & AI enthusiasts. They are data scientists, developers, founders, CTOs, engineers, architects, IT & product leaders, as well as tech-savvy business leaders. These creators, builders, makers and influencers value TWiML as an authentic, trusted and insightful guide to all that’s interesting and important in the world of machine learning and AI. Technologies covered include: machine learning, artificial intelligence, deep learning, natural language processing, neural networks, analytics, deep learning and more.

Skip carousel

More Episodes from The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Skip carousel

Related podcast episodes

Skip carousel

Discover this podcast and so much more

Coercing LLMs to Do and Reveal (Almost) Anything with Jonas Geiping - #678

Coercing LLMs to Do and Reveal (Almost) Anything with Jonas Geiping - #678

Description

Titles in the series (100)

More Episodes from The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Related podcast episodes