Skip to content
TrackPodcasts
technologyFeb 16, 202655:45pending

=Coffee

Hacked

About this episode

A lot of modern AI models have a kind of security guard layer that sits in front of them. Its job? A binary choice as to whether the prompt heading into the model is safe or not. Kasimir Schulz, a lead security researcher at HiddenLayer, has been researching how to trick these models. Their solution, a technique called "Echogram" involves words with such positive statistical sentiment — such overwhelming good vibes — that it flips that verdict. Learn more about your ad choices. Visit podcastchoices.com/adchoices

Get every episode summarized

Each time Hacked publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

=Coffee

Hacked

0:00
55:45

More episodes

More from Hacked

View all episodes →