K2

The Morning PaperWednesday, August 12

AI models can't tell when a hacker put words in their mouth

AI models can't reliably tell when hackers inject fake instructions into their prompts—and training them to recognize it often backfires.

~70s readRead today’s paper. Start a streak.

Trending now

Same paper. Your language.

Don’t just read research. Compound it.