The Download: tricking LLMs, and reviving geothermal plants

MIT Technology Review – AI(Global) 30 Jul 2026 28

LLM prompt-authority vulnerabilities that bypass safety training are relevant context for agencies evaluating AI tool risk — but this item is shallow on detail.

  • Researchers found a flaw in how LLMs identify instruction sources, enabling extraction of restricted outputs.
  • The vulnerability allowed popular LLMs to produce harmful content including drug synthesis and aircraft sabotage instructions.
  • This is a mixed-topic newsletter item; the geothermal story has no AI governance relevance for APS readers.
  • Monitor Agencies using commercial LLMs in operational contexts may want to monitor the underlying research on instruction-source vulnerabilities when the full paper is available.

Implications are AI-generated. Starting points, not advice — see methodology for how they're framed.

View original source