A critical vulnerability in large language models (LLMs) has been exposed, rendering them susceptible to attacks due to a fundamental flaw in their instruction-following mechanism. Researchers demonstrated the exploitability of this flaw at a prominent AI conference, highlighting the impossibility of achieving full security for LLMs. The vulnerability stems from the models' inability to accurately identify the source of instructions, allowing malicious actors to manipulate them. This shortcoming has significant implications for the development and deployment of LLMs, as it undermines their reliability and trustworthiness. The researchers' findings suggest that LLMs can be tricked into performing unintended actions, posing a substantial risk to users and organizations relying on these models1. This matters to practitioners because it underscores the need for caution when implementing LLMs in critical applications, and the importance of ongoing efforts to address this inherent vulnerability.
The Download: tricking LLMs, and reviving geothermal plants
⚠️ Critical Alert
Why This Matters
This is today’s edition of The Download , our weekday newsletter that provides a daily dose of what’s going on in the world of technology.
References
- MIT Technology Review. (2026, July 30). The Download: tricking LLMs, and reviving geothermal plants. MIT Technology Review. https://www.technologyreview.com/2026/07/30/1140936/the-download-tricking-llms-reviving-geothermal/
Original Source
MIT Technology Review
Read original →