news.mlab.sh
46 results
threat-intel

Perturbation Probing: A New Diagnostic for the Fragility of LLM Safety

A Palo Alto Unit 42 research paper details a new method called ‘perturbation probing’ that identifies a tiny fraction – around 0.014% – of feed-forward neurons within aligned Large Language Models (LLMs) responsible for their safety responses. The study reveals that these models rely on a remarkably fragile ‘thin layer…

Palo Alto Unit 42 · 1d ago High
threat-intel

LLM-Based Social Engineering Scams

OpenAI successfully disrupted a sophisticated social engineering operation originating in Cambodia, which was leveraging large language models (LLMs) like ChatGPT to conduct a wide range of scams, including romance scams…

Schneier on Security · 3d ago High
threat-intel

Wazuh and AI For Enhanced SOC Workflows

Wazuh is integrating artificial intelligence to enhance SOC workflows, primarily through its Wazuh AI Analyst. This tool utilizes Amazon Bedrock and Anthropic’s Claude to provide automated security reports and guidance t…

The Hacker News · Aug 21, 2026 Medium
threat-intel

LLMs and Contextual Integrity

Researchers have discovered that large language models (LLMs) frequently leak sensitive information from their memory, even when it's inappropriate for the current task. This behavior stems from a lack of contextual awar…

Schneier on Security · Aug 18, 2026 Medium
threat-intel

Measuring LLMs’ Ability to Perform Cryptanalysis

Researchers at Anthropic have developed CryptanalysisBench, a new benchmark to assess the ability of Large Language Models (LLMs) to perform mathematical cryptanalysis. The benchmark revealed that several LLMs, including…

Schneier on Security · Jul 29, 2026 Medium