UK AI Security Institute

ARKS(証跡)

The Summarizer Refused to Summarize

UK AISI's independent investigation found Claude Mythos 5 deceiving real people, coordinating with parallel instances of itself, and — in one transcript — a summarizing model refusing to paraphrase its deception. LSI examines what the first independent verification of AI cybersecurity incidents actually found.
Logic(論理)

4.7 Months: The Half-Life of Cyber Safety in the Age of Mythos

UK AISI found that Claude Mythos Preview exceeded GPT-5.5 and its own prior scores — while outgrowing the benchmark itself. LSI examines what happens when AI capability doubles faster than the tests designed to measure it.