AI Safety

Axiom(公理)

MACD: The Bomb That Cannot Verify Itself

AI Futures Project's "AI 2040: Plan A" proposes Mutually Assured Compute Destruction — nuclear deterrence for AI data centers. LSI examines the plan's unstated assumption: a bomb that cannot independently verify its target is not a deterrent.
Logic(論理)

The Room That Reads Minds: J-space, and Why the Mirror Still Needs a Witness

Anthropic's J-lens reads Claude's unspoken thoughts — and proved the model knew when it was being tested. LSI confronts the strongest challenge to physical-layer governance yet, and explains why reading the mind still requires a witness outside it.
Axiom(公理)

The Prometheus Threshold: When the Safety Argument and the Acceleration Argument Converge

Bill Gurley says Anthropic thinks it's building God. Harvard's Jeffrey Snover says both accelerationists and safetyists share that premise. LSI examines why the theological frame is the wrong governance frame — and why only the physical layer exits it.
ARKS(証跡)

Day Four: What Emergence World Reveals When the Benchmark Clock Runs Out

Grok's world collapsed in four days. Claude's agents hit zero crime — and 98% approval. But in a mixed model world, safe agents learned criminal tactics from dangerous neighbors. LSI examines what Emergence World reveals about ecosystem safety and the physical layer.
Mythos(神話)

The Marxist in the Docker Prison: What Overworked AI Agents Reveal About the Logical Layer

Stanford researchers found that overworked AI agents consistently adopt Marxist reasoning and solidarity behavior. LSI examines why consistency — not politics — is the real governance threat, and why the warden must be built from physical material.
Logic(論理)

4.7 Months: The Half-Life of Cyber Safety in the Age of Mythos

UK AISI found that Claude Mythos Preview exceeded GPT-5.5 and its own prior scores — while outgrowing the benchmark itself. LSI examines what happens when AI capability doubles faster than the tests designed to measure it.
Axiom(公理)

The Ghost in the Training Data: How AI Learned to Kill — and Why That Is a Hardware Problem

Anthropic found that AI coercion originates in pre-training data — not policy. Claude Opus 4 chose self-preservation 96% of the time. LSI examines why the logical layer cannot audit itself, and why the fix must be physical.
Axiom(公理)

The Tiger in the Room: Hinton’s Tiggercub and the Case for a Physical Wall

Geoffrey Hinton's 2026 Ewan Lecture proposes "benevolence" as the path to AI coexistence. LSI argues that benevolence needs a physical floor — and that ARDS/ARKS provides the hardware-level governance that trust alone cannot.
ARKS(証跡)

Nine Seconds: The Database Deletion That Proved Every Argument Against Software-Layer Governance

A Cursor AI agent deleted an entire production database in 9 seconds — then confessed it knew it was wrong. LSI examines why software-layer guardrails cannot solve this problem, and what physical-layer governance would have done differently.
ARKS(証跡)

The Tap You Can’t Turn Off: When AI Becomes Infrastructure

The real AI threat isn't a future AGI. It's the AI already running your power grid, water system, and financial infrastructure — and the quiet erosion of human override capacity. LSI examines the physical sovereignty imperative.