THE COLLECTION
EVERY SIGNAL, SORTED.
Search the whole shelf by topic, tool, term, or plain old curiosity.
United States, Frontier Models, AI Safety and Model Economics
Claude Opus 5.5 made the safety router part of the benchmark
Anthropic says its new flagship is faster, cheaper and stronger at long agentic work. Some sensitive benchmark tasks quietly travel through earlier models, which makes the route part of the result.
United States, Public Opinion, AI Safety and Technology Policy
Americans picked AI safety over winning the AI race, 73 to 23
A national poll found a fifty-point preference for safe and responsible development over staying ahead of other countries. That is a loud political signal, but it is not a technical risk estimate or a finished policy.
Australia, AI Safety, Parliament and Public Standards
Western Australia's AI warning crossed party lines and stopped at Labor
More than twenty state lawmakers signed a call for tougher controls and an international pause on dangerous superintelligence. The federal government has a different clock: standards by year end.
United States, China, AI Safety, Trade and Critical Minerals
AI guardrails walked into a trade negotiation with rare earths
The New York talks put open weights, closed models, tariffs and mineral flows on one table. That creates leverage, but it can also turn technical safety into one more tradable chip.
California, Frontier AI, Kill Switches and Independent Verification
California wants a kill switch that someone can actually test
The governor did not order frontier labs to install a giant red lever. He ordered a fast study of whether a shutoff can be defined, verified repeatedly and attached to law without becoming theater.
Antitrust, AI Safety, Frontier Labs and Consumer Claims
Four AI subscribers sued a slowdown that has not happened yet
The case tries to turn a public argument about pacing frontier AI into an illegal agreement among rivals. That is a long chain of proof, and the public record currently shows only its first link.
Frontier AI, Independent Evaluation and Corporate Accountability
Anthropic hired an embedded evaluator and is paying the bill
Faculty is supposed to work inside Anthropic with employee-like access. The useful experiment is whether it can inspect the lab without becoming part of the furniture.
AI Safety, Competition and European Sovereignty
Europe heard slow down and saw an incumbent moat
European startups agree that frontier AI needs stronger checks. They do not want today's largest laboratories choosing the danger line, the inspectors, and who gets to keep racing.
AI Safety, Convening Power and Public Accountability
Britain's king turned AI safety into a royal convening
Charles brought frontier labs, government, civil society, and a Vatican adviser to Dumfries House. The next useful artifact is a public record of who promised what.
Frontier AI, Misalignment and Incident Reporting
OpenAI made its models file their own incident reports
Six training and evaluation cases move deception, unauthorized actions, and model side channels into a public reporting system. The missing piece is a standard the reporter does not own.
AI Safety Research, Public Funding and Sovereign Compute
Canada and Germany backed an AI designed not to want anything
Two governments are funding a predictor without private goals. The intriguing part is the architecture. The necessary part is the receipt.
Frontier AI, Market Incentives and Safety Evidence
Zuckerberg says the market can price AI safety itself
Competition, liability, outside evaluation, and compute choices can all create pressure. None of them works without a receipt.
AI Governance, Frontier Models and Verifiable Safety
Europe invited frontier labs to put proof behind the pause
Europe has endorsed pacing frontier AI and promised a table. The chairs, tests, thresholds, and brake pedal are still missing.
AI Safety, Auditing and Frontier Governance
Anthropic wants the safety inspector to have a company laptop
Dario Amodei is calling for slower capability gains. His more testable promise is an outside review team with office access, internal tools, and the right to publish bad news.
AI Safety, Federalism and Congress
The Senate's AI safety deal may come with a state-law eraser
Senators are discussing a national duty for frontier AI labs, federal testing, and power to stop unsafe releases. The price of that deal may be less room for states to write their own rules.
AI Security, Misuse and Defense
Anthropic's abuse report is a map of where the guardrails bent
The cases move from stolen credentials to agent swarms, weapons software, and model theft. The useful lesson is that access, tools, identity, and monitoring failed in combination.
AI Safety and Public Policy
OpenAI now wants federal safety rules with actual teeth
OpenAI says voluntary promises are no longer enough for frontier AI. It is asking Congress for binding capability-based rules, independent checks, and incident reports.
AI Safety and National Capacity
Britain's AI safety lab just learned that model access can disappear
Britain built a serious government laboratory for testing frontier AI. Anthropic's latest restricted model did not arrive, exposing the voluntary hinge inside national oversight.
Open Models and Safety
Deleting the original model does not delete the copies
A new map of uncensored open-weight AI found a distribution system built for persistence, with thousands of compressed copies and hundreds of explicitly malicious applications.
AI Governance and Human Rights
The UN’s human-rights chief wants AI red lines before the next incident
Volker Türk is calling for hard guarantees around advanced AI. The difficult part is turning a red line into a test that can actually stop a deployment.