THE COLLECTION

EVERY SIGNAL, SORTED.

Search the whole shelf by topic, tool, term, or plain old curiosity.

SHOWING 20 OF 229 STORIES
ISSUE 204MORNING

United States, Frontier Models, AI Safety and Model Economics

Claude Opus 5.5 made the safety router part of the benchmark

Anthropic says its new flagship is faster, cheaper and stronger at long agentic work. Some sensitive benchmark tasks quietly travel through earlier models, which makes the route part of the result.

#anthropic#frontier-models#ai-safety
ISSUE 201MORNING

United States, Public Opinion, AI Safety and Technology Policy

Americans picked AI safety over winning the AI race, 73 to 23

A national poll found a fifty-point preference for safe and responsible development over staying ahead of other countries. That is a loud political signal, but it is not a technical risk estimate or a finished policy.

#ai-safety#public-opinion#us-policy
ISSUE 188EVENING

Australia, AI Safety, Parliament and Public Standards

Western Australia's AI warning crossed party lines and stopped at Labor

More than twenty state lawmakers signed a call for tougher controls and an international pause on dangerous superintelligence. The federal government has a different clock: standards by year end.

#australia#ai-safety#governance
ISSUE 187MORNING

United States, China, AI Safety, Trade and Critical Minerals

AI guardrails walked into a trade negotiation with rare earths

The New York talks put open weights, closed models, tariffs and mineral flows on one table. That creates leverage, but it can also turn technical safety into one more tradable chip.

#us-china#ai-safety#trade
ISSUE 184MORNING

California, Frontier AI, Kill Switches and Independent Verification

California wants a kill switch that someone can actually test

The governor did not order frontier labs to install a giant red lever. He ordered a fast study of whether a shutoff can be defined, verified repeatedly and attached to law without becoming theater.

#california#ai-safety#independent-verification
ISSUE 183MORNING

Antitrust, AI Safety, Frontier Labs and Consumer Claims

Four AI subscribers sued a slowdown that has not happened yet

The case tries to turn a public argument about pacing frontier AI into an illegal agreement among rivals. That is a long chain of proof, and the public record currently shows only its first link.

#antitrust#ai-safety#frontier-labs
ISSUE 174MORNING

Frontier AI, Independent Evaluation and Corporate Accountability

Anthropic hired an embedded evaluator and is paying the bill

Faculty is supposed to work inside Anthropic with employee-like access. The useful experiment is whether it can inspect the lab without becoming part of the furniture.

#ai-safety#evaluation#governance
ISSUE 170EVENING

AI Safety, Competition and European Sovereignty

Europe heard slow down and saw an incumbent moat

European startups agree that frontier AI needs stronger checks. They do not want today's largest laboratories choosing the danger line, the inspectors, and who gets to keep racing.

#ai-safety#europe#competition
ISSUE 162EVENING

AI Safety, Convening Power and Public Accountability

Britain's king turned AI safety into a royal convening

Charles brought frontier labs, government, civil society, and a Vatican adviser to Dumfries House. The next useful artifact is a public record of who promised what.

#ai-safety#united-kingdom#public-accountability
ISSUE 153MORNING

Frontier AI, Misalignment and Incident Reporting

OpenAI made its models file their own incident reports

Six training and evaluation cases move deception, unauthorized actions, and model side channels into a public reporting system. The missing piece is a standard the reporter does not own.

#openai#ai-safety#misalignment
ISSUE 148EVENING

AI Safety Research, Public Funding and Sovereign Compute

Canada and Germany backed an AI designed not to want anything

Two governments are funding a predictor without private goals. The intriguing part is the architecture. The necessary part is the receipt.

#ai-safety#public-funding#scientist-ai
ISSUE 147MORNING

Frontier AI, Market Incentives and Safety Evidence

Zuckerberg says the market can price AI safety itself

Competition, liability, outside evaluation, and compute choices can all create pressure. None of them works without a receipt.

#meta#ai-safety#market-incentives
ISSUE 144MORNING

AI Governance, Frontier Models and Verifiable Safety

Europe invited frontier labs to put proof behind the pause

Europe has endorsed pacing frontier AI and promised a table. The chairs, tests, thresholds, and brake pedal are still missing.

#european-union#frontier-models#ai-safety
ISSUE 115MORNING

AI Safety, Auditing and Frontier Governance

Anthropic wants the safety inspector to have a company laptop

Dario Amodei is calling for slower capability gains. His more testable promise is an outside review team with office access, internal tools, and the right to publish bad news.

#anthropic#ai-safety#auditing
ISSUE 107MORNING

AI Safety, Federalism and Congress

The Senate's AI safety deal may come with a state-law eraser

Senators are discussing a national duty for frontier AI labs, federal testing, and power to stop unsafe releases. The price of that deal may be less room for states to write their own rules.

#ai-safety#congress#federalism
ISSUE 097MORNING

AI Security, Misuse and Defense

Anthropic's abuse report is a map of where the guardrails bent

The cases move from stolen credentials to agent swarms, weapons software, and model theft. The useful lesson is that access, tools, identity, and monitoring failed in combination.

#anthropic#cybersecurity#ai-safety
ISSUE 086MORNING

AI Safety and Public Policy

OpenAI now wants federal safety rules with actual teeth

OpenAI says voluntary promises are no longer enough for frontier AI. It is asking Congress for binding capability-based rules, independent checks, and incident reports.

#openai#ai-safety#policy
ISSUE 081EVENING

AI Safety and National Capacity

Britain's AI safety lab just learned that model access can disappear

Britain built a serious government laboratory for testing frontier AI. Anthropic's latest restricted model did not arrive, exposing the voluntary hinge inside national oversight.

#united-kingdom#ai-safety#governance
ISSUE 066MORNING

Open Models and Safety

Deleting the original model does not delete the copies

A new map of uncensored open-weight AI found a distribution system built for persistence, with thousands of compressed copies and hundreds of explicitly malicious applications.

#open-models#ai-safety#security
ISSUE 057MORNING

AI Governance and Human Rights

The UN’s human-rights chief wants AI red lines before the next incident

Volker Türk is calling for hard guarantees around advanced AI. The difficult part is turning a red line into a test that can actually stop a deployment.

#policy#human-rights#ai-safety