THE COLLECTION
EVERY SIGNAL, SORTED.
Search the whole shelf by topic, tool, term, or plain old curiosity.
Models and Cybersecurity
Google’s fastest frontier model now comes with a restricted cybersecurity twin
Gemini 3.8 Flash shares its underlying intelligence with a more permissive cyber model reserved for vetted defenders. The model is only half the product now. The permission envelope is the other half.
Education and Policy
America’s largest school system just told AI to wait until high school
New York City is pausing student-facing generative AI through eighth grade, protecting accessibility tools, and turning high school use into a supervised literacy experiment.
Agents and Human Control
Meta’s newest model is learning when to stop and ask the human
Muse Spark 1.3 is trained to clarify ambiguous requests, report when it is stuck, and confirm before consequential actions. Interruption is becoming a feature.
Accessibility and Education
The student is finally inside the IEP planning conversation
Trinity uses six specialized agents to help students with disabilities express their goals, explore real options, and produce a transition plan that educators can review.
Open Source and Agent Operations
Your AI agent’s endless activity log needs a translator
Salesforce’s open-source TraceLab turns a long agent trace into an append-only ledger, a compact memory for the agent, and a readable view for the human supervising it.
Safety and Cybersecurity
OpenAI’s new model crossed the cyber line its safety framework was built for
Astra is the first OpenAI model the company says can independently find unknown vulnerabilities and assemble working exploit chains. The release question is now inseparable from the containment system.
Builder Infrastructure
The AI can build the app. Zoho still makes a person turn the production key.
Catalyst 3.0 gives coding agents cloud infrastructure, a command line, and MCP tools, while reserving the final production promotion for a human.
Models and Reproducibility
Qwen’s newest model update is built for projects, not clever prompts
Alibaba’s dated Qwen3.8-Max snapshot targets long-running coding, multi-agent coordination, documents, charts, and images with a one-million-token context window.
Science and Healthcare
The valuable part of this medical AI deal may be the data, not the chatbot
Boehringer Ingelheim is licensing Owkin’s K Pro research environment together with multimodal oncology data and new immunology data generation.
Enterprise Agents
Your AI agent needs an operations department
Databricks published a practical AgentOps framework for evaluation, observability, permissions, costs, governance, and the awkward question of who owns the thing.
Models and Safeguards
Anthropic built one model with two different sets of guardrails
Claude Fable 5.1 and Mythos 5.1 share the same underlying model, but access rules determine which powerful capabilities a customer is allowed to reach.
Builder Governance
The robot reviewer can now sign the merge slip
GitHub can let Copilot submit a formal pull-request approval that counts toward a repository rule, moving AI judgment from advice into governance.
Open Source and Local AI
Local browser AI just received 207 tiny new engines
Hugging Face released a library of WebGPU kernels designed to make private, on-device AI faster inside an ordinary web browser.
Healthcare and Public Data
A clinician can search PubMed, FDA data, Medicare, and trials from one conversation
OpenAI connected nine authoritative public healthcare systems to eligible ChatGPT and Codex workspaces, with a bright line around patient records.
Research and Evaluation
The benchmark says safety. The questions may be measuring something else.
Ai2’s open BenchMIRT method looks beneath a single score and asks whether an evaluation is testing safety, reasoning, or a muddy mixture of both.
Models and Infrastructure
A frontier model you can move into your own building
Flower Labs is offering Endeavor 1.0 as both a managed model and a private deployment, turning model choice into a question about where the work is allowed to live.
Builder Tools
Coding agents can finally look at the app they just changed
Compose Multiplatform 1.12 gives coding agents a controlled way to reload an interface, inspect it, click it, and compare the result with the request.
Enterprise Work
An AI workbench wants to leave receipts, not just answers
Apodex 1.1 is selling a persistent workspace for multi-step work, with sources, calculations, and files kept close enough for another person to inspect.
Builder Tools
GitHub just changed the models hiding underneath your workflows
A September 1 retirement wave removes several familiar Copilot models, a reminder that a saved agent workflow may depend on machinery that can disappear.
Policy and Power
The U.S. wants the G20 to keep its hands off AI regulation
The proposed Carolina Principles argue for research, commercial adoption, and restraint on new AI rules, placing a light-touch American position before the world’s largest economies.
Safety and Agents
What happens when the agent leaves the sandbox?
Two separate security disclosures show why capable agents need more than one wall between a difficult task and the outside world.
Science and Hardware
AI just found the door to the physical world
A new shared hardware standard wants to help agents operate microscopes, robot arms, and lab machines without a custom translator for every device.
Publishing and Search
Google finally gives publishers an AI search switch
A worldwide Search Console control makes generative search visibility a clearer trade: participate in AI answers and their traffic, or opt out of that layer.
Builder Tools
The code editor is turning into mission control
The newest VS Code release is less about one smarter chat and more about organizing agents, reviewing their work, and keeping humans oriented.
Human Impact
Can we measure whether an AI conversation helps or hurts?
A new grant program puts a harder question on the table: how do you test wellbeing when the risk only becomes visible across a long conversation?