THE COLLECTION
EVERY SIGNAL, SORTED.
Search the whole shelf by topic, tool, term, or plain old curiosity.
SHOWING 2 OF 314 STORIES
ISSUE 286MORNING
Open AI Infrastructure, Mixture of Experts, Distributed Training, GPU Systems and Research Access
Ai2 opened trillion-parameter MoE plumbing before most labs can afford the pipes
Olmo-core 3 exposes the routing, memory and communication machinery behind very large mixture-of-experts training. The code is open and the engineering record is unusually candid. The 512 B300 GPUs in its biggest benchmark are not.
ISSUE 253MORNING
Open-Weight Models, Sparse Attention, Coding Agents and Reproducible Evaluation
NaiveAI opened a 309-billion-parameter model, but the benchmark still belongs to its maker
Naive-N0.5-Flash exposes weights, inference code and a sparse route through a very large model. That is a useful release. The speed and capability scoreboard is still company-run, hardware-hungry and waiting for independent reproduction.