THE COLLECTION

EVERY SIGNAL, SORTED.

Search the whole shelf by topic, tool, term, or plain old curiosity.

SHOWING 2 OF 314 STORIES
ISSUE 286MORNING

Open AI Infrastructure, Mixture of Experts, Distributed Training, GPU Systems and Research Access

Ai2 opened trillion-parameter MoE plumbing before most labs can afford the pipes

Olmo-core 3 exposes the routing, memory and communication machinery behind very large mixture-of-experts training. The code is open and the engineering record is unusually candid. The 512 B300 GPUs in its biggest benchmark are not.

#open-source-ai#mixture-of-experts#ai-infrastructure
ISSUE 253MORNING

Open-Weight Models, Sparse Attention, Coding Agents and Reproducible Evaluation

NaiveAI opened a 309-billion-parameter model, but the benchmark still belongs to its maker

Naive-N0.5-Flash exposes weights, inference code and a sparse route through a very large model. That is a useful release. The speed and capability scoreboard is still company-run, hardware-hungry and waiting for independent reproduction.

#open-weights#mixture-of-experts#model-evaluation