Welcome to the AGI era of AI governance
It's a one-way door and we weren't ready for it.
Lo que el mundo de la IA está comentando ahora mismo, sacado directo de las fuentes — lo más nuevo primero.
It's a one-way door and we weren't ready for it.
One step further into the power politics of frontier AI systems.
A curated roundup of notable LLM research papers that came out this year
This was my last week at the Allen Institute for AI (Ai2), where I got the great privilege to work on the Olmo models, to grow, to learn, and to have broad lasting impacts.
Where marginally higher intelligence drives value, and where it doesn't.
Gemini Flash 3.5, Mythos, open-closed balance, America's open-source surge, emerging power struggles and more.
From Gemma 4 to DeepSeek V4, How New Open-Weight LLMs Are Reducing Long-Context Costs
A learning-oriented workflow for understanding new open-weight model releases
How coding agents use tools, memory, and repo context to make LLMs work better in practice
From MHA and GQA to MLA, sparse attention, and hybrid architectures
A Round Up And Comparison of 10 Open-Weight LLM Releases in Spring 2026
And an Overview of Recent Inference-Scaling Papers
A 2025 review of large language models, from DeepSeek R1 and RLVR to inference-time scaling, benchmarks, architectures, and predictions for 2026.
In June, I shared a bonus article with my curated and bookmarked research paper lists to the paid subscribers who make this Substack possible.
Understanding How DeepSeek's Flagship Open-Weight Models Evolved
Linear Attention Hybrids, Text Diffusion, Code World Models, and Small Recursive Transformers
Multiple-Choice Benchmarks, Verifiers, Leaderboards, and LLM Judges with Code Examples
A Detailed Look at One of the Leading Open-Source LLMs
And How They Stack Up Against Qwen3
From DeepSeek-V3 to Kimi K2: A Look At Modern LLM Architecture Design
A topic-organized collection of 200+ LLM research papers from 2025