Google’s Gemini lineup has a Pro-sized hole
PLUS: Sell a high-value AI workflow audit as a consultant
AI 开发圈现在正在聊什么,直接从来源抓取——最新的在最前面。
PLUS: Sell a high-value AI workflow audit as a consultant
PLUS: Anduril and Archer bring the ‘Thunder’
unsolved maths problems vs ai models, who'll win?
PLUS: Make AI note-taking easier with a “Captain’s Log” setup
The global implications on the AI ecosystem.
How LLMs Learn Low-, Medium-, and High-Effort Reasoning Modes
links to read over the weekend
new desktop app and hosted sites in ChatGPT
The most serious test to date of open source AI’s viability is happening right now.
More Fable, GPT-5.6 and new ChatGPT voice
what I want from a harness
An assessment of the open ecosystem and the motivations behind releasing models
Using Open-Weight Models in Local Coding Harnesses as an Alternative to Claude Code and Codex Subscriptions
A capability threshold I've been carefully monitoring.
This post was originally an op-ed co-authored with Kevin Xu of Interconnected for a general, non-technical audience.
About 3 years since I started writing weekly.
"Interview" #18
It's a one-way door and we weren't ready for it.
One step further into the power politics of frontier AI systems.
A curated roundup of notable LLM research papers that came out this year
This was my last week at the Allen Institute for AI (Ai2), where I got the great privilege to work on the Olmo models, to grow, to learn, and to have broad lasting impacts.
Where marginally higher intelligence drives value, and where it doesn't.
Gemini Flash 3.5, Mythos, open-closed balance, America's open-source surge, emerging power struggles and more.
From Gemma 4 to DeepSeek V4, How New Open-Weight LLMs Are Reducing Long-Context Costs
A learning-oriented workflow for understanding new open-weight model releases
How coding agents use tools, memory, and repo context to make LLMs work better in practice
From MHA and GQA to MLA, sparse attention, and hybrid architectures
A Round Up And Comparison of 10 Open-Weight LLM Releases in Spring 2026
And an Overview of Recent Inference-Scaling Papers
A 2025 review of large language models, from DeepSeek R1 and RLVR to inference-time scaling, benchmarks, architectures, and predictions for 2026.
In June, I shared a bonus article with my curated and bookmarked research paper lists to the paid subscribers who make this Substack possible.
Understanding How DeepSeek's Flagship Open-Weight Models Evolved
Linear Attention Hybrids, Text Diffusion, Code World Models, and Small Recursive Transformers