Feedforward
Blogs from the AI community
- Simon Willison's Weblogsimonwillison.net
There are no lossless transformations of natural-language text
- Rohan's Bytesrohan-paul.com
🗞️ Meta released Muse Glimmer, a 30B-parameter model built for always-on local agents on one consumer GPU
- Sebastian Raschka, PhDsebastianraschka.com
Muse Glimmer 30B Architecture Notes
- Interconnectsinterconnects.ai
5 useful things you'll learn in my new post-training textbook (shipping now!)
- Import AI jack-clark.net
Import AI 468: 23 RSI ideas; PostTrainBench+; and how trust and transparency interplay with AI racing
- Deep (Learning) Focuscameronrwolfe.substack.com
Notes on Midtraining
- Adjacent Possibleadjacentpossible.substack.com
The Uses Of Terror
- Understanding AIunderstandingai.org
Mathematicians are grappling with the possibility that AI might eclipse them
- Lossfunk Lettersletters.lossfunk.com
Diffusion models can see an illusion. Then they forget it before the picture comes out.
- Victoria Krakovnavkrakovna.wordpress.com
Using AI to analyze life patterns
- The Berkeley Artificial Intelligence Research Blogbair.berkeley.edu
From CUDA to MLX: How K-Search Brings Decades of Kernel Expertise to Apple Silicon
- Bounded Regretbounded-regret.ghost.io
Foundation Models for Oversight
- One Useful Thingoneusefulthing.org
An opinionated guide to which AI to use to do stuff
- Posts on Max Woolf's Blogminimaxir.com
LLMs break down in funny ways when told the Jacobian Conjecture counterargument
- ML Safety Newsletternewsletter.mlsafety.org
MLSN #22: Turning Cyber Vulnerabilities Into Exploits
- Maggie Appletonmaggieappleton.com
In Memoriam, Encountering World War I, and Thinkable Horrors
- Sorta Insightfulalexirpan.com
Which Tech CEOs Are Gamers?
- Hamel's Blog - Hamel Husainhamel.dev
Do Automated Evals Work?
- swyx's site RSS Feedswyx.io
What America has meant to me
- Lil'Loglilianweng.github.io
Harness Engineering for Self-Improvement
- Hamel’s Substackhamelhusain.substack.com
"It's Hard to Eval" Is a Product Smell
- Ponder on Franz Louis Cesistaleloykun.github.io
LoRA-Muon-OGD: Spectral Orthogonal Gradient Projection on the Low-Rank Manifold for LLM Continual Learning
- Eugene Yaneugeneyan.com
Patterns for Building Cybersecurity Evals
- Jeremy Jordanjeremyjordan.me
Exploring the age of continuous work.
- Token for Tokenblog.jxmo.io
Zen and the Art of AI Research
- âś°Vicki Boykisâś°vickiboykis.com
Running local models is good now
- Exploring Language Modelsnewsletter.maartengrootendorst.com
A Visual Guide to DiffusionGemma
- AI: A Guide for Thinking Humansaiguide.substack.com
On AI and "Jagged Intelligence"
- Shreya Shankarsh-reya.com
Exploring Agent-Assisted Qualitative Analysis
- Sander Dielemansander.ai
Learning the integral of a diffusion model
- karpathykarpathy.bearblog.dev
Sequoia Ascent 2026 summary
- Maharshi's blogmaharshi.bearblog.dev
Identity and ArithTuple Tensors in CuTeDSL
- Sam Altmanblog.samaltman.com
-
- Chris McCormickmccormickml.com
Optimizing Training with FlashAttention varlen
- flurries of latent creativityblog.singleton.io
Cheap turpentine
- zhengdongwang.comzhengdongwang.com
The means of some change
- Nicholas Carlininicholas.carlini.com
How to win a best paper award
- inFERENCeinference.vc
The Future of Software
- The Gradientthegradient.pub
After Orthogonality: Virtue-Ethical Agency and AI Alignment
- Andrej Karpathy blogkarpathy.github.io
microgpt
- antifragile systemsyongzx.substack.com
CoT monitorability: why g-means and not F1?
- Tim Dettmerstimdettmers.com
My Journey Towards Coding Agents: Building SERA
- Yi Tayyitay.net
2025 introspections: my year back at Google
- Nick’s Substacknickjiang.substack.com
Teaching Algorithms in Ethiopia
- Neel Nandaneelnanda.io
MATS Applications Open (Due Aug 29)
- Blog - Jason Weijasonwei.net
Life lessons from reinforcement learning
- atharva's blogksagar.bearblog.dev
how we accidentally solved robotics by watching 1 million hours of YouTube
- jeremybernste.injeremybernste.in
Deriving Muon
- Thonk From First Principlesthonking.ai
Why PyTorch is an amazing place to work... and Why I'm Joining Thinking Machines
- Chip Huyenhuyenchip.com
Common pitfalls when building generative AI applications
- FOR OUR POSTERITYforourposterity.com
SITUATIONAL AWARENESS: The Decade Ahead
- ruder.ioruder.io
The Evolving Landscape of LLM Evaluation
- Jascha’s blogsohl-dickstein.github.io
Neural network training makes beautiful fractals
- Kemal Erdem Blog RSS Feederdem.pl
Step by Step visual introduction to Diffusion Models.
- Brian Kitanoblog.briankitano.com
Llama from scratch (or how to implement a paper without crying)
- siboehmsiboehm.com
Can Function Inlining Affect Floating Point Outputs? Exploring FMA and Other Consistency Issues
- peterbloem.nlpeterbloem.nl
What design can teach us
- Greg Brockmanblog.gregbrockman.com
It's time to become an ML engineer
- Awni Hannunawni.github.io
An Introduction to Fisher Information
- Eric Jangblog.evjang.com
Robots Must Be Ephemeralized
- colah's blogcolah.github.io
Collaboration and Credit Principles