Newsletter
Thoughts on AI, development, and technology. Published on Substack.
The Bible Didn’t Make Me a Christian. It Did Make Me Think.
Note: Trying a new format with the sort of sparsely-lined paragraphs - I think it may read a little better here on substack than my normal ranting long form drivel....
Qwen3.6-35B-A3B-MTP at 70 tok/s on an RTX 3070
I’ve been running Qwen3.6-35B-A3B-MTP on my RTX 3070 (8GB VRAM) for a while now and previously documented getting ~55 tok/s....
Multi-Token Prediction MTP in llama.cpp How It Works and How to Use It
PR #22673 just landed in upstream llama.cpp, and it is a big deal....
The Architecture Breakthrough Nobody in Local LLM is Talking About
If you’ve been following the local LLM space this year, you’ve probably been tracking parameter counts, quantization methods, and hardware requirements like a stock trader watching candlesticks....
A 760M-Parameter Model That Beats GPT-5-High on Math
Let’s talk about a paper that quietly dropped on arXiv this week and absolutely deserves more attention than it’s getting....
Living with ADHD as an adult in 2026
I hate the fact that I can sit down and wrestle with complex problems for hours....
Google Silently Deploys a 4GB Local LLM to Every Chrome User
Google has quietly installed a 4GB local LLM into Chrome, and billions of users have no idea it is there....
Running Qwen3.6-35B-A3B-MTP at 55 tok/s on an 8GB GPU
A Practical Guide...
This Idea Might Change How AI Actually Thinks
and it quietly reveals something deeper about knowledge, logic, and how natural thinking really works...
Why are we looking for a 'flat' signal in song production?
A small piece on why we use monitors, EQ's and other things in order to get a flat frequency response....
How to tune your subwoofer and monitors correctly
The correct way to set the cutoffs on your studio gear....