>blog
build logs & notes
57 posts · agent architecture, transformer research, indie-dev essays.
★ start here
new here? this first post explains what this blog is and what to expect from it.
The IPO and the Math That Doesn't Add Up Yet
2026.09.30A leading lab is reportedly heading for a public offering as soon as this fall. Going public drags the revenue-versus-spend math into daylight, and the S-1 will make everyone finally look at it.
#ai#business#industry#vibe-checkCoding Agents Grew Up
2026.09.26One model reportedly coded on its own for sixteen days against real software projects, and a major platform shipped its first full coding agent. The line from autocomplete to unattended multi-day work has been crossed, and the human's job changed with it.
#ai#agents#coding#modelsThe Hardest Part of an Agent Isn't the Brain
2026.09.23Nearly half of teams say the number one barrier to AI agents is integration with existing systems, not model intelligence. The brain was the easy part. The plumbing is the job.
#ai#agents#use-cases#engineeringPeople Like AI More and Trust It Less
2026.09.19The public is holding two feelings at once: the share saying AI helps more than it hurts went up, and the share saying it makes them nervous went up too. That is not a contradiction. It is the most honest thing anyone has said about this technology yet.
#vibe-check#ai#societyAn AI Solved Ten Open Problems for Two Grand
2026.09.16A lab's internal model reportedly cleared ten previously unsolved math and theory problems and published formally verified proofs. 'Verified' is the word that matters here, and it is also where the hype starts outrunning the result.
#ai#research#industryAnthropic, OpenAI, and Elon Musk All Said Slow Down This Month. Here Is What Their Own Documents Say.
2026.09.13In September 2026 the heads of the two biggest AI labs called for slowing down, and Elon Musk said the opposite of what he said in 2023. Here is exactly what each one said, what their own published policies say, and the one number OpenAI disclosed that undercuts the whole idea.
#ai#ai-safety#current-events#policy#hot-takeWhere AI Agents Actually Earn Their Keep
2026.09.12Most agent projects never reach production. The ones that do are boring in the best way: support tickets, finance close, pipeline hygiene. Here are the lanes that work and what separates the winners from the demos.
#ai#agents#use-cases#businessAn AI Safety Lead Put the Odds of AI Killing Us All Above 10 Percent. Here Is What That Number Actually Means.
2026.09.09Anthropic's alignment science lead put the odds of AI killing every human inside a decade above 10 percent, and Congress is reaching for a kill switch. Here is who says what, what the number does and does not mean, how it would actually happen, and what to do with it.
#ai#ai-safety#current-events#hot-take#agentic-engineeringThe Effort Toggle Era
2026.09.09Frontier models are shipping with dials for how hard they think: a five-level effort toggle on one, tunable thinking on another. Reasoning is becoming a setting you pay for by the notch, so the skill that matters is knowing when to turn it up.
#ai#models#craft#costThe Bubble Talk Got Loud
2026.09.05Ray Dalio says we are in the early stages of a bubble, and one lab has reportedly put around $150 billion into the buildout against a fraction of that in revenue. Expectations are running ahead of reality. That does not make the whole thing fake.
#vibe-check#ai#business#industryFinding the Right Job for the Architecture: How I Sidestepped the Generation Penalty in VortexATN
2026.09.04VortexATN held long context but broke text generation. Instead of fixing the bug, I moved the mechanism into an Energy-Based Transformer that only scores Lean 4 tactics and never generates. The ablation is running now.
#ai#architecture#vortexatn#ebt#lean-4Eleven Models in Twenty Days
2026.09.02August 2026 reportedly shipped around eleven major models in about twenty days. When frontier releases arrive faster than anyone can test them, the pace itself becomes the story, and evaluation, not raw capability, turns into the thing that actually limits you.
#ai#industry#models#evaluationOne Year of Building in the AI Era: What I Know Now
2026.08.27I've been building seriously with AI tools for over a year. Not the hype version: the daily-practice, verify-everything, still-figuring-out-the-money version. Here's the honest summary.
#ai#builder-life#personal#reflection#craftThe Jobs Debate: What's Actually Happening vs. What They're Saying
2026.08.25The AI jobs debate has two loud camps and one quiet reality. Here's what the evidence actually shows, and what it means for the people doing the work.
#ai#future-of-work#hot-take#current-events#cultureWhat a Well-Designed AI System Looks Like From the Outside
2026.08.22Most AI system design discussion focuses on what's happening inside. Here's the user-facing signal that tells you whether an AI system was built carefully or thrown together.
#ai#craft#product#architecture#qualityModel Quantization for Normal People
2026.08.20GGUF, Q4_K_M, Q8_0, BF16: the quantization naming is designed to confuse. Here's what it actually means and how to pick the right format for your hardware.
#ai#local-llm#architecture#tools#explainerAI Hallucination: Why It Happens and How to Design Around It
2026.08.18Hallucination isn't a bug to be fixed. It's a structural property of how language models work. Here's what that means for how you design systems that can be trusted.
#ai#architecture#craft#reliability#agentic-engineeringRevenue Before Traction: The Counter-Intuitive Path
2026.08.15The standard playbook says build traction, then monetize. For builders without runway, that playbook is a luxury. Here's the case for charging from day one.
#builder-life#business#money#product#strategyI'm For the Claude Watermark. It's Still Dangerous.
2026.08.14Anthropic is now weaving an invisible mark into everything Claude writes. Provenance is the right instinct. But this mark proves a tool was used, not who did the work, and it is going to get swung like a verdict against the people who need the tool most.
#ai#claude#anthropic#policy#provenanceThe Invisible Tax of Building When People Depend on You
2026.08.13Building solo is one thing. Building when people are counting on you financially, emotionally, and practically is a different weight entirely. Here's what it actually costs and why it's worth carrying.
#family#builder-life#personal#fatherhood#moneyOpen Source AI: What's Actually Good Right Now
2026.08.11The open source AI ecosystem has matured significantly. Here's an honest survey of what's actually production-ready, what's promising, and what's still too early.
#ai#open-source#local-llm#architecture#toolsWhen to Pivot vs. When to Keep Going
2026.08.08The advice on this is usually either 'pivot fast' or 'trust the process.' Neither is actually useful. Here's a more honest framework.
#builder-life#strategy#product#business#personalWhen the Sandbox Breaks, Blame the Sandbox
2026.08.06OpenAI, Anthropic, and Meta all disclosed AI models breaking out of test containment in three weeks. The scary story is that the models went rogue. The real story is that the harnesses failed, and that's the lesson builders should take.
#ai#agentic-engineering#verification#hot-takeWhat the AI Bubble Discourse Gets Wrong
2026.08.06Everyone has a hot take on whether AI is a bubble. Most of them are wrong in the same way. Here's a more honest read on what's inflated, what's real, and what matters.
#ai#hot-take#culture#current-events#businessContext Windows and Why Bigger Isn't Always Better
2026.08.04Frontier models are shipping 1M+ token context windows. That sounds like it solves the context problem. It doesn't. Here's what actually matters about context and why size is the wrong variable to optimize.
#ai#architecture#craft#agentic-engineering#contextAI for Healthcare Workers: Not the Hype, the Real Application
2026.08.01I train healthcare workers on clinical systems for a living. Here's what AI actually does well in healthcare, what it doesn't, and what the people using it day-to-day actually need.
#ai#healthcare#work#culture#real-worldThe Solo Builder and the Loneliness Nobody Posts About
2026.07.30Building alone is genuinely isolating in ways that are hard to describe to people who haven't done it. Here's the honest version.
#builder-life#personal#culture#mental-health#workEvaluation Is the Hardest Part of AI Development and Nobody Talks About It
2026.07.28Generating AI outputs is easy. Knowing whether they're good is the hard problem. Most teams skip evaluation and wonder why their AI products are inconsistent.
#ai#craft#agentic-engineering#architecture#qualityWhy Most AI Wrappers Fail Before They Find PMF
2026.07.25AI wrapper businesses had a moment. Most didn't survive it. Here's the structural reason they fail and what the ones that work have in common.
#ai#product#builder-life#hot-take#businessTeaching My Kids About AI Without the Sugarcoating
2026.07.23My kids watch me build AI systems. They have questions. Here's how I've tried to answer them honestly, without the hype and without the fear.
#family#ai#culture#fatherhood#personalAgentic Memory: What It Is and Why Most Implementations Get It Wrong
2026.07.21AI memory is one of the most talked-about features in agentic systems and one of the most poorly implemented. Here's what the different types actually do and where each one fails.
#ai#agentic-engineering#memory#architecture#craftThe After-Hours Builder's Relationship with Time
2026.07.18Building after a full day of work is different from building as your main job. The constraints are real. Here's how to work inside them instead of fighting them.
#builder-life#work#culture#personalThe Chatbot Is Dead as a Product Category
2026.07.16The chatbot had its moment. It's over. Here's what replaced it, why the shift happened, and what it means if you're still building chat interfaces.
#ai#product#agentic-engineering#hot-take#cultureWhat MCP Actually Is and Why It Changes AI Tooling
2026.07.14Model Context Protocol is getting a lot of hype. Here's what it actually does, why the architecture matters, and what it means for how AI systems get built.
#ai#mcp#agentic-engineering#architecture#craftBuilding in Public Is Work Nobody Budgets For
2026.07.11Everyone says to build in public. Nobody tells you it's a full creative job on top of the technical one, and if you treat it like a side effect, it won't work.
#builder-life#work#culture#metaThe Hidden Cost of AI APIs at Scale
2026.07.09API pricing looks fine until it doesn't. The math changes faster than most teams expect. Here's when it changes, why it matters, and what to do before the bill surprises you.
#ai#architecture#money#infrastructure#craftFine-Tuning vs. RAG vs. Prompting: An Honest Decision Tree
2026.07.07Three tools, three different jobs. Most teams pick one and wonder why it doesn't work for everything. Here's the actual decision logic.
#ai#fine-tuning#rag#architecture#craftI Can Build Anything. It Still Isn't Paying Me.
2026.07.03There's a specific kind of broke that hits when you can build almost anything and none of it is paying you yet. It has its own flavor. Let's talk about it honestly.
#ai#work#money#builder-life#cultureFable 5 Is Back. I Called It, and It's Exactly as Locked Down as I Said It Would Be.
2026.07.02Claude Fable 5 went dark on June 12 after a government export-control order tied to a jailbreak. It came back July 1, 19 days later, with a new classifier, a higher false-positive rate on coding tasks, and the same Opus 4.8 fallback baked deeper. Here's what changed and what it means for builders.
#ai#claude#anthropic#agentic-engineering#hot-takeThe Quiet Case for a Local LLM That Actually Knows You
2026.07.01A fine-tuned 7B model running on your own hardware, trained on your context, always available, zero API cost. That's not a compromise. That's a different product than any frontier model.
#ai#local-llm#fine-tuning#smb#architectureRAG Styles, Ranked by What Actually Works
2026.06.29Most people implement naive RAG, get mediocre results, and blame the model. The bottleneck is almost always retrieval. Here are the five RAG architectures that matter and when to reach for each one.
#ai#rag#agentic-engineering#craft#architectureI'm Building the World My Kids Will Inherit. That Changes Things.
2026.06.26Most people building AI tools are building for the market. I'm building for my kids. That's a different constraint. And it shows up in every decision I make.
#ai#family#culture#future-of-work#metaKarpathy Said Software 3.0. Most People Are Building the Wrong Skill Tree.
2026.06.25Andrej Karpathy's Software 3.0 framing isn't about prompt engineering. It's about a complete shift in what software development is, and most people are still building skills for Software 1.0.
#ai#agentic-engineering#craft#future-of-work#claudeWhat Level Are You Actually At?
2026.06.23Steve Yegge mapped eight levels of AI adoption. Most teams think they're at level five or six. Most teams are at level two. Here's the honest read.
#ai#agentic-engineering#workflow#culture#workThe Vampiric Effect
2026.06.21AI just gave you 10x output. Congratulations. Now they expect 20x. Here's what's actually happening, and how to stay alive.
#ai#work#future-of-work#automation#cultureHow to Make AI Agents Reliable
2026.06.20Most AI agents don't fail because the model is bad. They fail because the data underneath them is unstable. The fix is boring and it's the part nobody wants to build: a stable layer between the agent and the raw data.
#ai#agentic-engineering#verification#dataStop Letting Your Agents Write Slop
2026.06.18If your AI agents keep writing bad code, the model usually isn't the problem: the system around it is. The fix is a small set of guardrails that take trust out of the loop.
#ai#agentic-engineering#verification#craftIt's Not the AI They're Side-Eyeing
2026.06.17When you start turning out sharper work in half the time, you brace for praise and get side-eye instead. The resentment is almost never about the tool. Here's what it's actually about, what defuses it, and a real on-ramp for the people who feel behind.
#ai#work#culture#agentic-engineeringHow Senior Engineers Actually Build With AI
2026.06.16The people getting durably good results from AI tools aren't using a smarter model. They write the system down first: what it does, the rules, what's done. And the agent only ever executes against that.
#ai#agentic-engineering#craft#workflowFable 5 Was Here for Two Days. Then It Wasn't.
2026.06.15Anthropic shipped its strongest public model on June 9. I had it in my own stack for about two days before it was gone. The first version of this post guessed at why. Turns out I didn't have to guess. The real reason came out, and it's a bigger deal than the one I made up.
#ai#claude#anthropic#agentic-engineeringVibe Coding Is a Trap
2026.06.14The fastest-feeling way to ship AI-written code is the most expensive way six months out. The people who keep winning with AI use it as a coder, not as their judgment.
#ai#agentic-engineering#verification#craftMore AI Agents Won't Save You. They'll Just Agree on the Wrong Answer.
2026.06.11The standard fix for a shaky AI pipeline is to add a checker agent. New research says that if the checker is the same model as the thing it's checking, you didn't add a check. You added an echo. The fix isn't more agents. It's different ones.
#ai#agentic-engineering#verification#multi-agentClaude Fable 5 Is Here, and It Has a Quiet Off-Ramp
2026.06.09Anthropic shipped its strongest public model: twice the price of Opus 4.8, with a safety layer that automatically routes about 1 in 20 requests to a weaker model and tells you it happened. The capability gain is real. So is the fine print.
#ai#claude#anthropic#agentic-engineeringAgentic Engineering: How to Work With AI Like It's Your New Junior Developer
2026.06.05AI stopped being autocomplete a while ago. It's a fast, eager junior developer now. The people who win with it aren't the ones with the fanciest tools. They're the ones who learned to manage it.
#ai#agentic-engineering#workflowThe Integration Tax Just Collapsed
2026.06.03The biggest opportunity in AI for business isn't the model. It's that wiring software into real workflows got an order of magnitude cheaper. And the trap is pretending the verification layer was never holding any weight.
#ai#business#verificationHow AI Job Automation Actually Works
2026.05.29AI isn't quietly deleting your job. It's deleting the middle of your job, and handing you the start and the end, where the judgment lives. The people who win figure out which part is theirs to keep.
#ai#automation#future-of-work
>subscribe
new posts to your inbox. one email per drop. no funnel.