<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Analysis on AI Coding Blog</title><link>https://mpklu.github.io/categories/analysis/</link><description>Recent content in Analysis on AI Coding Blog</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Sun, 07 Jun 2026 22:28:00 -0400</lastBuildDate><atom:link href="https://mpklu.github.io/categories/analysis/index.xml" rel="self" type="application/rss+xml"/><item><title>The Bugs Were Hiding Behind the Right Answers</title><link>https://mpklu.github.io/posts/bugs-behind-the-right-answers/</link><pubDate>Sun, 07 Jun 2026 22:28:00 -0400</pubDate><guid>https://mpklu.github.io/posts/bugs-behind-the-right-answers/</guid><description>&lt;p&gt;&lt;em&gt;Two colleagues who owned a critical subsystem left within a week of each other, and a leftover task in their area — one I barely knew — landed on me. The migration plan for it was freshly AI-generated, and what I didn&amp;rsquo;t fully trust was my own grasp of it. So I asked my AI to quiz me — not to grade me, just to see what I&amp;rsquo;d actually absorbed.&lt;/em&gt;&lt;/p&gt;</description></item><item><title>The Lattice Hypothesis</title><link>https://mpklu.github.io/posts/lattice/</link><pubDate>Sun, 12 Apr 2026 12:00:00 -0400</pubDate><guid>https://mpklu.github.io/posts/lattice/</guid><description>&lt;p&gt;&lt;em&gt;An engineer&amp;rsquo;s conjecture on distributed biological intelligence, dreams, and the uncomfortable trajectory of AI&lt;/em&gt;&lt;/p&gt;
&lt;hr&gt;
&lt;p&gt;In computing, thick clients carry their own processing power, storage, and local decision-making — but they become dramatically more capable when networked.&lt;/p&gt;
&lt;p&gt;A thought lodged in my head recently and wouldn&amp;rsquo;t leave:&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;A human brain looks remarkably like a thick client.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;Andy Clark and David Chalmers argued in their &lt;a href="https://consc.net/papers/extended.html"&gt;&amp;ldquo;Extended Mind&amp;rdquo; thesis&lt;/a&gt; (1998) that cognition doesn&amp;rsquo;t stop at the skull — it extends into notebooks, tools, and cultural artifacts. They didn&amp;rsquo;t use the client-server vocabulary, but the implication is the same: the brain is a node in a larger system.&lt;/p&gt;</description></item><item><title>The Most Dangerous Hallucination Is the One That Sounds Right</title><link>https://mpklu.github.io/posts/hallucination-pitfalls/</link><pubDate>Wed, 08 Apr 2026 12:00:00 -0400</pubDate><guid>https://mpklu.github.io/posts/hallucination-pitfalls/</guid><description>&lt;p&gt;&lt;em&gt;Your AI coding assistant quotes a specific line from a specific file. The file is real. The section name is close. The quote sounds exactly right. But the line doesn&amp;rsquo;t exist — and that&amp;rsquo;s precisely what makes it dangerous.&lt;/em&gt;&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="a-quote-that-never-was"&gt;A Quote That Never Was&lt;/h2&gt;
&lt;p&gt;I&amp;rsquo;ve been working with Claude Code on a large healthcare integration project — syncing an on-premise practice management system (I&amp;rsquo;ll call it &amp;ldquo;SimplePractice&amp;rdquo;) with a cloud FHIR platform. Over several weeks, we&amp;rsquo;ve built push/pull sync, environment tooling, documentation, and dozens of shell scripts. It&amp;rsquo;s been genuinely productive.&lt;/p&gt;</description></item><item><title>Gemma 4 Structured-Task Performance: Field Report from a Local-First App</title><link>https://mpklu.github.io/posts/gemma-4-benchmark/</link><pubDate>Mon, 06 Apr 2026 12:00:00 -0400</pubDate><guid>https://mpklu.github.io/posts/gemma-4-benchmark/</guid><description>&lt;h1 id="gemma-4-structured-task-performance-field-report-from-a-local-first-app"&gt;Gemma 4 Structured-Task Performance: Field Report from a Local-First App&lt;/h1&gt;
&lt;p&gt;&lt;em&gt;Benchmark data and prompt-format findings from deploying Gemma 4 E4B in a real application. Intended for LLM teams (Gemma, Ollama) and developers building structured-output pipelines on local models.&lt;/em&gt;&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="context"&gt;Context&lt;/h2&gt;
&lt;p&gt;We build &lt;a href="https://github.com/mpklu/gary"&gt;Gary&lt;/a&gt;, a privacy-first personal assistant CLI for macOS. It runs entirely locally — an encrypted database, a daemon process, and a local LLM via ollama. The LLM handles three structured tasks:&lt;/p&gt;</description></item><item><title>The Vibe Coding Trap</title><link>https://mpklu.github.io/posts/vibe-coding-trap/</link><pubDate>Mon, 23 Mar 2026 00:30:00 -0400</pubDate><guid>https://mpklu.github.io/posts/vibe-coding-trap/</guid><description>&lt;p&gt;&lt;em&gt;Let&amp;rsquo;s be real. When the first AI coding agents dropped, we all nodded solemnly and said, &amp;ldquo;Of course, a human will always review every single change. Safety first.&amp;rdquo;&lt;/em&gt;&lt;/p&gt;
&lt;p&gt;&lt;em&gt;We lied to ourselves. Or, more accurately, we succumbed to the seductive illusion of frictionless productivity—a &lt;a href="https://mpklu.github.io/posts/10x-illusion/"&gt;10x illusion&lt;/a&gt; where we feel like we&amp;rsquo;re coding faster, but we&amp;rsquo;re actually just accumulating debt we can&amp;rsquo;t afford to pay.&lt;/em&gt;&lt;/p&gt;
&lt;p&gt;&lt;em&gt;The recent &lt;a href="https://stackoverflow.blog/2026/03/19/ai-is-becoming-a-second-brain-at-the-expense-of-your-first-one/"&gt;Stack Overflow post on AI as a second brain&lt;/a&gt; identifies the core issue: we are offloading our judgment. This isn&amp;rsquo;t a future sci-fi risk; &lt;a href="https://stackoverflow.blog/2026/03/19/ai-is-becoming-a-second-brain-at-the-expense-of-your-first-one/"&gt;cognitive offloading&lt;/a&gt; is happening now, and it&amp;rsquo;s reshaping both our codebases and our minds.&lt;/em&gt;&lt;/p&gt;</description></item><item><title>10x Illusion</title><link>https://mpklu.github.io/posts/10x-illusion/</link><pubDate>Fri, 20 Mar 2026 18:43:42 -0400</pubDate><guid>https://mpklu.github.io/posts/10x-illusion/</guid><description>&lt;h1 id="the-10x-illusion-if-ai-codes-10x-faster-how-much-faster-do-projects-actually-ship"&gt;The 10x Illusion: If AI Codes 10x Faster, How Much Faster Do Projects Actually Ship?&lt;/h1&gt;
&lt;p&gt;&lt;em&gt;AI coding tools are getting shockingly good. So it&amp;rsquo;s natural to ask: if the coding part gets 10x faster, shouldn&amp;rsquo;t the whole project get 10x faster too?&lt;/em&gt;&lt;/p&gt;
&lt;p&gt;&lt;em&gt;The answer is surprisingly counterintuitive — and backed by a growing body of data.&lt;/em&gt;&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id="the-speed-is-real-the-extrapolation-is-not"&gt;The Speed Is Real. The Extrapolation Is Not&lt;/h2&gt;
&lt;p&gt;AI coding tools deliver genuine speed on implementation tasks. GitHub Copilot studies show developers completing isolated coding tasks &lt;strong&gt;55% faster&lt;/strong&gt;. AI agents can generate entire modules in minutes. The speed is not the illusion.&lt;/p&gt;</description></item></channel></rss>