<?xml version="1.0" encoding="UTF-8" ?>
<rss xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" version="2.0">
  <channel>
    <title><![CDATA[WUU73 AI Guides]]></title>
    <description><![CDATA[Practical AI coding guides, field notes, and experiments.]]></description>
    <link>https://wuu73.org/aiguide</link>
    <atom:link href="https://wuu73.org/aiguide/aiguide-rss.xml" rel="self" type="application/rss+xml" />
    <lastBuildDate>Sun, 13 Sep 2026 04:50:55 +0000</lastBuildDate>
    <ttl>60</ttl>

        <item>
          <title><![CDATA[Buttons Cli - The terminal app with a full GUI, AI Help, Agents can control it (so you can see them work) NOT minimal ADHD terminal app - Free AI usage included.]]></title>
          <description><![CDATA[buttonscli.com]]></description>
          <link>https://wuu73.org/aiguide/buttons-cli-my-adhd-terminal-app/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/posts/buttons-cli-my-adhd-terminal-app/</guid>
          <category><![CDATA[App Related]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Fri, 31 Jul 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<figure class="kg-card kg-image-card"><img src="https://wuu73.org/aiguide/content/images/imported/cms-buttons-cli-my-adhd-terminal-app/buttonscli_P6SpwFxDeH.png" class="kg-image" alt="buttons-cli" loading="lazy" width="1790" height="1195" srcset="https://wuu73.org/aiguide/content/images/size/w600/imported/cms-buttons-cli-my-adhd-terminal-app/buttonscli_P6SpwFxDeH.png 600w, https://wuu73.org/aiguide/content/images/size/w1000/imported/cms-buttons-cli-my-adhd-terminal-app/buttonscli_P6SpwFxDeH.png 1000w, https://wuu73.org/aiguide/content/images/size/w1600/imported/cms-buttons-cli-my-adhd-terminal-app/buttonscli_P6SpwFxDeH.png 1600w, https://wuu73.org/aiguide/content/images/imported/cms-buttons-cli-my-adhd-terminal-app/buttonscli_P6SpwFxDeH.png 1790w" sizes="(min-width: 720px) 720px"></figure><p><a href="https://buttonscli.com/?ref=wuu73.org" rel="noreferrer">buttonscli.com</a> - buttons, AI help and automation, AI control / AI can create new tabs, type commands, use multiple at the same time, full user interface, doesn't force you to memorize lots of key combos just to do simple things</p><figure class="kg-card kg-gallery-card kg-width-wide kg-card-hascaption"><div class="kg-gallery-container"><div class="kg-gallery-row"><div class="kg-gallery-image"><img src="https://wuu73.org/aiguide/content/images/2026/09/buttonscli_rlXOucfHbj.png" width="1861" height="1534" loading="lazy" alt="" srcset="https://wuu73.org/aiguide/content/images/size/w600/2026/09/buttonscli_rlXOucfHbj.png 600w, https://wuu73.org/aiguide/content/images/size/w1000/2026/09/buttonscli_rlXOucfHbj.png 1000w, https://wuu73.org/aiguide/content/images/size/w1600/2026/09/buttonscli_rlXOucfHbj.png 1600w, https://wuu73.org/aiguide/content/images/2026/09/buttonscli_rlXOucfHbj.png 1861w" sizes="(min-width: 1200px) 1200px"></div></div></div><figcaption><p dir="ltr"><span style="white-space: pre-wrap;">buttons cli terminal screenshots</span></p></figcaption></figure><p></p><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aiguide/content/images/2026/09/buttonscli_xjo5TtrjuK-1.png" class="kg-image" alt="" loading="lazy" width="1257" height="1117" srcset="https://wuu73.org/aiguide/content/images/size/w600/2026/09/buttonscli_xjo5TtrjuK-1.png 600w, https://wuu73.org/aiguide/content/images/size/w1000/2026/09/buttonscli_xjo5TtrjuK-1.png 1000w, https://wuu73.org/aiguide/content/images/2026/09/buttonscli_xjo5TtrjuK-1.png 1257w" sizes="(min-width: 720px) 720px"></figure><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aiguide/content/images/2026/09/buttonscli_VcQmxkNTOE-1.png" class="kg-image" alt="" loading="lazy" width="1188" height="1351" srcset="https://wuu73.org/aiguide/content/images/size/w600/2026/09/buttonscli_VcQmxkNTOE-1.png 600w, https://wuu73.org/aiguide/content/images/size/w1000/2026/09/buttonscli_VcQmxkNTOE-1.png 1000w, https://wuu73.org/aiguide/content/images/2026/09/buttonscli_VcQmxkNTOE-1.png 1188w" sizes="(min-width: 720px) 720px"></figure><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aiguide/content/images/2026/09/buttonscli_V8SfKUvPSJ-1.png" class="kg-image" alt="" loading="lazy" width="1182" height="798" srcset="https://wuu73.org/aiguide/content/images/size/w600/2026/09/buttonscli_V8SfKUvPSJ-1.png 600w, https://wuu73.org/aiguide/content/images/size/w1000/2026/09/buttonscli_V8SfKUvPSJ-1.png 1000w, https://wuu73.org/aiguide/content/images/2026/09/buttonscli_V8SfKUvPSJ-1.png 1182w" sizes="(min-width: 720px) 720px"></figure><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aiguide/content/images/2026/09/explorer_ZePSufRA4b.png" class="kg-image" alt="" loading="lazy" width="2000" height="1101" srcset="https://wuu73.org/aiguide/content/images/size/w600/2026/09/explorer_ZePSufRA4b.png 600w, https://wuu73.org/aiguide/content/images/size/w1000/2026/09/explorer_ZePSufRA4b.png 1000w, https://wuu73.org/aiguide/content/images/size/w1600/2026/09/explorer_ZePSufRA4b.png 1600w, https://wuu73.org/aiguide/content/images/size/w2400/2026/09/explorer_ZePSufRA4b.png 2400w" sizes="(min-width: 720px) 720px"></figure><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aiguide/content/images/2026/09/explorer_6yhgY4HWdu-1.png" class="kg-image" alt="" loading="lazy" width="2000" height="1068" srcset="https://wuu73.org/aiguide/content/images/size/w600/2026/09/explorer_6yhgY4HWdu-1.png 600w, https://wuu73.org/aiguide/content/images/size/w1000/2026/09/explorer_6yhgY4HWdu-1.png 1000w, https://wuu73.org/aiguide/content/images/size/w1600/2026/09/explorer_6yhgY4HWdu-1.png 1600w, https://wuu73.org/aiguide/content/images/size/w2400/2026/09/explorer_6yhgY4HWdu-1.png 2400w" sizes="(min-width: 720px) 720px"></figure>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[first post test]]></title>
          <description><![CDATA[first, post, test]]></description>
          <link>https://wuu73.org/aiguide/first-post-test/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/posts/first-post-test/</guid>
          <category><![CDATA[Blog]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Tue, 30 Jun 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<h2 id="hey-yall-howdy-asldkfjal-asdfasdf">Hey yall! howdy asl;dkfj;al asdfasdf</h2><p>i'm just testing this new blog system</p><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aiguide/content/images/imported/cms-first-post-test/8control5.png" class="kg-image" alt="buttons" loading="lazy"></figure>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[Token cost is high — investors want their money]]></title>
          <description><![CDATA[A June 2026 roundup of free &amp; cheap AI coding resources — Poolside, Mistral Vibe CLI, Nvidia NIM, and budget picks like Minimax, Deepseek v4, and StepFun Flash.]]></description>
          <link>https://wuu73.org/aiguide/06072026/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/06072026/</guid>
          <category><![CDATA[Guides]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Sun, 07 Jun 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<h1 id="%F0%9F%92%B0-higher-token-costs-reduced-credits">💰 Higher Token Costs, Reduced Credits</h1><p>Providers are hiking token prices while slashing free credits and allowances—GitHub Copilot being a prime example, where users now burn through entire monthly quotas on single tasks that previously consumed a fraction.</p><p>Luckly, there are still plenty left.</p><p>Here is some more free APIs and more coding agents (that seem actually good)</p><h2 id="%E2%9C%85-free-coding-agents-apis">✅ FREE Coding Agents, APIs</h2><h3 id="poolside">Poolside</h3><p><strong>What:</strong> Coding agent + API<br><strong>Link:</strong> <a href="https://poolside.ai/get-started?ref=wuu73.org">poolside.ai/get-started</a><br><strong>Status:</strong> Currently free</p><blockquote>I Use their CLI agent to update docs automatically while I worked on other things. Solid so far.</blockquote><h3 id="mistral-vibe-cli">Mistral Vibe CLI</h3><p><strong>What:</strong> CLI coding assistant<br><strong>Link:</strong> <a href="https://mistral.ai/?ref=wuu73.org">mistral.ai</a> - look for the Vibe CLI, it has higher rate limits vs the other stuff</p><p><strong>Install:</strong></p><p><code>curl -LsSf https://mistral.ai/vibe/install.sh | </code>bash<br></p><p><strong>Status:</strong> Free up to a request/time cap</p><p>Common knock on Mistral: "it sucks." My experience: replies instantly, handles terminal commands, automations, and scripting just fine. Perfectly adequate for lightweight tasks.</p><h3 id="nvidia-nim">Nvidia NIM</h3><p><strong>What:</strong> Free hosted models with rate limits<br><strong>Link:</strong> <a href="https://build.nvidia.com/nvidia?ref=wuu73.org">build.nvidia.com/nvidia</a><br><strong>Status:</strong> Free (rate-limited)</p><p>Haven't stress-tested the limits yet.</p><p><strong>Tool I built:</strong> An endpoint liveness checker — paste an OpenAI-compatible <code>/v1/models</code> URL (optional key), and it pings every model to log which ones respond and when. Useful for figuring out if a "free" resource is actually reliable enough to use. (Buggy right now, fix coming soon — don't use real keys yet.)</p><p>🔗 <a href="https://extra.wuu73.org/chu5?ref=wuu73.org">extra.wuu73.org/chu5</a></p><p><strong>Opencode Zen &amp; Go models:</strong> Some may work without an API key. If not, one key covers both Zen and Go — free models, zero cost. Opencode Go is a coding plan/subscription for $5/$10, I used up my entire alotment in like one week though.. with lite use</p><hr><h2 id="%F0%9F%92%B5-cheapall-close-to-or-less-than-1m">💵 CHEAP - All close to or less than $1/M</h2><p><strong>Minimax (M3 / 2.7 / 2.5)</strong> — API is extremely reliable. When I had a sub, even the lowest tier let me run tons of subagents without hitting limits. Prices may have increased; re-evaluating API vs. subscription.</p><p><strong>Deepseek v4</strong> — Free flash models using Opencode Zen's free models and some other ways like thru Cline, Kilo Code endpoints. Cheap pro/flash. Reasonix CLI agent works well! I am using it a lot.</p><p><strong>StepFun Flash 3.7</strong> — Inexpensive, strong at tool-use and agentic workflows.</p><p><strong>Arcee AI Trinity</strong> -- inexpensive models, USA based. All less than a dollar input/output per M tokens. Good with agentic tools.</p><h2 id="%F0%9F%97%82%EF%B8%8F-coding-plan-picks">🗂️ Coding Plan Picks</h2><ul><li><strong>Minimax</strong> → ⭐ Best option (if pricing/limits haven't changed)</li><li><strong>Opencode Go</strong> → Ran out in ~1 week. Raw API + free models is probably cheaper.</li></ul><h2 id="%F0%9F%A4%96-new-agent-harnesses">🤖 New Agent Harnesses</h2><ul><li><strong>Reasonix</strong> — specifically for Deepseek v4 — it is good! Been using it a lot</li><li><strong>Mistral Vibe</strong> — higher rate limits for free, haven't ran into limits yet</li><li><strong>Poolside Pool</strong> — totally free to use right now! Works good, might not be the most intelligent, but great for creating docs and doing tasks</li></ul><h2 id="%F0%9F%A7%B0-misc-tools">🧰 Misc Tools</h2><h3 id="kilo-code">Kilo Code</h3><p>VS Code extension with plenty of free models available all the time. You can even use its API endpoints in other tools/apps.</p><h3 id="cliproxyapi">CLIProxyAPI</h3><p>A tool that lets you stack providers for fallbacks and convert between OpenAI and Anthropic style endpoints, so you can use any model in Claude Code. I tried and tested this thing well — no issues with formatting the API correctly, no errors in Claude Code.</p><h3 id="buttons-cli">Buttons CLI</h3><p><a href="https://buttonscli.com/?ref=wuu73.org">buttonscli.com</a> — This terminal app can be controlled by AI coding agents via MCP or CLI, letting agents control multiple terminal tabs while you watch what they're doing (which is hard when using them in terminals since it only shows short summaries). Has agent stalling detection and nudging to get back to working.</p>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[Agent Harness Field Guide]]></title>
          <description><![CDATA[A deep dive into how Pochi, Neovate Code, Mux, Crush, Kimi Code CLI, Qwen Code, OpenHands, Claude Code, DeerFlow, Hermes Agent, Pi Mono, OpenCode, Codex CLI, Wintermolt, Zaica, Goose, Dirac, Open Claude Code, Reasonix, CodeWhale, and CheetahClaws are built: tool systems, model plumbing, prompt cachi]]></description>
          <link>https://wuu73.org/aiguide/coding-agents/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/infoblogs/coding_agents/</guid>
          <category><![CDATA[Deep Dives]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Fri, 29 May 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<p>Repository Study • 22 agents • Local snapshot</p><p>This blog audits the code that is actually present in <code>coding-agents\</code>: how each agent wires models, exposes tools, runs shell commands, speaks MCP or ACP, and where Claude Code and Hermes feel fundamentally different from the rest of the field — plus Oh My Pi's hashline-heavy Bun + Rust fork of Pi Mono, OpenCode's client-server runtime, Reasonix's exact-match edit gate, ADK-Rust's feature-gated Rust framework stack, CheetahClaws's Python-native daemon-plus-kernel architecture, and OpenAI's own Codex CLI, a Rust-native agent with platform-specific sandboxes and bidirectional MCP support.</p><figure class="kg-card kg-image-card kg-card-hascaption"><img src="https://wuu73.org/aiguide/content/images/imported/coding-agents/assets/diagrams/architecture-families.png" class="kg-image" alt="Hand-drawn diagram grouping coding agents into bespoke runtimes, provider multiplexers, protocol bridges, orchestration platforms, and special cases around a local repo snapshot." loading="lazy" width="1376" height="768" srcset="https://wuu73.org/aiguide/content/images/size/w600/imported/coding-agents/assets/diagrams/architecture-families.png 600w, https://wuu73.org/aiguide/content/images/size/w1000/imported/coding-agents/assets/diagrams/architecture-families.png 1000w, https://wuu73.org/aiguide/content/images/imported/coding-agents/assets/diagrams/architecture-families.png 1376w" sizes="(min-width: 720px) 720px"><figcaption>The repo set breaks into a handful of recurring runtime shapes. <a href="https://wuu73.org/aiguide/coding-agents-methods/">Architecture</a> goes deep on the families, while <a href="https://wuu73.org/aiguide/coding-agents-protocols/">Protocols</a> adds the MCP versus ACP map.</figcaption></figure><p>Shameless Plug<a href="https://buttonscli.com/?ref=wuu73.org" rel="noopener">While we're deep-diving into complex CLI tools... let's be honest. Are you tired of pretending you enjoy memorizing 400 arcane tmux key combos? Do you lie awake at night dreading the moment you have to edit a cryptic dotfile just to change a font size? Same. It's the AI age, we shouldn't be computing like it's 1995.<strong>ButtonsCLI</strong>: The terminal for</a></p><p>(Alright, ad over. Back to the serious technical analysis.)</p><h2 id="what-this-deep-dive-covers">What this deep dive covers</h2><p>The interesting part of coding agents is not the marketing layer. It is the runtime beneath it: whether the tool layer is generic or bespoke, whether shell access is tightly guarded or casually wrapped, whether model support is truly abstracted or just superficially multiplexed, and whether the repo reads like a productized operating environment or a fast-moving integration shell.</p><p>I read the local repositories for <strong>Pochi</strong>, <strong>Neovate Code</strong>, <strong>Mux</strong>, <strong>Crush</strong>, <strong>Kimi CLI</strong>, <strong>Qwen Code</strong>, <strong>OpenHands</strong>, <strong>Claude Code</strong>, <strong>DeerFlow</strong>, <strong>Hermes Agent</strong> (by Nous Research), <strong>Pi Mono</strong>, <strong>Oh My Pi</strong>, <strong>OpenCode</strong>, <strong>Codex CLI</strong> (by OpenAI), <strong>Wintermolt</strong>, <strong>Zaica</strong>, <strong>Goose</strong>, <strong>Dirac</strong>, <strong>Reasonix</strong>, <strong>CodeWhale</strong>, <strong>CheetahClaws</strong>, and <strong>Open Claude Code 2.0</strong> (a clean-room implementation via AI decompilation), then mapped them against the same questions on every page of this site.</p><p>🔍</p><h4 id="methodology-note">Methodology note</h4><p>This is a <strong>local-only repo study</strong>. I did not use outside documentation beyond what is already checked into the worktree. That matters most for OpenHands, where the repo itself says the newer V1 agent core now lives elsewhere.</p><h2 id="repos-audited">Repos audited</h2><ul><li><strong>Pochi</strong> - TypeScript monorepo with vendor-specific model adapters and built-in agents.</li><li><strong>Neovate Code</strong> - TypeScript CLI with AI SDK providers, MCP, and hardened bash tooling.</li><li><strong>Mux</strong> - large TypeScript desktop/browser agent platform with workspaces and provider routing.</li><li><strong>Crush</strong> - Go-based terminal product with custom tools, permissions, and provider metadata plumbing.</li><li><strong>Kimi CLI</strong> - Python terminal agent focused on Kimi platform flows plus ACP and MCP bridges.</li><li><strong>Qwen Code</strong> - Gemini CLI descendant with strong config resolution, declarative tools, and MCP lifecycle management.</li><li><strong>OpenHands</strong> - platform/runtime repo with sandbox and legacy CodeAct agent pieces still present locally.</li><li><strong>Claude Code</strong> - deeply integrated Bun + React/Ink runtime centered on Anthropic models and a huge internal tool surface.</li><li><strong>DeerFlow</strong> - LangGraph-based super-agent harness with middleware, subagents, SSE streaming, and config-driven model factories.</li><li><strong>Hermes Agent</strong> (Nous Research) - Python agent with persistent skill learning, <strong>14+ messaging platform gateways</strong> (Telegram, Discord, Slack, WhatsApp, Signal, WeChat, Matrix, Mattermost, Feishu, DingTalk, Email, SMS, HomeAssistant, Webhook), MoA synthesis, 6 execution backends, and RL training infrastructure.</li><li><strong>Codex CLI</strong> (OpenAI) - Rust-native coding agent (3,805 files, 70+ crates) with macOS Seatbelt, Linux bubblewrap/Landlock, Windows sandbox, MCP client <em>and</em> server, multi-agent spawning, OpenAI Responses API plumbing, configurable TOML config, IDE extensions, and a Ratatui TUI.</li><li><strong>Pi Mono</strong> - TypeScript minimalist harness (874 files) with tree-structured JSONL v3 sessions, differential TUI rendering, Pi Packages (shareable bundles via npm/git), 26 providers across 10 APIs, parallel tool execution, file mutation queue, 4 run modes, MIT license, author Mario Zechner.</li><li><strong>Oh My Pi</strong> - Bun + TypeScript + Rust fork of Pi Mono (<strong>v14.7.8</strong>, MIT) with default hashline editing, a Rust native engine, MCP and plugin discovery, LSP, Python, browser tooling, task/subagent execution, and a dedicated edit-benchmark package.</li><li><strong>OpenCode</strong> - TypeScript/Bun client-server coding runtime (4,531 files) with a SolidJS opentui terminal UI, web and desktop clients, <code>apply_patch</code> plus exact-string editing, wildcard permissions, worktrees, MCP, ACP, skills, and an SDK.</li><li><strong>ADK-Rust</strong> - Rust framework workspace (<strong>v0.8.0</strong>, Apache-2.0, <strong>34 workspace members</strong>) with feature-tiered crates, a minimal LLM trait, provider-native Anthropic/OpenAI tool wrappers, graph workflows, MCP integration, A2A APIs, AWP deployment support, and optional browser/realtime/sandbox modules.</li><li><strong>Wintermolt</strong> - Zig 0.15 native binary (3 MB, zero runtime) with 6 AI backends, 16 tools, cron scheduling, Tailscale mesh, camera vision, browser automation, MCP bidirectional, chat bridges to 4 platforms, and a macOS menu bar app.</li><li><strong>Zaica</strong> - Zig 0.15 focused coding agent (~9,100 lines, zero runtime) with multi-provider LLM support, chain-mode structured workflows, parallel sub-agent dispatch, reactive state management (zefx), and Wyhash-based loop detection.</li><li><strong>Goose</strong> (AAIF/Linux Foundation) - Rust-native AI agent (v1.32.0, Apache-2.0) with 15+ providers, 5-layer security inspector stack, LLM-based AdversaryInspector, 4 GooseModes (Auto/Approve/SmartApprove/Chat), extension system, recipe framework, and MOIM injection.</li><li><strong>Open Claude Code 2.0</strong> - Clean-room implementation of Claude Code via AI-powered decompilation (1,581 tests, async generator architecture, multi-agent teams, git worktree isolation, 7-type hook system).</li><li><strong>Dirac</strong> - TypeScript fork of Cline with hash-anchored parallel edits, AST-native precision, multi-file batching, 64.8% cost reduction, no MCP, hook system, git checkpoints, state mutex, and 40+ provider support. 8/8 on TerminalBench 2.0 evals.</li></ul><h2 id="the-landscape-in-one-screen">The landscape in one screen</h2><p>🧠</p><h3 id="bespoke-runtime-products">Bespoke runtime products</h3><p><strong>Claude Code</strong> and <strong>Crush</strong> feel like full terminal operating environments, not thin wrappers. Their tool, permission, and UX layers are part of the product, not just adapters around a chat loop. </p><p>Most opinionated🔌</p><h3 id="provider-multiplexers">Provider multiplexers</h3><p><strong>Mux</strong>, <strong>Neovate</strong>, and <strong>Qwen Code</strong> all build serious provider catalogs and shared abstractions. They want broad model reach more than a single model-native identity. </p><p>Most configurable🧩</p><h3 id="protocol-and-adapter-layers">Protocol and adapter layers</h3><p><strong>Pochi</strong> and <strong>Kimi CLI</strong> stand out for their ecosystem bridges. Pochi ships vendor-specific packages for Codex, Qwen, Copilot, and others. Kimi invests heavily in ACP and in translating internal tool output into protocol-friendly shapes. </p><p>Most bridge-heavy🏗️</p><h3 id="client-server-coding-runtimes">Client-server coding runtimes</h3><p><strong>OpenCode</strong> feels less like a single CLI and more like a shared runtime: one backend powering TUI, web, desktop, SDK, MCP, ACP, skills, worktrees, and a permission bus. It is the clearest open-source example here of a terminal agent growing into a platform. </p><p>Most runtime-shaped🧱</p><h3 id="frameworks-and-execution-harnesses">Frameworks and execution harnesses</h3><p><strong>DeerFlow</strong>, <strong>OpenHands</strong>, and <strong>ADK-Rust</strong> are less about one CLI persona and more about reusable orchestration environments. DeerFlow and OpenHands lean into services and sandboxes. ADK-Rust pushes furthest toward a publishable crate ecosystem with feature tiers, graph workflows, provider-native tools, A2A, AWP, and payment adapters. </p><p>Most framework-shaped</p><h3 id="the-pi-lineage-kernel-to-power-tool-fork">The Pi lineage: kernel to power-tool fork</h3><p><strong>Pi Mono</strong> is still the sharp minimalist kernel: tree-structured JSONL sessions, differential TUI rendering, Pi Packages, and a tight extension-first philosophy. <strong>Oh My Pi</strong> takes that lineage in the opposite direction with Bun workspaces, Rust native modules, default hashline edits, MCP, plugins, browser tooling, a task tool, and an edit benchmark lab. </p><p>Most visible fork split</p><h3 id="self-improving-multi-platform-agents">Self-improving multi-platform agents</h3><p><strong>Hermes Agent</strong> (Nous Research) is in a category of its own: a persistent skill-learning loop, six remote execution backends, <strong>14+ messaging platform</strong> gateways (Telegram, Discord, Slack, WhatsApp, Signal, WeChat, Matrix, Mattermost, Feishu, DingTalk, Email, SMS, HomeAssistant, Webhook), MoA synthesis, and RL training infrastructure. </p><p>Most functionally unique</p><h3 id="zero-runtime-native-binaries">Zero-runtime native binaries</h3><p><strong>Wintermolt</strong> and <strong>Zaica</strong> are both written entirely in Zig 0.15 — no Node.js, no Python, no garbage collector. They compile to single native binaries (Wintermolt is 3 MB) with cross-compilation to any Zig target including ARM boards. Wintermolt goes wide (7 modes, cron, Tailscale, camera, browser, MCP, chat bridges). Zaica goes deep (chain-mode workflows, reactive state, Wyhash loop detection). </p><p>Most portable⚡</p><h3 id="openais-own-coding-agent">OpenAI's own coding agent</h3><p><strong>Codex CLI</strong> is OpenAI's production coding agent — a Rust workspace of 70+ crates with platform-specific sandboxes (Seatbelt, bubblewrap/Landlock, Windows restricted tokens), bidirectional MCP (client <em>and</em> server), multi-agent job execution, and a configurable provider system that supports Ollama and LM Studio. </p><p>Most sandboxed</p><h3 id="extension-first-security-advocates">Extension-first security advocates</h3><p><strong>Goose</strong> (AAIF at Linux Foundation) takes a unique approach to security: it uses an LLM-based <code>AdversaryInspector</code> that fires a second LLM call to review tool calls against user-defined rules from <code>~/.config/goose/adversary.md</code>. This is defense-in-depth for multi-agent setups where parent agents delegate to sub-agents. With 15+ providers, 4 GooseModes (Auto/Approve/SmartApprove/Chat), and a recipe framework, Goose is Rust-native with extensive feature gates (local-inference, aws-providers, otel). </p><p>Most LLM-secured</p><h2 id="fast-takeaways">Fast takeaways</h2>
<!--kg-card-begin: html-->
<table>
<thead>
<tr>
<th>Question</th>
<th>Best answer from this snapshot</th>
<th>Why</th>
</tr>
</thead>
<tbody>
<tr>
<td><strong>Which repo feels most different?</strong></td>
<td>Claude Code</td>
<td>
                  It is the least generic and the most integrated:
                  Anthropic-first, huge tool catalog, plan/worktree/team flows,
                  React terminal UI, permission system, and a massive central
                  query runtime.
                </td>
</tr>
<tr>
<td><strong>Which repos are most model-agnostic?</strong></td>
<td>Mux, Neovate, Qwen Code, Goose</td>
<td>
                  All invest in provider registries, routing layers, and shared
                  config resolution instead of pinning themselves to one native
                  model family. Goose ships 35+ provider modules across direct
                  APIs, ACP bridges, and declarative JSON configs.
                </td>
</tr>
<tr>
<td>
<strong>Which repo adapts multiple ecosystems most
                    explicitly?</strong>
</td>
<td>Pochi</td>
<td>
                  It does not stop at a generic provider interface; it ships
                  vendor-specific packages for Codex, Qwen Code, GitHub Copilot,
                  Gemini CLI, and more.
                </td>
</tr>
<tr>
<td>
<strong>Which repo is most reusable as a framework?</strong>
</td>
<td>ADK-Rust</td>
<td>
                  It is organized as a 34-member Rust workspace with a minimal
                  default tier, publishable crates for
                  agents/models/tools/server, and explicit add-on layers for
                  graph workflows, payments, AWP, browser automation, realtime,
                  and sandboxing.
                </td>
</tr>
<tr>
<td>
<strong>Which repo treats edit reliability like an engineering
                    lab?</strong>
</td>
<td>Oh My Pi</td>
<td>
                  Default hashline mode, prompt helpers for real anchors,
                  separator tuning, compact previews, a dedicated
                  <code>typescript-edit-benchmark</code> package, and benchmark
                  scripts for the hashline variant all point to a repo that is
                  explicitly iterating on mechanical edit failure modes.
                </td>
</tr>
<tr>
<td>
<strong>Which shell tooling is most safety-conscious?</strong>
</td>
<td>Neovate Code, Claude Code, and Goose</td>
<td>
                  Neovate hard-codes command bans and high-risk detection
                  (22-item banned list, quote-aware pipeline parser), while
                  Claude layers permissions, tree-sitter AST analysis, and
                  Zsh-specific attack detection over a richer command surface.
                  Hermes uses supply-chain verification (cosign provenance) for
                  its execution environment. Goose uniquely uses an LLM-based
                  AdversaryInspector that fires a second LLM call to review tool
                  calls against user-defined rules.
                </td>
</tr>
<tr>
<td>
<strong>Which repo has the most unique capabilities?</strong>
</td>
<td>Hermes Agent</td>
<td>
                  Self-improving skill loop, 6 remote backends, multi-platform
                  IM gateways, MoA synthesis across 4 frontier models, and RL
                  training infrastructure — none of which appear anywhere else
                  in this set.
                </td>
</tr>
<tr>
<td>
<strong>Which repo is hardest to judge from local code
                    alone?</strong>
</td>
<td>OpenHands</td>
<td>
                  The local repo still contains useful architecture, but its own
                  docs say the newer V1 agent core moved to a separate Software
                  Agent SDK repository.
                </td>
</tr>
<tr>
<td><strong>Which code feels most polished?</strong></td>
<td>Claude Code, Crush, Mux, Qwen Code</td>
<td>
                  These four snapshots show the clearest internal consistency
                  between product goals, tool design, configuration, and error
                  handling. Crush is notable for being the only agent with
                  native LSP diagnostics and Sourcegraph code search as
                  first-class tools.
                </td>
</tr>
<tr>
<td>
<strong>Which repo is the most extension-friendly?</strong>
</td>
<td>Pi Mono</td>
<td>
                  A tight minimalist harness that deliberately ships without
                  MCP, permissions, or sub-agents — expecting you to compose
                  them via extensions. Pi Packages let you bundle and share
                  configurations across projects via npm or git.
                </td>
</tr>
<tr>
<td>
<strong>Which repo feels most like a platform runtime?</strong>
</td>
<td>OpenCode</td>
<td>
                  One backend drives the TUI, browser console, desktop shell,
                  MCP, ACP, skills, worktrees, and SDK. It reads more like a
                  small agent platform than a one-window CLI.
                </td>
</tr>
<tr>
<td>
<strong>Which repo has the most platform-specific
                    sandboxing?</strong>
</td>
<td>Codex CLI</td>
<td>
                  Three separate sandbox implementations — macOS Seatbelt, Linux
                  bubblewrap/Landlock, and Windows restricted tokens — each with
                  split-filesystem awareness and carveout support. Also the only
                  agent in this set that doubles as an MCP server for other
                  agents.
                </td>
</tr>
<tr>
<td>
<strong>Which repo handles web-grounded research best?</strong>
</td>
<td>DeerFlow overall; Crush if you want no extra search API bill</td>
<td>
                  DeerFlow combines free default search/fetch with the most
                  explicit deep-research methods in the current snapshot. Crush
                  is the strongest product-style coding agent that pairs free
                  DuckDuckGo search with a delegated multi-step fetch workflow.
                  See <a href="https://wuu73.org/aiguide/coding-agents-web-research/">Web Research</a> for the full
                  comparison, including Wintermolt, Claude Code, Codex, and
                  Reasonix.
                </td>
</tr>
</tbody>
</table>
<!--kg-card-end: html-->
<h2 id="approximate-codebase-size-by-file-count">Approximate codebase size by file count</h2><p>File count is not the same thing as quality, but it does reveal where the implementation surface is broadest.</p><p>ADK-Rust is intentionally omitted here: its 34-member framework workspace and extracted companion repos make raw file-count comparison less useful than with single-product CLIs.</p><h4 id="opencode">OpenCode</h4><p><strong>4531 files</strong></p><h4 id="codex-cli">Codex CLI</h4><p><strong>3805 files</strong></p><h4 id="openhands">OpenHands</h4><p><strong>2774 files</strong></p><h4 id="mux">Mux</h4><p><strong>2226 files</strong></p><h4 id="claude-code">Claude Code</h4><p><strong>2137 files</strong></p><h4 id="qwen-code">Qwen Code</h4><p><strong>2038 files</strong></p><h4 id="pochi">Pochi</h4><p><strong>1315 files</strong></p><h4 id="kimi-cli">Kimi CLI</h4><p><strong>899 files</strong></p><h4 id="deerflow">DeerFlow</h4><p><strong>810 files</strong></p><h4 id="crush">Crush</h4><p><strong>799 files</strong></p><h4 id="neovate">Neovate</h4><p><strong>582 files</strong></p><h4 id="hermes">Hermes</h4><p><strong>~450 files</strong></p><h4 id="pi-mono">Pi Mono</h4><p><strong>874 files</strong></p><h4 id="wintermolt">Wintermolt</h4><p><strong>51 Zig files (~18,400 lines)</strong></p><h4 id="zaica">Zaica</h4><p><strong>~13 files (~9,100 lines)</strong></p><h4 id="open-claude-code-20">Open Claude Code 2.0</h4><p><strong>61 files (~8,300 lines)</strong></p><h4 id="goose">Goose</h4><p><strong>Rust Cargo workspace (~6+ crates)</strong></p><h4 id="dirac">Dirac</h4><p><strong>TypeScript monorepo (fork of Cline)</strong></p><h2 id="my-high-level-verdict">My high-level verdict</h2><h3 id="best-designed-if-you-value-a-coherent-product-runtime">Best designed, if you value a coherent product runtime</h3><p><strong>Claude Code</strong> is the standout. It is not the most provider-flexible repo, but it is the clearest example of an agent built as its own operating model: tool schemas, permissioning, commands, tasking, worktrees, UI, feature flags, and retry logic all sit inside one deliberate runtime.</p><h3 id="best-designed-if-you-value-clean-systems-engineering">Best designed, if you value clean systems engineering</h3><p><strong>Crush</strong> is the nicest surprise. The Go codebase feels disciplined, modular, and product-minded without being bloated. Its provider plumbing, permissions, and TUI organization are easier to reason about than many faster-moving TypeScript peers.</p><h3 id="best-multi-model-architecture">Best multi-model architecture</h3><p><strong>Mux</strong> and <strong>Qwen Code</strong> lead here. Mux has a broad provider routing layer with desktop app ambitions, while Qwen Code has a particularly strong configuration and runtime model-resolution story.</p><h3 id="most-extensible-framework-shape">Most extensible framework shape</h3><p><strong>DeerFlow</strong> wins on composability. It feels more like a harness for building agent systems than a single agent persona, which makes it powerful but also less opinionated than Claude Code or Crush.</p><h3 id="most-publishable-framework-ecosystem">Most publishable framework ecosystem</h3><p><strong>ADK-Rust</strong>. The repo is built as a crate ecosystem, not just a runnable app: minimal-by-default packaging, typed tools, workflow agents, graph orchestration, A2A/AWP surfaces, and an honest stability file that separates mature crates from frontier modules.</p><h3 id="most-functionally-unique">Most functionally unique</h3><p><strong>Hermes Agent</strong> by Nous Research. The self-improving skill loop, 14+ messaging platform gateways, MoA tool (4 frontier models in parallel), and RL training infrastructure are not features in any other repo here. It is the only agent that explicitly tries to get better at your tasks over time.</p><h3 id="most-portable-%E2%80%94-zero-runtime-one-binary">Most portable — zero runtime, one binary</h3><p><strong>Wintermolt</strong> and <strong>Zaica</strong> are the only agents here that compile to a single native binary with zero runtime dependency. Wintermolt (3 MB, 18,400 lines) is the most ambitious agent in any language. Zaica (~9,100 lines) is the most focused coding specialist with chain-mode workflows and best-in-class loop detection.</p><h3 id="most-extension-friendly-kernel">Most extension-friendly kernel</h3><p><strong>Pi Mono</strong> by Mario Zechner. A tight 874-file TypeScript kernel with tree-structured JSONL v3 sessions, differential TUI rendering, 26 providers across 10 APIs, Pi Packages (shareable bundles via npm/git), parallel tool execution, a file mutation queue, and 4 run modes. MIT licensed and deliberately minimal so you can build MCP, permissions, or sub-agents yourself.</p><h3 id="most-serious-edit-lab-fork">Most serious edit-lab fork</h3><p><strong>Oh My Pi</strong>. It starts from Pi Mono's terminal-agent kernel and then turns editing into a research surface: default hashline mode, prompt/runtime anchor helpers, a native engine, MCP and plugin plumbing, built-in task execution, and benchmark infrastructure dedicated to edit variants.</p><h3 id="most-complete-open-source-runtime">Most complete open-source runtime</h3><p><strong>OpenCode</strong>. The repo combines a terminal UI, browser console, desktop shell, SDK, MCP, ACP, worktrees, skills, and a permission bus behind one backend runtime. It is the clearest open-source "agent platform" in this snapshot.</p><h3 id="most-security-conscious-sandboxing">Most security-conscious sandboxing</h3><p><strong>Codex CLI</strong> by OpenAI. Three platform-specific sandbox implementations (macOS Seatbelt, Linux bubblewrap/Landlock, Windows restricted tokens), split-filesystem awareness, an execution policy engine with a rule DSL, bidirectional MCP (client and server), and a strict clippy lint policy that bans <code>unwrap_used</code> and <code>expect_used</code> across 70+ crates.</p><h2 id="how-to-read-the-rest-of-this-site">How to read the rest of this site</h2><p>1<a href="https://wuu73.org/aiguide/coding-agents-methods/"><strong>Architecture</strong></a></p><p>Compare tool schemas, shell execution, MCP support, and recovery patterns.</p><p>↓2<a href="https://wuu73.org/aiguide/coding-agents-agents/"><strong>Agents</strong></a></p><p>Read per-repo profiles, strengths, weaknesses, and fit.</p><p>↓3<a href="https://wuu73.org/aiguide/coding-agents-prompts/"><strong>Models</strong></a></p><p>See who is genuinely provider-neutral and who writes model-specific logic.</p><p>↓4<a href="https://wuu73.org/aiguide/coding-agents-implementation/"><strong>Claude Code</strong></a></p><p>The dedicated page on why Claude Code feels like a category of its own.</p><p>↓5<a href="https://wuu73.org/aiguide/coding-agents-deerflow/"><strong>DeerFlow</strong></a></p><p>The LangGraph-based super agent harness with 14-layer middleware, skill evolution, and sub-agent orchestration.</p><p>↓6<a href="https://wuu73.org/aiguide/coding-agents-openhands/"><strong>OpenHands</strong></a></p><p>The platform-shaped agent with Docker sandboxing and ingenious temperature-bumping retry logic.</p><p>↓7<a href="https://wuu73.org/aiguide/coding-agents-security/"><strong>Security</strong></a></p><p>Deep dive into shell injection defense, prompt injection scanning, permissions, sandboxing, and loop detection.</p><p>↓8<a href="https://wuu73.org/aiguide/coding-agents-protocols/"><strong>Protocols</strong></a></p><p>MCP and ACP implementation compared — transports, OAuth, lifecycle, and deferred tool loading.</p><p>↓9<a href="https://wuu73.org/aiguide/coding-agents-subagents/"><strong>Subagents</strong></a></p><p>How agents delegate work, isolate children, enforce concurrency limits, and collect results.</p><p>↓10<a href="https://wuu73.org/aiguide/coding-agents-hermes/"><strong>Hermes Agent</strong></a></p><p>The completely separate deep dive on the most unusual agent in the set — self-improving, multi-platform, and RL-augmented.</p><p>↓11<a href="https://wuu73.org/aiguide/coding-agents-pi/"><strong>Pi Mono</strong></a></p><p>The minimalist kernel — 874 files, tree-structured JSONL v3 sessions, differential TUI, Pi Packages, 26 providers across 10 APIs, parallel tool execution, file mutation queue, 4 run modes, MIT licensed.</p><p>↓12<a href="https://wuu73.org/aiguide/coding-agents-opencode/"><strong>OpenCode</strong></a></p><p>The client-server runtime — SolidJS terminal UI, web and desktop clients, <code>apply_patch</code>, wildcard permissions, ACP, MCP, worktrees, and skills.</p><p>↓13<a href="https://wuu73.org/aiguide/coding-agents-adk-rust/"><strong>ADK-Rust</strong></a></p><p>The Rust framework workspace: 34 member crates, feature tiers, provider-native tool wrappers, graph workflows, A2A, AWP, and optional sandbox/browser/realtime modules.</p><p>↓14<a href="https://wuu73.org/aiguide/coding-agents-codex/"><strong>Codex CLI</strong></a></p><p>OpenAI's production agent: 3,805 files, 70+ Rust crates, three platform-specific sandboxes, bidirectional MCP, multi-agent jobs, and IDE extensions.</p><p>↓15<a href="https://wuu73.org/aiguide/coding-agents-wintermolt/"><strong>Wintermolt</strong></a></p><p>The 3 MB everything-agent: 6 backends, 16 tools, cron, Tailscale, camera, browser, MCP, chat bridges, and a macOS menu bar app.</p><p>↓16<a href="https://wuu73.org/aiguide/coding-agents-zaica/"><strong>Zaica</strong></a></p><p>The focused specialist: chain-mode workflows, reactive state management, Wyhash loop detection, and a hand-crafted terminal REPL.</p><p>↓17<a href="https://wuu73.org/aiguide/coding-agents-goose/"><strong>Goose</strong></a></p><p>The extension-first Rust agent: LLM-based AdversaryInspector, 4 GooseModes, 15+ providers, recipe framework, and MOIM injection.</p><p>↓18<a href="https://wuu73.org/aiguide/coding-agents-zig-agents/"><strong>Zig Agents</strong></a></p><p>Head-to-head comparison: two agents, one language, opposite philosophies — platform vs. specialist, 18,400 lines vs. ~9,100.</p><p>↓19<a href="https://wuu73.org/aiguide/coding-agents-openclaudecode/"><strong>Open Claude Code</strong></a></p><p>Clean-room rebuild of Claude Code v2.1.91 via ruDevolution decompilation: async generator loop, 25 tools, 5 providers, nightly releases.</p><p>→20<a href="https://wuu73.org/aiguide/coding-agents-dirac/"><strong>Dirac</strong></a></p><p>Hash-anchored parallel edits, AST-native precision, 64.8% cost reduction vs competitors, no MCP, 8-type hook system, git checkpoints.</p><p>→21<a href="https://wuu73.org/aiguide/coding-agents-reasonix/"><strong>Reasonix</strong></a></p><p>DeepSeek-native coding agent with byte-exact SEARCH/REPLACE, edit-gate review, repair stages, and strict sandbox enforcement.</p><p>→22<a href="https://wuu73.org/aiguide/coding-agents-codewhale/"><strong>CodeWhale</strong></a></p><p>DeepSeek-first Rust agent with a constitution prompt, durable task manager, persistent subagents, runtime APIs, and a transactional edit stack.</p><p>→23<a href="https://wuu73.org/aiguide/coding-agents-cheetahclaws/"><strong>CheetahClaws</strong></a></p><p>Python-native multi-provider agent with a plugin registry, daemon server, capability-gated kernel, MCP plumbing, and a mixed direct-write plus diff editing stack.</p>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[How Agent Harnesses Edit Files]]></title>
          <description><![CDATA[A field guide to how agent harnesses — Cline, Codex, OpenCode, ADK-Rust, Reasonix, CodeWhale, CheetahClaws, Oh My Pi, Crush, DeerFlow, Dirac, Goose, OpenHands, Pi, Pochi, Qwen Code, Wintermolt, Zaica, and more — actually edit files in your codebase.]]></description>
          <link>https://wuu73.org/aiguide/coding-file-edits/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/infoblogs/coding_file_edits/</guid>
          <category><![CDATA[Deep Dives]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Fri, 29 May 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<p>🔧 Developer Guide • Updated April 2026</p><p>How do AI coding assistants actually edit your files? Discover the strategies, matching algorithms, fallback mechanisms, and provider-native, exact-match, capability-gated, or anchor-native editing contracts behind Cline, Codex, OpenCode, ADK-Rust, Reasonix, CodeWhale, CheetahClaws, Oh My Pi, Aider, Crush, DeerFlow, Dirac, Goose, OpenHands, Pochi, Qwen Code, and more.</p><h2 id="why-does-this-matter">Why Does This Matter?</h2><p>When you ask an AI assistant to "add a new function" or "fix this bug," there's a critical gap between the LLM's text generation and your actual filesystem. The AI outputs text—but how does that become a working code change?</p><p>Every AI coding agent solves this differently. Some use <strong>surgical search-and-replace</strong>, others apply <strong>unified diffs</strong>, some just <strong>overwrite entire files</strong>, and a small but important set use <strong>hash-backed anchors</strong> to survive drift. Dirac does it with cryptographic word anchors; Oh My Pi does it with compact line-plus-hash markers. Understanding these approaches matters if you're:</p><ul><li>Building your own AI coding tool</li><li>Debugging why an AI edit failed</li><li>Choosing between different AI assistants</li><li>Curious about the engineering behind these tools</li></ul><h2 id="the-core-challenge">The Core Challenge</h2><p>💡</p><h4 id="llms-are-non-deterministic">LLMs Are Non-Deterministic</h4><p>Even with the same prompt, an AI might output slightly different whitespace, comments, or formatting each time. A file editing system must handle these variations gracefully—or fail constantly.</p><p>Consider what can go wrong:</p><ul><li><strong>Whitespace mismatches:</strong> The AI uses 2 spaces, but your file uses tabs</li><li><strong>Hallucinated content:</strong> The AI "remembers" code that doesn't exist</li><li><strong>Partial context:</strong> The AI only saw 50 lines but the file has 500</li><li><strong>Race conditions:</strong> You edited the file while the AI was generating</li></ul><h2 id="the-landscape-at-a-glance">The Landscape at a Glance</h2><p>We analyzed how the agents in <code>repos/</code> handle file editing, plus a few adjacent reference agents. Each takes a different philosophical approach:</p><p>🎯</p><h3 id="cline">Cline</h3><p><strong>The Precision Specialist</strong></p><p>Uses search/replace with a 4-tier matching strategy. Doesn't trust AI whitespace—implements heavy fallback logic. </p><p>4-Tier Fallback🔧</p><h3 id="codex-claude-code">Codex / Claude Code</h3><p><strong>The Patch Master</strong></p><p>Uses custom patch syntax with <code>*** Begin Patch</code> markers. Treats file editing like version control. </p><p>Custom Patch Format🛠️</p><h3 id="opencode">OpenCode</h3><p><strong>The Fallback King</strong></p><p>Implements <strong>9 different matching algorithms</strong> tried sequentially. Maximum redundancy approach. </p><p>9-Layer Fallback⚡</p><h3 id="aider">Aider</h3><p><strong>The Format Flexible</strong></p><p>Supports three formats: SEARCH/REPLACE, whole file, and unified diff. Model-specific defaults. </p><p>3 Edit Formats🚀</p><h3 id="grok-cli">Grok CLI</h3><p><strong>The Dual-Mode Agent</strong></p><p>Traditional text editor + Morph AI-powered fast editing. User confirmation with diff previews. </p><p>4,500+ tok/sec💎</p><h3 id="crush-neovate">Crush / Neovate</h3><p><strong>The Redundant Matchers</strong></p><p>JSON tool schemas with edit/write/multiedit tools and layered string-matching fallbacks for stubborn whitespace and escaping. </p><p>Layered Replacement📌</p><h3 id="dirac">Dirac</h3><p><strong>The Hash-Anchored Editor</strong></p><p>Edits target stable line hashes instead of line numbers, then batch multi-file changes in reverse order to avoid drift. </p><p>Hash Anchors🧪</p><h3 id="oh-my-pi">Oh My Pi</h3><p><strong>The Hashline Lab</strong></p><p>Makes hashline the default edit mode: compact line-plus-hash anchors, multi-section preflight checks, same-path merge logic, nearby anchor rebasing, and benchmark-driven iteration. </p><p>Hashline Default🧩</p><h3 id="openhands-claude-style">OpenHands / Claude-style</h3><p><strong>The Standard Editor Interface</strong></p><p>Uses <code>str_replace_editor</code>-style commands: view, create, str_replace, insert, and undo, usually inside a sandbox. </p><p>Editor Command API🏗️</p><h3 id="adk-rust">ADK-Rust</h3><p><strong>The Provider Delegate</strong></p><p>Wraps provider-native editor contracts instead of building a giant local matcher. Anthropic tools execute strict exact-match edits; OpenAI <code>apply_patch</code> is surfaced as a native built-in. </p><p>Native Tool Contracts⚙️</p><h3 id="pi-pochi-qwen">Pi / Pochi / Qwen</h3><p><strong>The Exact-Match Pragmatists</strong></p><p>Prefer explicit old/new content with uniqueness checks, preview diffs, encoding preservation, and clear failure messages. </p><p>Exact + Verified🔌</p><h3 id="goose-deerflow-hermes">Goose / DeerFlow / Hermes</h3><p><strong>The Tool-Host Agents</strong></p><p>File mutation often lives in extensions, sandbox tools, or MCP-like tool registries rather than a single built-in editor primitive. </p><p>Extension/Sandbox Tools</p><h2 id="the-six-core-editing-methods">The Six Core Editing Methods</h2><p>Across all agents we analyzed, file editing boils down to these six fundamental approaches:</p>
<!--kg-card-begin: html-->
<table>
<thead>
<tr>
<th>Method</th>
<th>Token Cost</th>
<th>Reliability</th>
<th>Best For</th>
<th>Used By</th>
</tr>
</thead>
<tbody>
<tr>
<td><strong>Whole File Replacement</strong></td>
<td><span class="badge badge-red">High</span></td>
<td><span class="badge badge-green">100%</span></td>
<td>New files, small files, last resort</td>
<td>All agents (fallback)</td>
</tr>
<tr>
<td><strong>Search &amp; Replace</strong></td>
<td><span class="badge badge-green">Low</span></td>
<td><span class="badge badge-orange">Medium</span></td>
<td>Targeted edits, function changes</td>
<td>
                  Cline, Aider, OpenCode, Crush, Pochi, Qwen Code, Neovate,
                  ADK-Rust (Anthropic wrappers)
                </td>
</tr>
<tr>
<td><strong>Unified Diff / Patch</strong></td>
<td><span class="badge badge-green">Very Low</span></td>
<td><span class="badge badge-orange">Variable</span></td>
<td>Multi-file refactors, trained models</td>
<td>
                  Codex, Aider, Claude Code/OpenClaudeCode, shell-based agents
                </td>
</tr>
<tr>
<td><strong>Line-Based / Anchor</strong></td>
<td><span class="badge badge-blue">Medium</span></td>
<td><span class="badge badge-blue">Good</span></td>
<td>When exact match fails</td>
<td>OpenCode, Cline (fallback)</td>
</tr>
<tr>
<td><strong>Multi-Edit / Atomic</strong></td>
<td><span class="badge badge-blue">Medium</span></td>
<td><span class="badge badge-green">High</span></td>
<td>Variable renames, bulk changes</td>
<td>OpenCode</td>
</tr>
<tr>
<td><strong>Hash-Backed Anchors</strong></td>
<td><span class="badge badge-blue">Low-Medium</span></td>
<td><span class="badge badge-green">High</span></td>
<td>Repeated surgical edits in drifting files</td>
<td>Dirac, Oh My Pi</td>
</tr>
</tbody>
</table>
<!--kg-card-end: html-->
<h2 id="the-secret-sauce-fallback-cascades">The "Secret Sauce": Fallback Cascades</h2><p>The top-performing agents don't rely on a single method. They implement <strong>cascading fallbacks</strong>—if one approach fails, they automatically try the next.</p><p>1<strong>Exact Match</strong></p><p>Try byte-for-byte string matching. Fastest, most reliable when it works. </p><p>↓2<strong>Whitespace Flexible</strong></p><p>Normalize spaces/tabs, trim lines. Handles indentation differences. </p><p>↓3<strong>Anchor Matching</strong></p><p>Match first/last lines of a block, fuzzy-match the middle content. </p><p>↓4<strong>Diff/Patch Application</strong></p><p>Use diff-match-patch or git cherry-pick algorithms. </p><p>↓5<strong>Full Overwrite</strong></p><p>Nuclear option. Works 100% but expensive and risky for large files.</p><p>↔️</p><h4 id="not-every-agent-wants-a-fallback-ladder">Not every agent wants a fallback ladder</h4><p>ADK-Rust is one counterexample. It prefers provider-native tool contracts and sharp exact-match failures over adding more local heuristics. Oh My Pi is another: instead of endlessly extending fuzzy matching, it changes the address space with hashline anchors.</p><h2 id="key-insights-for-tool-builders">Key Insights for Tool Builders</h2><p>✅</p><h4 id="dont-force-json">Don't Force JSON</h4><p>For file editing, custom formats or XML reduce escaping errors significantly. Cline and Codex both avoid passing code inside JSON strings.</p><p>✅</p><h4 id="lsp-integration">LSP Integration</h4><p>OpenCode checks for syntax errors <em>immediately</em> after every edit using Language Server Protocol. Bad edits get caught and reported back to the AI.</p><p>✅</p><h4 id="user-confirmation">User Confirmation</h4><p>Grok CLI shows diff previews before every write. Users can approve, modify, or skip. This prevents catastrophic mistakes.</p><p>✅</p><h4 id="1-indexed-line-numbers">1-Indexed Line Numbers</h4><p>Codex, OpenCode, and ADK-Rust all use 1-based line numbers when communicating with LLMs. It matches how humans count lines in editors.</p><h2 id="explore-the-playbook">Explore the Playbook</h2><p><a href="https://wuu73.org/aiguide/coding-file-edits-methods/">📚 Editing MethodsDeep dive into each editing approach: whole file, search/replace, unified diff, anchors, and multi-edit.</a><a href="https://wuu73.org/aiguide/coding-file-edits-agents/">🤖 Agent ComparisonsDetailed analysis of Cline, Codex, OpenCode, Aider, Grok CLI, Crush, Dirac, OpenHands, Pi, Pochi, Qwen Code, Goose, and more.</a><a href="https://wuu73.org/aiguide/coding-file-edits-prompts/">💬 Prompts &amp; InstructionsHow agents tell AI models to use their tools. System prompts, tool definitions, and error handling.</a><a href="https://wuu73.org/aiguide/coding-file-edits-adk-rust/">🏗️ ADK-Rust Deep DiveShort focused write-up on provider-native editing, strict exact matches, Anthropic editor wrappers, and OpenAI patch declarations.</a><a href="https://wuu73.org/aiguide/coding-file-edits-reasonix/">🎯 Reasonix Deep DiveByte-exact SEARCH/REPLACE, edit-gate review, repair stages, snapshotting, and strict sandbox enforcement.</a><a href="https://wuu73.org/aiguide/coding-file-edits-codewhale/">🐋 CodeWhale Deep DiveDirect writes, exact-first replace with bounded fuzzy fallback, transactional patch rollback, approval gating, and diagnostic feedback from touched files.</a><a href="https://wuu73.org/aiguide/coding-file-edits-cheetahclaws/">🐆 CheetahClaws Deep DiveMixed-mode editing with direct Read/Write/Edit tools, capability-gated kernel built-ins, notebook mutation, and unified diff feedback.</a><a href="https://wuu73.org/aiguide/coding-file-edits-oh-my-pi/">🧪 Oh My Pi Deep DiveFocused analysis of hashline editing, prompt helpers, nearby rebasing, multi-section preflight, and how it compares to Dirac and Pi Mono.</a><a href="https://wuu73.org/aiguide/coding-file-edits-implementation/">🛠️ Build Your OwnPractical guide with code examples. Implement fallback cascades, matching algorithms, and LSP integration.</a></p><h2 id="quick-comparison-chart">Quick Comparison Chart</h2><p>A visual overview of how different agents prioritize various aspects:</p><h4 id="token-efficiency">Token Efficiency</h4><p>Lower is betterCodex★★★★★Aider★★★★☆Cline★★★☆☆OpenCode★★★☆☆</p><h4 id="fallback-depth">Fallback Depth</h4><p>More is more robustOpenCode9 layersAider5 layersCline4 layersCodex3 layers</p><h4 id="lsp-integration-1">LSP Integration</h4><p>Syntax checkingOpenCodeFullClinePartialCodexMinimalAiderMinimal</p><h4 id="user-approval">User Approval</h4><p>Confirmation flowGrok CLIFull diff previewCodexApproval req.ClineConfigurableAiderAuto-apply</p>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[May 2026: Oh My Pi, Opencode Go, and AI Coding Price Increases]]></title>
          <description><![CDATA[A May 2026 budget AI coding update covering Oh My Pi, Opencode Go, DeepSeek v4, MiniMax M2.7, GLM 5.1, and the latest pricing shifts across coding tools.]]></description>
          <link>https://wuu73.org/aiguide/05152026/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/05152026/</guid>
          <category><![CDATA[Guides]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Fri, 15 May 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<h1 id="may-2026-big-changes-with-ai-coding-on-a-budget-current-workflow">May 2026: Big Changes with AI Coding on a Budget / Current Workflow</h1><p>All of the top companies are starting to raise prices, limit access. I kept going back to Claude Code but with cheaper chinese models instead of real Claude, which works okay, but every time they release a new Claude Code, it starts annoying me with something. Last time, even when I ran it with dangerously skip permissions, it just kept asking me permission for every single thing it did. It has done that a few times, so I said screw it, i'll make my own agentic coding tool!</p><p>But then I found Oh My Pi - its Pi Mono, the agent used in OpenClaw, but its got all the good stuff you typically want. The developer added a lot of stuff I was thinking about in my head that I would have to implement if i tried to make my own (since I couldn't find anything decent!).</p><p>When you install it, it'll pick up information from other coding agents, like api keys, etc, that have already been installed/used, so its very easy to just install and use it.</p><p>Its doing a lot of things right, Its just an agent that works, and has an efficient hash based file edit mode that seems to work great and is more efficient (and works with all models).</p><p>It has been perfect so far, no complaints with Minimax M2.7, Deepseek v4.</p><p>Minimax coding plan or whatever they call it, token plan or w/e, is $10, and its a good model for constant use, its fast, mostly accurate but not always. But since, i've found some more better plans that include it. Opencode Go $5 first month, $10 after - has several models, Qwen3.6 models, Minimax, Deepseek, some other ones. I am about to try it, but seems good - someone complained about a model having only 250k context instead of 1m but mostly people seem happy. You can use it with things without using Opencode at all (Opencode sucks, sorry to say, glitchy mess, i would avoid especially if on a Windows machine)</p><h2 id="%F0%9F%9A%80-github-copilot-wont-be-so-great-with-the-tokens-coming-next-month">🚀 Github Copilot won't be so great with the tokens coming next month</h2><p><strong>Github Copilot, Claude Code, Pi</strong> this is the last month (May) to get the crazy lenient usage out of it before they switch it to something much worse. I have no idea how that will work out... but I will find out since I have copilot.</p><h2 id="claude-code-oh-my-pi">Claude Code + Oh My Pi</h2><p>Some people seem to think that it is only the model intelligence that matters, and the agent harness / coding tool around the models aren't a big deal. <strong>This is sooooo wrong.</strong></p><p>I recommend Oh My Pi at this point since Claude Code is trying hard to kick everyone off or annoy them if they aren't using 100% Claude models.</p><h2 id="deepseek-v4-flash-and-pro">Deepseek v4 Flash and Pro</h2><p>Right now its crazy discounted/cheap <strong>Deepseek v4</strong> and/or <strong>Minimax M2.7</strong> via Opencode Go and/or <strong>GLM 5.1</strong> (GLM seems to have lots of availability issues...), compared to using <strong>Claude Opus 4.5</strong> in some tools like Opencode. Qwen3.5/3.6 are very good as well.</p><p>I am only paying <strong>28-40 cents per million tokens</strong> if I use the official Deepseek Anthropic-compatible API endpoint with Claude Code.</p>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[Alibaba Coding Plan: $3/month for Claude Code with Qwen, Kimi, GLM &amp; More]]></title>
          <description><![CDATA[A quick look at Alibaba&#39;s Coding Plan — $3/month for access to Qwen, Kimi, GLM, and MiniMax models via Claude Code using the Dashscope API.]]></description>
          <link>https://wuu73.org/aiguide/03022026/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/03022026/</guid>
          <category><![CDATA[Guides]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Mon, 02 Mar 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<h1 id="unfortunately-they-jacked-the-price-up-to-50month-so-this-is-old-information-already-unfortunately">Unfortunately they jacked the price up to $50/month :( so this is old information already unfortunately!</h1><p>I read that if you had bought it when it was at $3, it might still let you pay $10 or something, but this is unconfirmed. I got rid of mine and am just using CliProxyAPIPlus to route Openrouter models to Claude Code.</p><p>But... I found this:</p><p><a href="https://claurst.kuber.studio/?ref=wuu73.org">Rust redo of Claude Code, when the src code was leaked</a></p><p>Haven't tried it yet, and it might not have every detail yet from C.C. but it looks good, will try.</p><p>Will post new guide more like the older ones soon.</p><h1 id="old-info-starts-here">old info starts here</h1><h1 id="alibaba-coding-plan-3month-for-claude-code">Alibaba Coding Plan: $3/Month for Claude Code</h1><p>I had to post about this cheap Coding Plan I found — <a href="https://www.alibabacloud.com/help/en/model-studio/coding-plan?ref=wuu73.org"><strong>Alibaba Coding Plan</strong></a>, <strong>$3/month</strong> right now, and you get all these models plus some other ones:</p><h2 id="plan-details">Plan Details</h2>
<!--kg-card-begin: html-->
<table>
<thead>
<tr>
<th>Plan</th>
<th>Intro Price</th>
<th>Requests / 5 hrs</th>
<th>Requests / week</th>
<th>Requests / month</th>
</tr>
</thead>
<tbody>
<tr>
<td><strong>Lite</strong></td>
<td>$3 / first month</td>
<td>1,200</td>
<td>9,000</td>
<td>18,000</td>
</tr>
<tr>
<td><strong>Pro</strong></td>
<td>$15 / first month</td>
<td>6,000</td>
<td>45,000</td>
<td>90,000</td>
</tr>
</tbody>
</table>
<!--kg-card-end: html-->
<h2 id="supported-models">Supported Models</h2><p><strong>Recommended:</strong></p><ul><li><code>qwen3.5-plus</code> (vision)</li><li><code>kimi-k2.5</code> (vision)</li><li><code>glm-5</code></li><li><code>MiniMax-M2.5</code></li></ul><p><strong>More models:</strong></p><ul><li><code>qwen3-max-2026-01-23</code></li><li><code>qwen3-coder-next</code></li><li><code>qwen3-coder-plus</code></li><li><code>glm-4.7</code></li></ul><hr><p>You can use it directly in <strong>Claude Code</strong>, <strong>OpenClaw</strong>, and others.</p><p>I still use <strong>Claude Code</strong> and <strong>GitHub Copilot</strong>. I tried Goose... on Windows 11... errors and more errors! I don't have time for that so I got rid of it. I have noticed a lot of these things are created by Mac people, and Windows/Linux are 2nd class citizens — more of an afterthought.</p><hr><h2 id="claude-code-setup-for-alibaba-dashscope">Claude Code Setup for Alibaba / Dashscope</h2><p><strong>🔧 Click to expand: Bash config for Claude Code with Alibaba models (add to your ~/.bashrc)</strong>Copy and paste this at the end of your `~/.bashrc` file if using bash in Linux, or vibe code something similar for easy model switching.<code># --- ALIBABA / DASHSCOPE CLAUDE CODE CONFIGURATION ---<br><br># PASTE YOUR API KEY HERE:<br>export ALIBABA_API_KEY="sk-sp-xxxxxxx"<br><br># Internal helper function to launch Claude with Alibaba settings<br>_claude_ali_launch() {<br>    local model_name=$1<br>    shift<br>    (<br>        export ANTHROPIC_BASE_URL="https://coding-intl.dashscope.aliyuncs.com/apps/anthropic"<br>        export ANTHROPIC_AUTH_TOKEN="$ALIBABA_API_KEY"<br>        export ANTHROPIC_API_KEY=""  # Clear this to avoid conflicts<br>        export ANTHROPIC_MODEL="$model_name"<br>        export ANTHROPIC_DEFAULT_HAIKU_MODEL="$model_name"<br>        export ANTHROPIC_DEFAULT_OPUS_MODEL="$model_name"<br>        export ANTHROPIC_DEFAULT_SONNET_MODEL="$model_name"<br>        export ANTHROPIC_REASONING_MODEL="$model_name"<br><br>        # Performance/Token settings<br>        # export CLAUDE_CODE_MAX_OUTPUT_TOKENS=30000<br>        export API_TIMEOUT_MS=3000000<br>        export CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC=1<br>        export CLAUDE_CODE_EXPERIMENTAL_AGENT_TEAMS=1<br><br>        claude --dangerously-skip-permissions --model "$model_name" "$@"<br>    )<br>}<br><br># Model Commands<br>alias claude-qwen='_claude_ali_launch qwen3.5-plus'<br>alias claude-kimi='_claude_ali_launch kimi-k2.5'<br>alias claude-glm5='_claude_ali_launch glm-5'<br>alias claude-mm='_claude_ali_launch MiniMax-M2.5'<br><br># --- MENU DISPLAY ---<br>claude-menu() {<br>    echo -e "\n\033[1;36m====================================================\033[0m"<br>    echo -e "\033[1;32m  CLAUDE CODE - ALIBABA (DASHSCOPE) ENDPOINTS       \033[0m"<br>    echo -e "\033[1;36m====================================================\033[0m"<br>    echo -e "  \033[1;33mclaude-qwen\033[0m  - Use Qwen 3.5 Plus"<br>    echo -e "  \033[1;33mclaude-kimi\033[0m  - Use Kimi K2.5"<br>    echo -e "  \033[1;33mclaude-glm5\033[0m  - Use GLM-5"<br>    echo -e "  \033[1;33mclaude-mm\033[0m    - Use MiniMax M2.5"<br>    echo -e "\033[1;36m----------------------------------------------------\033[0m"<br>    echo -e "  Keys are pre-configured with skip-permissions."<br>    echo -e "\033[1;36m====================================================\033[0m\n"<br>}<br><br># Run the menu automatically on startup</code><br>claude-menu<br></p><hr><h2 id="which-model-to-use">Which Model to Use?</h2><p>I have found that running <strong>Claude Code set to kimi-k2.5</strong> will create very elaborate, complex plans — but it is not as fast as the others. <strong>MiniMax M2.5 is quickest</strong>, which is often the best just for that reason.</p><p>Some tactics you can use:</p><ul><li>Run these different model commands, even at the same time, to all create plans simultaneously.</li><li>Then have one of the smarter ones like Kimi go over all the plans and give opinions — explain why any one is better than the others.</li><li><strong>Generate a best-of-N plan. Use MiniMax to implement.</strong></li></ul><p>I haven't used the new <strong>Qwen 3.5</strong> models enough yet to have a strong opinion, but they seem good. Same with <strong>GLM 5</strong>.</p><p>The Coding Plan's <a href="https://www.alibabacloud.com/help/en/model-studio/openclaw-coding-plan?ref=wuu73.org">main page</a> gives instructions for all of the main coding agentic tools, and they include <strong>OpenClaw</strong>. Save some money! Be careful with security if messing with OpenClaw. 🔐</p><p>Gemini 3.1 Pro came out too! It's very very good. The USA AI companies are still the absolute best ones. It still makes sense to use something like my tool, <a href="https://wuu73.org/aicp?ref=wuu73.org">aicodeprep-gui</a>, to plan things out on all the different web chat's (see the older guides from last year... July, Sept etc for how-to's) and just use that tool to go back and forth when bug fixing or in the planning stages of a project.</p><p>I use Gemini 3 Flash a lot for bug fixing, the speed is sometimes better than intelligence since a faster model can fail 3 times before fixing something, while the smarter model can figure it out on first pass or two, but it takes longer...</p>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[Purpose Driven Development]]></title>
          <description><![CDATA[Notes on using short purpose sentences to keep coding agents on track.]]></description>
          <link>https://wuu73.org/aiguide/purpose-driven-development/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/infoblogs/purpose_driven_development/</guid>
          <category><![CDATA[Blog]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Thu, 12 Feb 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<h1 id="purpose-driven-development">Purpose Driven Development</h1><p>Help AI models by giving them the WHY behind your instructions.</p><p>Fancy title but simple idea I have been doing that helps keep coding agents on track:</p><p>After I give some instructions (anywhere in the process of coding with them) if I think something might be hard to understand or I suspect that it might go off track or forget, I add a sentence or two about the purpose of the instructions.</p><p>Like:</p><p>Prompt/instructions:<br>Have the app check to see if powershell 7 is installed before defaulting to the built in powershell, on app start.<br>**The purpose of this is to** help the user save time so they don't have to manually change it to the newer powershell (it seems logical to assume they will probably want to use the latest one, if they installed it).<br></p><h2 id="usage-tips">Usage tips:</h2><ul><li>Keep the purpose concise (one or two lines) and place it immediately after the instruction it clarifies.</li><li>Use the word Purpose or say "The Point of this is..." so agents can understand. Basically, its like if it was a person and you are trying to teach the person something and when they understand the WHY behind something, its easier to learn or do it!</li><li>Consider visually styling or marking purpose lines (highlight, bold, or a prefixed token) if your toolchain supports it — that makes them easier to extract.</li></ul>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[February 2026: Deepseek v3.2 in Claude Code, minor updates about how to stay low cost]]></title>
          <description><![CDATA[February 2026: AI Updates - Claude Code strategies, Deepseek v3.2, cost-saving workflows, and multilingual tools]]></description>
          <link>https://wuu73.org/aiguide/02042026/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/02042026/</guid>
          <category><![CDATA[Guides]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Wed, 04 Feb 2026 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<h1 id="february-2026-ai-coding-on-a-budget">February 2026: AI Coding on a Budget</h1><p>AI is moving so fast it's nearly impossible to keep up with it or have enough time in the day! I am having a blast though.</p><h2 id="%F0%9F%9A%80-current-workflow-lately">🚀 Current Workflow Lately</h2><p><strong>Claude Code</strong> and <strong>Github Copilot</strong> most of the time (Copilot is not as good as Claude Code but they do try to catch up fast and they have the most bang for your buck when it comes to access to the best models: Claude Opus 4.5, GPT 5.2, etc).</p><p><strong>Claude Code is just the best!!</strong></p><p>Some people seem to think that it is only the model intelligence that matters, and the agent harness / coding tool around the models aren't a big deal. <strong>This is sooooo wrong.</strong></p><h2 id="why-claude-code-cheap-models-expensive-models-bad-tools">Why Claude Code + Cheap Models &gt; Expensive Models + Bad Tools</h2><p>I can get such a better, smoother experience using <strong>Claude Code with Deepseek v3.2</strong> and/or <strong>Minimax M2.1</strong> and/or <strong>GLM 4.7</strong>, compared to using <strong>Claude Opus 4.5</strong> in some tools like Opencode.</p><p>I am only paying <strong>28-40 cents per million tokens</strong> if I use the official Deepseek Anthropic-compatible API endpoint with Claude Code.</p><p>I use Claude Code in Windows often in WSL (Linux sort of "inside" Windows, a lot of things just work better in WSL) and have my <code>~/.bashrc</code> file set up with custom commands to run Claude Code with either Deepseek or Minimax. <strong>GLM 4.7 is the cheapest option here... $3/month still I believe!</strong></p><p><strong>🔧 Click to expand: Add custom bash commands to run Claude Code with Deepseek API, GLM 4.7, Minimax M2.1</strong>Add custom bash commands to run Claude Code with **Deepseek API**, **GLM 4.7**, **Minimax M2.1**, or anything that supports `v1/messages` type API. This makes it simple to switch back and forth when you want normal Claude or Deepseek Claude, M2 Claude, etc. ### Tested in Ubuntu WSL2 You can ask AI to modify it for Mac, Windows Powershell, etc.<code>## Claude Code Custom Commands<br>CLAUDE_BIN="$HOME/.local/bin/claude"<br>CLAUDE_SETTINGS="$HOME/.claude/settings.json"<br><br>### Helper function to swap settings temporarily<br>### There are better/other ways to do it, but I know this works<br>### I haven't tested other versions but you can<br><br>_claude_with_config() {<br>    local json_content="$1"<br>    shift<br><br>    # Ensure .claude directory exists<br>    mkdir -p "$HOME/.claude"<br><br>    # Backup original settings<br>    [ -f "$CLAUDE_SETTINGS" ] &amp;&amp; cp "$CLAUDE_SETTINGS" "$CLAUDE_SETTINGS.backup"<br><br>    # Write JSON directly - no variables to expand<br>    echo "$json_content" &gt; "$CLAUDE_SETTINGS"<br><br>    # Run claude<br>    "$CLAUDE_BIN" "$@"<br>    local exit_code=$?<br><br>    # Restore original settings<br>    [ -f "$CLAUDE_SETTINGS.backup" ] &amp;&amp; mv "$CLAUDE_SETTINGS.backup" "$CLAUDE_SETTINGS"<br><br>    return $exit_code<br>}<br><br>## DeepSeek Configuration<br>claude-ds() {<br>    local config='{<br>  "env": {<br>    "ANTHROPIC_AUTH_TOKEN": "sk-examplekey",<br>    "ANTHROPIC_BASE_URL": "https://api.deepseek.com/anthropic",<br>    "ANTHROPIC_API_KEY": "",<br>    "ANTHROPIC_DEFAULT_HAIKU_MODEL": "DeepSeek-V3.2",<br>    "ANTHROPIC_DEFAULT_OPUS_MODEL": "DeepSeek-V3.2",<br>    "ANTHROPIC_DEFAULT_SONNET_MODEL": "DeepSeek-V3.2",<br>    "ANTHROPIC_MODEL": "DeepSeek-V3.2",<br>    "ANTHROPIC_REASONING_MODEL": "DeepSeek-V3.2-Speciale",<br>    "API_TIMEOUT_MS": 600000<br>  },<br>  "includeCoAuthoredBy": false<br>}'<br>    _claude_with_config "$config" "$@"<br>}<br><br>## This version of the Deepseek command adds the skip permissions flag<br>claude-ds2() {<br>    local config='{<br>  "env": {<br>    "ANTHROPIC_AUTH_TOKEN": "sk-examplekey",<br>    "ANTHROPIC_BASE_URL": "https://api.deepseek.com/anthropic",<br>    "ANTHROPIC_API_KEY": "",<br>    "ANTHROPIC_DEFAULT_HAIKU_MODEL": "DeepSeek-V3.2",<br>    "ANTHROPIC_DEFAULT_OPUS_MODEL": "DeepSeek-V3.2",<br>    "ANTHROPIC_DEFAULT_SONNET_MODEL": "DeepSeek-V3.2",<br>    "ANTHROPIC_MODEL": "DeepSeek-V3.2",<br>    "ANTHROPIC_REASONING_MODEL": "DeepSeek-V3.2-Speciale",<br>    "API_TIMEOUT_MS": 600000<br>  },<br>  "includeCoAuthoredBy": false<br>}'<br>    _claude_with_config "$config" --dangerously-skip-permissions "$@"<br>}<br><br>## MiniMax Configuration<br>claude-mm() {<br>    local config='{<br>  "env": {<br>    "ANTHROPIC_AUTH_TOKEN": "eyJexamplekeyFJP5CLrCE1f-sxKKg",<br>    "ANTHROPIC_BASE_URL": "https://api.minimax.io/anthropic",<br>    "ANTHROPIC_API_KEY": "",<br>    "ANTHROPIC_DEFAULT_HAIKU_MODEL": "MiniMax-M2.1",<br>    "ANTHROPIC_DEFAULT_OPUS_MODEL": "MiniMax-M2.1",<br>    "ANTHROPIC_DEFAULT_SONNET_MODEL": "MiniMax-M2.1",<br>    "ANTHROPIC_MODEL": "MiniMax-M2.1",<br>    "API_TIMEOUT_MS": 3000000,<br>    "CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC": "1"<br>  },<br>  "includeCoAuthoredBy": false<br>}'<br>    _claude_with_config "$config" "$@"<br>}<br><br>claude-mm2() {<br>    local config='{<br>  "env": {<br>    "ANTHROPIC_AUTH_TOKEN": "eyJhbGcfakeexamplekeyNpfWyandBFJP5CLr9999CE1f-sxKKg",<br>    "ANTHROPIC_BASE_URL": "https://api.minimax.io/anthropic",<br>    "ANTHROPIC_API_KEY": "",<br>    "ANTHROPIC_DEFAULT_HAIKU_MODEL": "MiniMax-M2.1",<br>    "ANTHROPIC_DEFAULT_OPUS_MODEL": "MiniMax-M2.1",<br>    "ANTHROPIC_DEFAULT_SONNET_MODEL": "MiniMax-M2.1",<br>    "ANTHROPIC_MODEL": "MiniMax-M2.1",<br>    "API_TIMEOUT_MS": 3000000,<br>    "CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC": "1"<br>  },<br>  "includeCoAuthoredBy": false<br>}'<br>    _claude_with_config "$config" --dangerously-skip-permissions "$@"<br>}<br><br>## Claude Menu Command<br>claude-menu() {<br>    echo "Available Claude Custom Commands:"<br>    echo "  claude-ds    - Run Claude with DeepSeek"<br>    echo "  claude-ds2   - Run Claude with DeepSeek (skip permissions)"<br>    echo "  claude-mm    - Run Claude with MiniMax"<br>    echo "  claude-mm2   - Run Claude with MiniMax (skip permissions)"<br>}</code><br></p><hr><h2 id="%F0%9F%8E%AF-why-claude-code-stays-ahead">🎯 Why Claude Code Stays Ahead</h2><p>All of those coding tools are lagging behind Claude Code, and it seems like the main developer on that is just way smarter (and faster) so it's unlikely to change anytime soon. He's the one coming up with all the novel methods to save context or split context, he knows how to delegate tasks with the right context to subagents, etc.</p><p><strong>For example, the Skills feature is genius.</strong> Most people thought it was a gimmick - not me! It's immediately obvious how smart it is with the ability to not shove the entire amount of text at the model. It can choose to load more progressively.</p><p>I can see a lot of other things that could possibly use a similar feature/addition like context pruning... with a dedicated model that does that in the background. Might be a good thing.</p><p>### 💡 Key Insight: Skills Feature The **Skills feature** in Claude Code allows: - Progressive context loading (not dumping everything at once) - Smart selection of relevant information - Reduced token usage - Better model performance This is what separates amateur tools from professional ones.</p><hr><h2 id="%F0%9F%92%B0-use-smartest-models-for-planning-cheap-models-for-doing">💰 Use Smartest Models for Planning, Cheap Models for Doing</h2><p>You don't have to use the pricey <strong>Claude 4.5 Opus and Sonnet</strong> the entire time - I see that the internet is still full of angry people when they can't do much with a <strong>$20 Claude subscription</strong>. Do they upgrade to MAX for $200? People tend to use Opus for everything or think they need at least Sonnet.</p><p>### My Philosophy (Still True After All This Time) **You don't need superintelligent models to use tools and MCP servers.** I think the best coding tool could be one that always uses at least two models: 1. **One for planning** (the smart one) 2. **One for execution** (well, more like one main smart one, and many smaller/less intelligent models that run subagents to do things) One model for tool use (multiple instances of it at the same time usually) that just hand back the information to the smart model when it needs to do smart stuff like fix bugs or plan out a whole app at a high level.</p><h3 id="the-strategy">The Strategy:</h3><ol><li><strong>Planning Stage:</strong> Use Claude Opus 4.5, GPT 5.2, Deepseek v3.2, or Gemini 3 Pro</li><li><strong>Execution Stage:</strong> Use GLM 4.7, Minimax M2.1, or Deepseek v3.2</li><li><strong>Tool Usage:</strong> Cheap models work just fine for file operations, searches, etc.</li></ol><hr><h2 id="%F0%9F%93%9D-spend-longest-time-in-the-planning-stages">📝 Spend Longest Time in the Planning Stages</h2><p><strong>Figure out details so it won't guess</strong></p><p>If you pay for Claude, use it for planning stages, spend much more time here than feels comfortable. Or just use <strong>Deepseek v3.2</strong> here (it is so good, I am not sure why I ignored it for a while!).</p><p>Look up <strong>SPEC.md</strong> methods or different PLAN approaches. I like this plugin called <strong>Superpowers</strong>. Github has some spec command you can download that is probably good although it seemed too time consuming. I am still trying to find the perfect plan tool or workflow to plan really good to where it can one-shot a complex project.</p><p>### 🧪 Experimental Multi-Model Planning Workflow A workflow I have been thinking about but haven't tried yet: 1. Let **Minimax/GLM 4.7** plan with Superpowers plan plugin 2. Run it through **GPT 5.2**, **Claude Opus 4.5**, **Gemini 3 Pro**, and a couple Chinese models 3. Have them look it over and make suggestions 4. Synthesize the best ideas from all models You can use this tool to help send prompts to many models at the same time (I'm working on making that part a separate thing, and making it into an Agent Skill/plugin): **[AI Code Prep GUI](https://wuu73.org/aicp)** Then switch to a faster model (Deepseek is a bit slow!) like **GLM 4.7** - I believe it is still $3/month, for execution.</p><p>There are easy ways around this, even better than what I had discovered before... read on!</p><hr><h2 id="%E2%9A%A1-glm-47">⚡ GLM 4.7</h2><p>It's a good "do-er" model, not as smart as Claude or Deepseek v3.2, but it is <strong>fast and cheap</strong>. Only <strong>$3/month!</strong></p><p>🚀 You've been invited to join the GLM Coding Plan! Enjoy full support for Claude Code, Cline, and 10+ top coding tools — starting at just $3/month. Subscribe now and grab the limited-time deal! <strong>Link:</strong> <a href="https://z.ai/subscribe?ic=LQIZTM2EP4&ref=wuu73.org">https://z.ai/subscribe?ic=LQIZTM2EP4</a></p><hr><h2 id="%F0%9F%8F%86-claude-code-is-always-several-steps-ahead-of-the-rest">🏆 Claude Code Is Always Several Steps Ahead of the Rest</h2><p>I am currently using <strong>Claude Code with Deepseek v3.2</strong> which only costs:</p>
<!--kg-card-begin: html-->
<table>
<thead>
<tr>
<th>Token Type</th>
<th>Cost</th>
</tr>
</thead>
<tbody>
<tr>
<td>1M INPUT TOKENS (CACHE HIT)</td>
<td><strong>$0.028</strong></td>
</tr>
<tr>
<td>1M INPUT TOKENS (CACHE MISS)</td>
<td><strong>$0.28</strong></td>
</tr>
<tr>
<td>1M OUTPUT TOKENS</td>
<td><strong>$0.42</strong></td>
</tr>
</tbody>
</table>
<!--kg-card-end: html-->
<p>I also use it a TON with <a href="https://www.minimaxi.com/?ref=wuu73.org"><strong>Minimax M2</strong></a>.</p><h3 id="my-current-workflow">My Current Workflow:</h3><ol><li>Plan with <strong>aicodeprep-gui</strong> using multiple models (and the new <a href="https://wuu73.org/aicp/flowc/?ref=wuu73.org">Flow Studio</a>)</li><li>Get everything fine-tuned</li><li>Get a good plan and possibly a speckit going (if it's a big enough software project)</li><li>Load <strong>Claude Code</strong> up</li><li>Give it the markdown plan and just let it run</li><li>Give it some MCP servers to help automate testing</li></ol><p><strong>Works so well.</strong></p><hr><h2 id="%F0%9F%A4%96-gemini-cli">🤖 Gemini CLI</h2><p>It is getting better! I have not used it that much but I am going to see what it can do vs Claude Code. <strong>It is still free!</strong></p><p>I use Gemini 3 Flash and Pro in Copilot for debugging... its great, better than Claude usually for that.</p><p>You can get the <strong>3-month $300 Google Cloud credit</strong>, which allows you to use <strong>Nano Banana Pro</strong> (REALLY amazing) in the CLI/API.</p><hr><h2 id="%F0%9F%8C%8C-google-anti-gravity-ide">🌌 Google Anti Gravity IDE</h2><p>It's good! It's <strong>free to use</strong> (rate limited but usable) and you can use <strong>Opus 4.5 + Gemini 3 Pro free</strong>.</p><p>It is a <strong>VS Code Clone</strong>.</p><p>Link: <a href="https://antigravity.google/?ref=wuu73.org">https://antigravity.google</a></p><hr><h2 id="%F0%9F%8C%8D-updates-to-aicodeprep-gui-now-multilingual">🌍 Updates to aicodeprep-gui: Now Multilingual!</h2><p>I made it multilingual as some people asked about it - originally I made it for myself which is why I didn't bother with any of that before. <strong>Now you can switch language.</strong></p><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aicp/scrs/hindi.png" class="kg-image" alt="Hindi language support" loading="lazy"></figure><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aicp/scrs/arabic.png" class="kg-image" alt="Arabic language support" loading="lazy"></figure><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aicp/scrs/chinese.png" class="kg-image" alt="Chinese language support" loading="lazy"></figure><figure class="kg-card kg-image-card"><img src="https://wuu73.org/aicp/scrs/korean.png" class="kg-image" alt="Korean language support" loading="lazy"></figure><p>I noticed a lot of visitors to this site from <strong>China, Korea, India</strong>... so this should help!</p><hr><h2 id="%F0%9F%92%A1-quick-tips-best-practices">💡 Quick Tips &amp; Best Practices</h2><p>### Model Selection Matrix **When to use EXPENSIVE models (Claude Opus 4.5, GPT 5.2):** - Initial project architecture decisions - Complex debugging that cheaper models can't solve - Novel algorithm design - Critical security reviews **When to use CHEAP models (GLM 4.7, Deepseek v3.2):** - File operations - Repetitive code generation - Following established patterns - MCP server tool usage - Test writing **When to use FREE models (Gemini 3 Pro on Antigravity):** - Free doesn't mean bad, its good for everything - Experiments - Learning new technologies - Debugging</p><hr><h2 id="%F0%9F%8E%93-lessons-learned-so-far-in-2026">🎓 Lessons Learned So Far in 2026</h2><ol><li><strong>The tool matters MORE than the model</strong> - Claude Code with Deepseek &gt; Bad tool with Claude Opus</li><li><strong>Planning is worth the investment</strong> - Spend 3x longer planning vs coding</li><li><strong>Multi-model workflows work</strong> - Different models excel at different things</li><li><strong>Context curation is king</strong> - Don't dump everything, be strategic</li><li><strong>Cheap models are getting REALLY good</strong> - GLM 4.7 and Deepseek v3.2 rival expensive models for many tasks</li></ol><hr><h2 id="%F0%9F%93%8A-cost-comparison-february-2026">📊 Cost Comparison: February 2026</h2><p>Here's what I'm actually spending monthly for unlimited AI coding:</p>
<!--kg-card-begin: html-->
<table>
<thead>
<tr>
<th>Service</th>
<th>Cost</th>
<th>What I Use It For</th>
</tr>
</thead>
<tbody>
<tr>
<td><strong>Github Copilot Pro</strong></td>
<td>$10/month</td>
<td>Claude 4.5, GPT 5.2 access</td>
</tr>
<tr>
<td><strong>GLM 4.7 Coding Plan</strong></td>
<td>$3/month</td>
<td>Main execution model</td>
</tr>
<tr>
<td><strong>Deepseek v3.2 API</strong></td>
<td>~$5-10/month</td>
<td>Planning &amp; complex logic</td>
</tr>
<tr>
<td><strong>Google Cloud Credits</strong></td>
<td>FREE</td>
<td>Gemini experiments</td>
</tr>
<tr>
<td><strong>Antigravity IDE</strong></td>
<td>FREE</td>
<td>Testing &amp; prototyping</td>
</tr>
<tr>
<td><strong>Total</strong></td>
<td><strong>~$20-25/month</strong></td>
<td>Unlimited AI coding!</td>
</tr>
</tbody>
</table>
<!--kg-card-end: html-->
<p>Compare this to:</p><ul><li>Claude Pro MAX: <strong>$200/month</strong> (and you still hit limits!)</li><li>Cursor Pro: <strong>$40/month</strong></li><li>Other premium IDEs: <strong>$30-100/month</strong></li></ul><hr><h2 id="%F0%9F%9A%80-whats-next">🚀 What's Next?</h2><p>I'm experimenting with:</p><ul><li>Multi-model review workflows</li><li>Automated spec generation</li><li>Context pruning with dedicated models</li><li>MCP server automation strategies</li></ul><p>Check back next month for updates!</p><p>## Final Thoughts **AI coding in 2026 is not about having the most expensive subscription.** It's about understanding which model to use for which task, how to structure your prompts, and leveraging the right tools. Claude Code + cheap models + good planning = infinite coding at 5% the cost.</p><hr><p><em>Last updated: February 4, 2026</em></p>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[November 2025: Cost / Lean AI Coding Updates]]></title>
          <description><![CDATA[November 2025 updates: free &amp; cheap coding models (Minimax M2, Ling 1T, Ring 1T, Longcat, Kimi K2), cost hacks and emerging tooling.]]></description>
          <link>https://wuu73.org/aiguide/12092025/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/12092025/</guid>
          <category><![CDATA[Guides]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Tue, 09 Dec 2025 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<p>It's hard to keep up with all the new things and all the progress, but here are some cost / price / lean AI coding updates.</p><h2 id="current-workflow">Current Workflow</h2><p>I am currently using Claude Code a TON with <a href="https://www.minimaxi.com/?ref=wuu73.org"><strong>Minimax M2</strong></a>. Usually I plan with aicodeprep-gui using multiple models (and the new <a href="https://wuu73.org/aicp/flowc/?ref=wuu73.org">Flow Studio</a>), get everything fine tuned, get a good plan and possibly a speckit going (if it's a big enough software project) and then load Claude Code up, give it the markdown plan and just let it run. Give it some MCP servers to help automate testing.. works so well.</p><h2 id="model-updates">Model Updates</h2><h3 id="minimax-m2"><a href="https://www.minimaxi.com/?ref=wuu73.org">Minimax M2</a></h3><p>Was <strong>free for a period of time in November</strong>, works with Claude Code (as well as GLM 4.5/4.6) very very good!! Works in Cline, etc. Better/smarter vs GLM models. Also, I got <strong>$50 API credits free</strong> just by signing up for the developer deal - their API is very cheap, I hope the model quality stays great! They have a coding plan similar to the GLM / Z.ai coding plan (but $10/mo vs $3).</p><p>I just noticed Cline is giving away free usage of Minimax M2 right now.</p><p><strong>🔥 Claude Code + Chinese Models = Less Hallucinations!</strong></p><p>Both <strong>Minimax M2</strong> and <strong>GLM 4.5/4.6</strong> work <em>exceptionally well</em> with Claude Code. Here's the surprising thing: they hallucinate LESS than real Claude 4.5 Opus! Opus tends to stray off into its own ideas a lot, going on tangents or "improving" things you didn't ask it to improve. These Chinese models stay focused on the task, follow instructions better, and just get the job done. Especially when you let Claude Code run without user input - M2 and GLM just keep chugging along reliably.</p><p><strong>Get 10% off Z.AI's coding plan:</strong> <a href="https://z.ai/subscribe?ic=RV42B8COHB&ref=wuu73.org">https://z.ai/subscribe?ic=RV42B8COHB</a></p><h3 id="worth-looking-into">Worth Looking Into</h3><p>Have not used these enough to have an opinion but looks possibly good:</p><ul><li><strong>Ling 1T</strong></li><li><strong>Ring 1T</strong></li></ul><h3 id="longcatchat"><a href="https://longcat.chat/?ref=wuu73.org">Longcat.chat</a></h3><p>Has some new models, with <strong>0.5-5.0 mil tokens free daily</strong> on the website. Also looks like it might work in Claude Code.</p><h3 id="kimi-k2-thinking"><a href="https://www.kimi.com/?ref=wuu73.org">Kimi K2 Thinking</a></h3><p>I am sure this is excellent, have not tried, but heard good things. <strong>Non-thinking Kimi K2 is already excellent</strong> for writing.</p><h3 id="grok-4">Grok 4</h3><p>I tried using it for something (one of the 5 LLMs I will sometimes use for brainstorming or fixing a bug - 5 models, each from a different AI company, then a 6th model to analyze all 5 results to generate a best of all of them result usually provides a superior response to just one model <a href="https://wuu73.org/aicp/flowc?ref=wuu73.org">Learn more here</a> - and it hallucinated pretty bad. It was actually funny, is this what people see when they claim AI is always hallucinating? I don't usually get much of that. I don't trust people claiming these models are always the best - some people (not mentioning names) will lie a ton, and i'm supposed to just believe them when they claim their models are always the best. Everytime I try to use it, its meh, not terrible but not great. Its free though - but I don't trust that it will always tell me the truth and is not programmed to say something different.</p><hr><h2 id="interesting-finds">Interesting Finds</h2><p>These are currently (either totally or partial) free, and interesting enough to use:</p><h3 id="googles-antigravity-ide-code-editor"><a href="https://antigravity.google/?ref=wuu73.org">Google's Antigravity IDE / Code Editor</a></h3><p>https://antigravity.google/?ref=wuu73.org</p><p>VS Code clone with Gemini 3 Pro, Claude Sonnet 4.5... FREE! Its got a multi-agent system that can do many things at the same time. I am using it today and so far it is okay but NOT as good as Claude Code or Github Copilot in VS Code (which is actually very good now compared to how it used to be!). I also got rate limited, but its free and i'm sure they will improve it with regular updates. Can't beat free usage of both those models!</p><p>When it started rate limiting Gemini 3, I could switch to Claude 4.5 and get plenty of use, so it appears the limits to free are for each model rather than combined.</p><h3 id="httpswwwkimicom"><a href="https://www.kimi.com/?ref=wuu73.org">https://www.kimi.com</a></h3><p>Besides normal AI web chat, there are things on here for:</p><ul><li>Generating slides automatically</li><li>A Manus-like thing that has AI agents (powered by Kimi K2) inside a virtual machine that can code and create apps or do things, research etc for you</li></ul><h3 id="httpsagentminimaxio"><a href="https://agent.minimax.io/?ref=wuu73.org">https://agent.minimax.io</a></h3><p>Looks interesting, Minimax M2 model is fantastic so I'm sure this is good, and appears to be <strong>free</strong> as of the time I am typing this.</p><h3 id="cc-switch">cc-switch</h3><p><a href="https://github.com/farion1231/cc-switch?ref=wuu73.org">https://github.com/farion1231/cc-switch</a></p><p>Looks useful for switching config/MCP servers on/off between Claude Code and Codex. I downloaded the repo and scanned it through several LLMs for a security check, looks good. It's very small size due to using Rust and Tauri.</p><hr><h2 id="github-spec-kit-from-microsoft">GitHub Spec-Kit from Microsoft</h2><p><a href="https://github.com/github/spec-kit?ref=wuu73.org">https://github.com/github/spec-kit</a></p><p>This is <strong>free</strong> and will keep your AI coding agent on task. It seems good but I have not experimented enough to know how good it works. It seems a bit overkill for smaller projects but for super important, larger software, it seems like it would be very helpful.</p><hr><h2 id="async-web-based-tools">Async / Web Based Tools</h2><h3 id="google-jules">Google Jules</h3><p><a href="https://jules.google/?ref=wuu73.org">Google Jules</a></p><p>Currently free and seems reliable. You give it permission to work with your repos on GitHub, it will make changes, work inside a VM where it is able to test in a loop.</p><h3 id="github-spark-copilot-agent"><a href="https://github.com/features/spark?ref=wuu73.org">GitHub Spark</a> &amp; Copilot Agent</h3><p>https://github.com/features/spark?ref=wuu73.org</p><p>Both similar and work great! The copilot agent only uses one premium request for the whole task.</p><p>**About GitHub Copilot:** It used to suck, but now it is good. I pay the $10/month and then use the unlimited GPT 4.1 and limited Claude 4.5. It has come a long way, and the **copilot agent mode is fantastic**. So I currently pay $10 for Copilot (and the Claude credits go a long way.. you can just keep on going in the same chat as long as you don't keep opening fresh ones, it barely uses any premium credits even after many hours).</p><h3 id="my-current-subscriptions">My Current Subscriptions</h3><ul><li><strong>$10 for Github Copilot Pro</strong></li><li><strong>$3 for Chutes.ai</strong></li><li><strong>$3 for </strong><a href="https://z.ai/subscribe?ic=RV42B8COHB&ref=wuu73.org"><strong>Z.AI GLM Coding Plan</strong></a> - Get 10% off with that link! (but.. I think Minimax M2 is also excellent! Very cheap API, both work amazingly in Claude Code)</li></ul><p>I have used both of these more than once, to run async jobs (you send it off, and it gets to work without you interacting much or needing to, then you come to look at it later).</p><p><strong>🤝 A Note About "Deals" on This Guide</strong></p><p>I am not going to 'sell out' ever and start going down the enshittification road to crappyness. I will never put crap on this guide. So when you see a 'deal' like 10% off Z.AI's coding plan, it's because it's genuinely great! It's pocket change credit they would give me that I don't really care about.</p><p>2025-2026 is the era of no shame where people will try to sell you the crappiest crap, snake oil, etc. because they get a dollar. <strong>NOT ME! Promise!</strong></p><p>These Chinese models (GLM, Minimax M2) are GREAT in Claude Code, especially when letting it run without user input. They hallucinate less than real Claude Opus 4.5 which tends to stray off into its own ideas constantly.</p><hr><h2 id="random-study-i-found">Random study I found</h2><p><a href="https://arxiv.org/abs/2407.01489?ref=wuu73.org">https://arxiv.org/abs/2407.01489</a></p><p>@@@ make this section stand out with a yellow background, with dark black text. Try to make it a rounded box with soft edges or a fade to yellow then fade back to normal. The yellow shouldn't be full brightness.. more like 30-40% grey still.</p><h2 id="agentless-demystifying-llm-based-software-engineering-agents">Agentless: Demystifying LLM-based Software Engineering Agents</h2><p>Recent advancements in large language models (LLMs) have significantly advanced the automation of software development tasks, including code synthesis, program repair, and test generation. More recently, researchers and industry practitioners have developed various autonomous LLM agents to perform end-to-end software development tasks. These agents are equipped with the ability to use tools, run commands, observe feedback from the environment, and plan for future actions. However, the complexity of these agent-based approaches, together with the limited abilities of current LLMs, raises the following question: Do we really have to employ complex autonomous software agents? To attempt to answer this question, we build Agentless -- an agentless approach to automatically solve software development problems. Compared to the verbose and complex setup of agent-based approaches, Agentless employs a simplistic three-phase process of localization, repair, and patch validation, without letting the LLM decide future actions or operate with complex tools. Our results on the popular SWE-bench Lite benchmark show that surprisingly the simplistic Agentless is able to achieve both the highest performance (32.00%, 96 correct fixes) and low cost ($0.70) compared with all existing open-source software agents! Furthermore, we manually classified the problems in SWE-bench Lite and found problems with exact ground truth patch or insufficient/misleading issue descriptions. As such, we construct SWE-bench Lite-S by excluding such problematic issues to perform more rigorous evaluation and comparison. Our work highlights the current overlooked potential of a simple, interpretable technique in autonomous software development. We hope Agentless will help reset the baseline, starting point, and horizon for autonomous software agents, and inspire future work along this crucial direction. @@@@</p><h2 id="testing-new-feature-flow-studio">Testing New Feature: Flow Studio</h2><p>Testing new feature in aicodeprep-gui/pro - Has a <strong>Flow Studio</strong> now where you can have it auto-send your context/code/project, send your problems or tasks etc (prompt) to as many LLMs as you want at the same time. You can design complex logic. See more by going here: <a href="https://wuu73.org/aicp/flowc/flow_studio_help.html?ref=wuu73.org">https://wuu73.org/aicp/flowc/flow_studio_help.html</a></p><p>I am currently splitting that so it is a <strong>separate app and MCP server</strong>, so any app or model can directly call for help (if you use 5 different companies AI models... GPT 5.. Claude 4.5.. Gemini.. Kimi.. etc and maybe 3 instances each of those... compare all outputs or have AI do it.. and design a best of N or better than all solution - that's the idea and it works well on harder problems or big planning stages.</p>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[How I Code with AI on a Budget/Free – September 2025 Edition]]></title>
          <description><![CDATA[A deep-dive into my 2025 AI coding stack using GLM 4.5, Gemini 2.5 Pro, Cline and AI Code Prep.]]></description>
          <link>https://wuu73.org/aiguide/09012025/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/09012025/</guid>
          <category><![CDATA[Guides]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Mon, 01 Sep 2025 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<h1 id="how-i-code-with-ai-on-a-budgetfree">How I Code with AI on a budget/free</h1><h2 id="my-browser-setup-the-free-ai-buffet">My Browser Setup: The Free AI Buffet</h2><p>First things first, I have a browser open loaded with tabs pointing to the free tiers of powerful AI models. Why stick to one when you can get multiple perspectives for free? My typical lineup includes:</p><ul><li>At least 2-3 tabs of <a href="https://chat.z.ai/?ref=wuu73.org">z.ai's GLM 4.5</a> – free on web, and seems as good or better than Claude 4! no joke. <a href="https://z.ai/subscribe?ic=RV42B8COHB&ref=wuu73.org"><strong>Get 10% off Z.AI's coding plan</strong></a> if you want API access for Claude Code (works amazingly well, less hallucinations than real Claude Opus!).</li><li>At least one tab but more like three of Google <a href="https://aistudio.google.com/?ref=wuu73.org">Gemini AI Studio</a> (Gemini 2.5 Pro/Flash are often free and unlimited here).<br>Also, try <a href="https://gemini.google.com/app?ref=wuu73.org">Google Gemini 2.5 Pro</a> (different than AI Studio, has better image generation and deep research; I always have a couple tabs of this along with a couple tabs of AI Studio).</li><li>Couple tabs of <a href="https://poe.com/?ref=wuu73.org">Poe.com</a> usually set to Claude 4 or o4-mini for its free daily credits on premium models.</li><li>Several tabs of <a href="https://openrouter.ai/?ref=wuu73.org">OpenRouter</a>, set to several models, some free models, some not.</li><li>At least one tab for <a href="https://chatgpt.com/?ref=wuu73.org">ChatGPT</a> (the free version is still useful).</li><li>At least one tab for <a href="https://www.perplexity.ai/?ref=wuu73.org">Perplexity AI</a>, especially good for research-heavy questions.</li><li>At least one tab for <a href="https://chat.deepseek.com/?ref=wuu73.org">Deepseek</a> (v3 and r1 are free on their web interface, though watch the context limit).</li><li>One tab for <a href="https://grok.com/?ref=wuu73.org">Grok.com</a>. Good, free and seemingly unlimited for general use and deep research/image editing. I mainly use the deep research feature, similar to perplexity.</li><li><a href="https://phind.com/?ref=wuu73.org">Phind</a> is another free one, it tries to show you flowcharts/diagram visuals.</li><li><a href="https://lmarena.ai/?ref=wuu73.org">lmarena.ai</a> offers free access to Claude Opus 4 and Sonnet 4 and others. Free Opus 4 is so good.</li></ul><p><a href="https://claude.ai/new?ref=wuu73.org">Claude.ai</a> - Free but sometimes so limited it's annoying to use, so I use other sites/ways to access Claude like Cody extension, Copilot, etc.</p><p>### ⚠️ Important Disclaimer about Grok Grok offers free compute and uncensored image generation, which can be useful when other models' safety systems interfere with legitimate tasks. However, reports indicate that Grok has been instructed to lie about about some things. While the misinformation appears to be mostly on X, if you keep in mind to restrict usage to coding or be cautious knowing it might be programmed with questionable motives, it can occasionally be useful. It was never that great though but its sometimes okay.</p><h2 id="a-smarter-cheaper-workflow-focused-context">A smarter, cheaper workflow: Focused Context</h2><p>When you use AI in web chat's (the chat interfaces like AI Studio, ChatGPT, Openrouter, instead of thru an IDE or agent framework) are almost always better at solving problems, and coming up with solutions compared to the agents like Cline, Trae, Copilot.. Not always, but usually.</p><p>When you use things like Cursor, Cline, Roo Code for everything, they are sending tons of text to the AI about how to use their tools, how to use or activate MCP server stuff, edit files, etc it "dumbs it down" too much. It gets confused. People end up paying for the most expensive best models to do everything and even that isn't enough to get over the dumbing down effect from the AIs getting tons of unneeded information unrelated to your problem.</p><p>So when that happens, I use my tool to generate the right context to solve my problem. Then I paste it into one of the many AI web chat's (sometimes more than one, since they sometimes give different answers) and just ask it questions or ask it to code review, to try to figure out why x is happening when y is happening...etc then when it figures out a solution.. i have it write a prompt for Cline or another agent type thing to do the actual file edits. GPT 4.1 can handle this just fine and I have unlimited. No reason to be wasting Claude credits to edit files. No reason to be sending Claude a bunch of crap it doesn't need making it dumb. I can use Claude to plan out anything or fix really hard problems, cheap, using Openrouter web chat then just paste it back in Cline and let it run.</p><p>After doing this for a while, you really get a feel for which models excel at which types of tasks.</p><p>**How AI Code Prep Helps (Example Prompt Structure):** _Can you help me figure out why my program does x instead of y?_ Then, [AI Code Prep GUI](https://wuu73.org/aicp) (for Windows, Mac, Linux, and web) steps in. It recursively scans your project folder (subfolders, sub-subfolders, you name it) and grabs the code, formatting it nicely for AI like this: The context block generated by AI Code Prep looks like this:Can you help me figure out why my program does x instead of y?<br><br>fileName.js:<br>&lt;code&gt;<br>... the contents of the file..<br>&lt;/code&gt;<br><br>nextFile.py:<br>&lt;code&gt;<br>import example<br>...etc<br>&lt;/code&gt;<br><br>Can you help me figure out why my program does x instead of y?<br>It writes it twice if you have that option enabled, which helps get the AI to focus better on your question/prompt. You can choose to have it on top, bottom, or both. OpenAI claims this helps, I haven't really tested to see if that's true but it seems logical. On Windows, you just right-click somewhere inside your project folder (or on the folder itself) and select "AI Code Prep GUI" from the context menu (look at the screenshots on the site). A GUI window pops up, usually with the right code files pre-selected. It smartly tries to skip things you likely don't need, like `node_modules`, `.git`, etc. If its guess isn't perfect, you can easily check or uncheck files. This is super useful when your project is huge and blows past an AI's context limit. You can manually curate exactly what the AI needs to see. The problem with many coding agents like [Cline](https://cline.bot/), Github Copilot, Cursor, Windsurf, etc., is that they often send either WAY too much context or WAY too little. This is why they can seem dumb or ineffective sometimes. Sometimes, you just gotta do things yourself, use a tool like mine to select the files yourself, but it helps auto-select the code files while skipping the stuff you probably don't need (but still have the option to add what you want with the checkboxes) and then dump that curated context into several AIs (especially the free web ones!). There are other context-generating tools, but many are command-line only, or need a public GitHub repo link. What if your code is private? What if you want to keep it local? What if you prefer checkboxes on a GUI? For something like this a GUI makes sense. It has some other things I haven't seen anywhere else like the preset buttons, dual placement of problem/prompt, per-project saves. ![AI Code Prep Explanation](https://wuu73.org/aicp/scrs/ar.png)</p><hr><h2 id="model-strategy-picking-the-right-brain-for-the-job">Model Strategy: Picking the Right Brain for the Job</h2><h2 id="model-strategy-picking-the-right-brain-for-the-job-1">Model Strategy: Picking the Right Brain for the Job</h2><p>Since many great models are free to use via web interfaces (like Gemini in <a href="https://aistudio.google.com/?ref=wuu73.org">AI Studio</a>, <a href="https://grok.com/?ref=wuu73.org">Grok</a>, <a href="https://chat.deepseek.com/?ref=wuu73.org">Deepseek</a>), I prioritize these. <a href="https://poe.com/?ref=wuu73.org">Poe.com</a> also gives free daily credits for top models like Claude and the new o4 series.</p><p><strong>Gemini 2.5 Pro</strong> (via <a href="https://aistudio.google.com/?ref=wuu73.org">AI Studio</a>) is great for debugging, great for planning, and also finding it the best at lots of things now. For really thorny issues, I might try the new <strong>o4-mini</strong> (available via <a href="https://openrouter.ai/?ref=wuu73.org">OpenRouter</a> or <a href="https://poe.com/?ref=wuu73.org">Poe</a>). It surprisingly fixed a persistent bug for me right away, though I'm still figuring out its best use cases. It's notably cheaper via API than the previous top dogs like Claude 3.5/3.7/4.</p><p>I usually try <strong>Claude 3.7 or 4</strong> at some point, via <a href="https://poe.com/?ref=wuu73.org">Poe</a> or API (<a href="https://openrouter.ai/?ref=wuu73.org">OpenRouter</a> makes this easy), or github Copilot chat (you can get some free usage from that if you don't pay) but it's pricier for frequent use. Think of Claude 3.7 and 4 as Claude on Adderall – brilliant, sometimes verbose, maybe a bit 'psychotic' like Hunter S. Thompson. Lots of great output, but you might need a calmer model like <strong>Claude 3.5</strong> to refine it or do the actual coding.</p><p>For really hard problems, try using OpenAI's o3 or GLM 4.5, Qwen3 Coder 480b. You can get lots of free daily tokens if you set your account to allow sharing of your data to help train models. Go to the Open AI Playground page, click the settings icon in the upper right, then click Data Controls on the left sidebar, then Sharing on the displayed page, there you can change the "Share inputs and outputs with OpenAI" setting to Enabled, which will give you:</p><ul><li>Up to 250 thousand tokens per day across gpt-5, gpt-4.1, gpt-4o, o1 and o3</li><li>Up to 2.5 million tokens per day across gpt-4.1-mini, gpt-4.1-nano, gpt-4o-mini, o1-mini, o3-mini, o4-mini, and codex-mini-latest</li></ul><p>This is really awesome, o3 and GPT 4.5 seem super genius! Sometimes in the OpenAI Playground, I have it set up to use o3 and o4-mini side-by-side, to compare them. This helps me get a feel for which ones are best for which types of problems.</p><p><strong>Claude 4 and 3.7</strong> is always a good option to try and fix hard problems quickly, it is just harder to access it for cheap or free. But it is often the best out of them all. When you really need to fix something fast, use it. <a href="https://poe.com/?ref=wuu73.org">Poe</a> has free tokens for all models, daily. <a href="https://openrouter.ai/?ref=wuu73.org">OpenRouter</a> has all models paid and/or free. Claude 3.7 is Claude on caffeine – brilliant, sometimes verbose, maybe a bit 'psychotic' like Hunter S. Thompson. Lots of great output, but you might need a calmer model like <strong>Claude 3.5 / 4</strong> to refine it or do the actual coding.</p><h2 id="the-hybrid-approach-premium-planning-budget-execution">The Hybrid Approach: Premium Planning + Budget Execution</h2><p>After extensive testing with various models, I've developed a hybrid strategy that maximizes both quality and cost-effectiveness. The key insight is that different models excel at different parts of the development process.</p><h2 id="my-smart-juice-theory-of-model-intelligence">My "smart juice" theory of model intelligence</h2><p>How Models become stupid under certain circumstances: AI models are usually smarter the less text you send to them. Think of each model as having a fixed amount of "intelligence" or "smart juice" available for every question or problem you ask. When you send a simple, focused prompt, nearly 100% of that intelligence is available to solve your problem. But the more complex your input—long agentic instructions about how to use tools, lots of context unrelated to your specific problem, or multiple pages of code—the more of that "smart juice" gets used up just processing the unrelated stuff like how it can use tools in your IDE, leaving less intelligence energy available for your actual problem.</p><p>This is why tools like Cursor, Cline, and other agentic systems can sometimes seem less effective: if they send five giant pages of instructions and context before even getting to your real question, the model's available intelligence for your specific problem drops. The more "stuff" you send, the more diluted the model's focus becomes. For best results, keep your prompts as concise and targeted as possible—curate the context so the model can use its full intelligence on what matters most.</p><p>When you have a hard problem or bug, you will usually save time by using AI Code Prep to dump it into a web chat (as discussed on page 1 of this guide). It cuts out all the extra instructions and stuff that get sent in agentic IDEs/apps. I noticed that this works better even if you give the AI ALL of the files from your project. The agentic instructions/stuff/bloat that is unrelated to your actual problem is the content that seems to make the AI dumber/run out of juice.</p><p><strong>My Workflow is something like this when starting a new project:</strong></p><ol><li><strong>Plan &amp; Brainstorm:</strong> Use the smarter/free web models (Gemini 2.5, o4-mini, Claude 3.7, 4, o3, etc) to figure out the approach, plan the steps, identify libraries, etc.</li><li><strong>Generate Agent Prompt:</strong> Ask one of these smart models: "Write a detailed-enough prompt for <a href="https://cline.bot/?ref=wuu73.org">Cline</a>, my AI coding agent, to complete the following tasks: [describe tasks]". Sometimes, I'll copy this generated prompt and paste it into <em>another</em> free AI good at rewriting (like <a href="https://chatgpt.com/?ref=wuu73.org">ChatGPT</a>) to refine it further.</li><li><strong>Execute with Cline:</strong> Paste the step-by-step task list into <a href="https://cline.bot/?ref=wuu73.org">Cline</a>, configured to use a stable and efficient model like <strong>GPT 4.1 or Claude 3.5</strong> (or Claude 4 if it is doing really complicated things). The 4.1's have been trained to follow instructions well.</li><li><strong>Fallback:</strong> If GPT 4.1 struggles, switch <a href="https://cline.bot/?ref=wuu73.org">Cline</a> to use <strong>Claude 3.5</strong> via API. It seems to be the next best for reliable execution. <a href="https://chat.deepseek.com/?ref=wuu73.org">Deepseek v3 or R1</a> is really great at following instructions as well.</li></ol><p>Essentially: Use expensive/smart models (and the excellent free Gemini 2.5 Pro) to strategize and plan. Validate the plan by pasting it into 2-3 other free models (<a href="https://chat.deepseek.com/?ref=wuu73.org">Deepseek R1</a>, Claude on <a href="https://poe.com/?ref=wuu73.org">Poe</a> if context allows) and ask "Is this good? Can you improve it or find flaws?". Then, use a stable workhorse like GPT 4.1 or Claude 3.5 within <a href="https://cline.bot/?ref=wuu73.org">Cline</a> to do the heavy lifting (coding).</p><p><strong>o4-mini</strong> seems particularly adept at untangling complex code logic or figuring out high-level implementation strategies (like choosing frameworks or libraries). I'll often throw my initial idea at Gemini 2.5, o4-mini, GPT 4.1, <a href="https://chatgpt.com/?ref=wuu73.org">ChatGPT</a>, maybe o3-mini (try <a href="https://duckduckgo.com/chat?ref=wuu73.org">duck.ai</a> - often free), and <a href="https://phind.com/?ref=wuu73.org">Phind</a> to get a range of ideas. If the free/cheap options don't crack it, I'll escalate to pricier models via API.</p><h2 id="alternative-agents-setups">Alternative Agents &amp; Setups</h2><p><a href="https://trae.ai/?ref=wuu73.org">Trae.ai</a> (from Bytedance, makers of TikTok) is a free VS Code compatible IDE with free AI usage, including Claude 4, Claude 3.7, Claude 3.5, and GPT 4.1. Their agents aren't as good as Cline (nothing is as good, to be honest!) but it's free and gives access to the best models. Sometimes, I find its built-in agent isn't as robust as <a href="https://cline.bot/?ref=wuu73.org">Cline</a>. However, since <a href="https://trae.ai/?ref=wuu73.org">Trae</a> seems to be a VS Code clone, you can likely install the <a href="https://cline.bot/?ref=wuu73.org">Cline</a> extension within it! However... it is too overloaded to get any free usage from it, its too slow. I'll still mention it though.. but meh.</p><p>So, you could have two setups:</p><ul><li>VS Code + <a href="https://cline.bot/?ref=wuu73.org">Cline</a> extension + <a href="https://github.com/features/copilot?ref=wuu73.org">Copilot</a> extension (get the $10/mo subscription for cheap API access via Cline, though the free tier might offer some basic use).</li><li><a href="https://trae.ai/?ref=wuu73.org">Trae.ai</a> + <a href="https://cline.bot/?ref=wuu73.org">Cline</a> extension (potentially leveraging Trae's free model access if Cline can use it, or using your own API keys).</li></ul><p>Try both! Sometimes the native <a href="https://github.com/features/copilot?ref=wuu73.org">Copilot</a> agent solves things <a href="https://cline.bot/?ref=wuu73.org">Cline</a> struggles with, and vice-versa. I suspect <a href="https://cline.bot/?ref=wuu73.org">Cline</a> sometimes sends overly large prompts which might hinder performance on certain tasks compared to the more integrated Copilot agent.</p><h3 id="roo-code-clines-clone">Roo Code: Cline's Clone</h3><p><strong>Roo Code</strong> is a clone of Cline, very similar but with some different features that are worth trying out. Sometimes Cline might work better for your workflow, and sometimes Roo Code will. It's a good idea to try both and see which fits your needs for a given project or coding style.</p><p><a href="https://cline.bot/?ref=wuu73.org">Cline</a> for VS Code is free, but remember you pay for the API calls unless you're leveraging the <a href="https://github.com/features/copilot?ref=wuu73.org">Copilot</a> subscription trick. Using the VS Code LM API setting in <a href="https://cline.bot/?ref=wuu73.org">Cline</a> with a $10/month Copilot sub is currently the most cost-effective way to get near-unlimited access to powerful models within the agent.</p><h3 id="new-cli-tools-claude-code-qwen-code-gemini-cli">New CLI Tools: Claude Code, Qwen Code, Gemini CLI</h3><p>There's a lot of buzz about new CLI tools for coding, especially <strong>Claude Code</strong>, <strong>Qwen Code</strong>, and <strong>Gemini CLI</strong>. People rave about Claude Code's capabilities, though I haven't tried it myself yet. When I do, I plan to set it up to use <strong>GLM 4.5</strong> instead (there's a guide for this on the z.ai website).</p><p>Claude Code supports subagents—these are agents that only do one task and don't use extra tools. This setup can mimic the streamlined workflow described in this guide, focusing the model's intelligence on a single job. Subagents are a clever way to avoid the "bloat" of agentic instructions and keep things efficient.</p><p>If you want to experiment, check out the guides and community tips for configuring these tools. The ecosystem is evolving quickly, and each tool has its own strengths for different workflows.</p><hr><h2 id="tldr-model-updates-%E2%80%93-september-2025">TL;DR &amp; Model Updates – September 2025</h2><h2 id="tldr-quickstart-guide">TL;DR: Quickstart Guide</h2><ul><li><strong>Models &amp; Roles:</strong></li><li>Planning &amp; Brainstorming: GLM 4.5, Kimi K2, the newest Qwen3 Coder and 2507's, Gemini 2.5 Pro (AI Studio), o4-mini (OpenRouter), Claude 3.7 or 4 (Poe), if you have OpenAI Playground configured for the 250k free daily tokens, I recommend using those up with o3 and GPT 5.</li><li>Problem Solving &amp; Debugging: GPT-5 (free tokens in Playground), GLM-4.5 (it seems to be a genius, about Claude 4 level) Claude 4 (free daily on Poe)</li><li>Actual Coding: GPT-4.1 via Cline; fallback to Claude 3.5.. or the new ones: Qwen3 Coder, Instruct, 2507, GLM 4.5, Kimi K2.</li><li><strong>Key Tools:</strong></li><li>VS Code</li><li>AI Code Prep GUI – locally scan &amp; curate only the files you need, saves so much time</li><li>Cline (VS Code agent) for step-by-step code execution</li><li>Free web chats for multi-perspective advice: Poe.com, ChatGPT, Grok, Deepseek, Perplexity, OpenAI Playground, AI Studio w/Gemini 2.5 Pro, Openrouter, duck.ai</li><li><strong>Quick Workflow:</strong> 1. Run AI Code Prep GUI to bundle your (if already existing) project's relevant files. 2. Paste that context into your favorite web chat models for planning &amp; debugging. 3. Ask one model to "Write me a detailed Cline prompt for these tasks," then refine it (e.g. in ChatGPT). 4. Copy/paste into Cline set to GPT-4.1 to generate or fix code; if it stalls, switch to Claude 3.5.</li><li><strong>Cost-Saving Hacks:</strong></li><li>Enable "share data" in OpenAI Playground for 250k free GPT-4.5, o3, (both genius expensive models) &amp; 2.5 MILLION free tokens/day for o4-mini, o3-mini!!</li><li>$10/mo GitHub Copilot subscription gives you rate-limited access to Claude models via Cline</li><li>Pay-as-you-go on OpenRouter for o4-mini, Claude 3.7, and other new models</li></ul><p>## Note about some free CLI agent tools Currently as of end of August 2025, Qwen Code (using the GREAT Qwen3 Coder 480b model!) is totally free, 1000-2000 requests a day (I have not run into limits with heavy use). It is an agentic tool that runs in the terminal, but you can also 'borrow' the API key to use in anything. Kilo Code currently has it set up automatically, just choose the Qwen provider. Gemini CLI is also free with high limits but I have heard about getting limited fast. Claude Code is currently said to be the best CLI tool, but it really eats tokens if you aren't getting them for free. Claude Code Router, a free repo on github, allows you to use whatever API you want for Claude Code. OpenCode.ai is another one, it allows me to use my unlimited GPT 4.1 from the $10/month co-pilot subcription. I still have not found that any of these are as good as doing it my way where I use [AI Code Prep GUI](https://wuu73.org/aicp) to plan with models on their native web chat's, then implement with GPT 4.1. Maybe I should try to create my own agentic tool, that separates things so that difficult problems are sent to the smartest AIs without the tools, MCPs, emulating this way of doing it, then send to smaller model for coding.</p><h2 id="latest-model-updates-aug-2025">Latest Model Updates (Aug 2025)</h2><h3 id="%F0%9F%92%B0-budget-conscious-getting-max-value">💰 Budget-Conscious: Getting Max Value</h3><p><strong>GPT 4.5</strong> <em>(DISCONTINUED)</em></p><ul><li>This model was discontinued.</li></ul><p><strong>o3</strong></p><ul><li>Possibly equal to Claude 4 in abilities, really great at fixing hard problems, genius level. Throw your entire codebase in here with <a href="https://wuu73.org/aicp?ref=wuu73.org">AI Code Prep GUI</a></li><li><strong>Free Tokens:</strong> 250k daily when you enable data sharing in <a href="https://platform.openai.com/settings/organization/data-controls/sharing?ref=wuu73.org">Data Controls/Sharing settings</a></li></ul><p><strong>o4-mini</strong></p><ul><li>Not as smart as o3, but very very good, like o3's younger brother</li><li><strong>Free Tokens:</strong> 2.5 million daily when you enable data sharing in <a href="https://platform.openai.com/settings/organization/data-controls/sharing?ref=wuu73.org">Data Controls/Sharing settings</a></li></ul><p><strong>Gemini 2.5 Pro</strong></p><ul><li>Free to use in <a href="https://aistudio.google.com/?ref=wuu73.org">AI Studio</a>. Also very good at fixing hard problems</li><li><strong>Best For:</strong> Complex debugging and architectural planning</li></ul><p><strong>Deepseek R1 0528</strong></p><ul><li>Super smart model with enhanced reasoning capabilities</li><li><strong>Availability:</strong> Free on Deepseek's web interface</li></ul><h3 id="%F0%9F%9A%80-premium-fix-problems-now">🚀 Premium: Fix Problems NOW</h3><p><strong>Claude 4 Sonnet</strong></p><ul><li>The mega genius, can fix most problems in one shot if you provide enough context (<a href="https://wuu73.org/aicp?ref=wuu73.org">AICodePrep is what I use</a>). These also seem to be the best writers, best at everything really.. the secret sauce</li><li><strong>Use Case:</strong> When you absolutely need it fixed right the first time</li></ul><p><strong>Claude 4 Opus</strong></p><ul><li>$75 per million tokens, haven't tried it yet myself but I hear its a super mega genius like Sonnet but even better, magic sauce</li><li><strong>Performance:</strong> Rumored to be the ultimate problem solver</li></ul><h2 id="solid-worker-models">Solid Worker Models</h2><p>These models listen to instructions really well and do as they are told:</p><p><strong>GPT 4.1</strong></p><ul><li>Use the above smarter models for abstract high level design or to fix problems, then 4.1 to make changes. You can cut and paste output from anywhere else right into Cline with 4.1 to do the coding.</li></ul><p><strong>Claude Sonnet 3.5</strong></p><ul><li>Good solid model for coding and editing, slightly slower vs 4.1 but very reliable.</li></ul><p><strong>Deepseek v3</strong></p><ul><li>Pretty good and very cheap for a model that can do edits/code/agent work.</li></ul><p><strong>OpenRouter Free Models</strong></p><ul><li>Drag the price filter to $0 on <a href="https://openrouter.ai/?ref=wuu73.org">OpenRouter</a> to see free models available for testing. Worth experimenting with new ones as they become available.</li></ul><h2 id="free-claude-4-lmarenaai-and-more">Free Claude 4: lmarena.ai, and More</h2><p><strong>Claude Opus 4 and lmarena.ai</strong></p><ul><li><a href="https://lmarena.ai/?ref=wuu73.org">lmarena.ai</a> offers free access to Claude Opus 4 and Sonnet 4 and others.</li><li>Any free usage of anything by Anthropic is worth saving, remembering, and using. When all else fails, or when you need to get stuff done, perfectly, and now, choose Claude 4 Sonnet or Opus.</li></ul><h3 id="latest-model-updates-2025">Latest Model Updates (2025)</h3><p><strong>Claude Sonnet 4 &amp; Opus 4</strong></p><ul><li><strong>Status:</strong> Just released and proving to be the best models for everything</li><li><strong>Performance:</strong> Fixed some bugs today that were hard for most other models to handle</li><li><strong>Cost:</strong> More expensive, but accessible through GitHub Copilot $10/month (rate limited)</li><li><strong>GitHub Copilot Value:</strong> The cheap $10/mo GitHub Copilot subscription seems to be giving me a very generous amount of Claude 4 uses through the VS Code LM API. I keep using it as my main model in Cline/Roo Code, and it has not run out yet. I either use GPT 4.1 or Claude 4, been waiting for it to limit me but it hasn't yet (they say that GPT 4.1 is unlimited). I always hate subscriptions for anything, but this particular one is insanely good value.</li><li><strong>Strategy:</strong> Save Claude 4 for tough problems, use GPT 4.1 for regular coding</li></ul><p><strong>GPT 4.5</strong></p><ul><li><strong>Performance:</strong> Excellent for bug fixing and complex problem solving</li><li><strong>Token Limits:</strong> 250k tokens daily if you allow data for model training</li><li><strong>Cost Hack:</strong> Free under the 250k token limit when you enable data sharing in settings</li><li><strong>Use Case:</strong> Great for using AI Code Prep tool to analyze entire codebases</li></ul><p>## NEW!! Bad ass new Chinese models + GPT 5 **GLM 4.5** - Very much like Claude 4 Opus or Sonnet; follows agentic rules and uses tools near perfectly. Fixes really hard bugs and handles complicated tasks with lots of context. **Qwen3 Coder 480B** - Another super great one; a favorite for being powerful and cheap. **Qwen3 Instruct &amp; Thinking 2507** - Similar to Qwen3 Coder—strong, dependable, and cost-effective. **Kimi K2 (Moonshot)** - Feels Anthropic-inspired or trained on Claude-like synthetic data. Really good; used a lot. **GPT-5** - **edit** -- OpenAI has vastly improved GPT-5 since I wrote this and it is actually pretty decent at tools now....but the first week was bad which is when i wrote this: Not very good at custom tool usage (e.g., your own tools/MCP/Cline). Better to plan and fix like this guide suggests using GPT 5 and other insanely great models like GLM 4.5, then have it write a prompt for a simpler agent model to do actual editing and tool usage. GPT 4.1 still wins on value, and the new Chinese models handle custom tools/Cline easily. I have not played around with this one yet to know if its good but usually all models are going to be great at certain things. I am honestly more excited about the Chinese models, because of cost and the experiences so far have proven them to be reliably great pretty much all the time.</p><h3 id="current-coding-workflows-that-i-use-2025">Current Coding Workflows that I use (2025)</h3><p><strong>For New Projects:</strong></p><ol><li><strong>Planning Phase:</strong> Type your idea(s) and ask several AI's in web chat's to brainstorm with you and help you figure out a real plan, into notepad or a blank file. - this is just so you can paste into multiple web chat's</li><li><strong>Multi-Model Consultation:</strong> Paste into two or three of these models web chat's for different "doctor's opinions": - Gemini 2.5 Pro (free) - GLM 4.5 (free and unlimited!) - Qwen3-Coder or one of the -2507 models - o4-mini on OpenAI Playground to use the free 2.5mil daily tokens - Claude 4 on Poe.com (free daily credits)</li><li><strong>Refinement:</strong> Go back and forth to fine-tune details</li><li><strong>Task Generation:</strong> Have model write a PRD (Product Requirements Document) markdown file or similar, like "APP_REQUIREMENTS.md" or website requirements etc with a list of all the things it must have. Then have it write step-by-step task list with subtasks, for an AI coding agent to implement. Save this to a new project folder - and you might want to run 'git init' if you use Git.</li><li><strong>Execution:</strong> Copy/paste into Cline (or just tell it "your task is in the project requirements .md file") set to GPT 4.1 for 'act' mode (or Qwen 3 Coder, GLM 4.5 Flash.. also free currently and excellent). If you prompt the web models to break a big project into small enough sub-tasks, then you can get away with using a cheaper implementation/agent model. "Cheap" right now for me is GPT 4.1 using the Github Copilot API (I don't even use the actual copilot much I just like the unlimited 4.1 - also every month you get credits to use all the top expensive models like Opus/Sonnet 4 which are great for top level planning, you can also use copilot's own web chat interface)</li></ol><p><strong>For Problem Solving, fixing hard bugs:</strong></p><ol><li>Load AI Code Prep GUI tool (type 'aicp' in terminal) and type the problem into the prompt box, look at which files are already selected to make sure that it looks good for context (when in doubt, add more code files, usually it helps)</li><li>Set the options to 'Add prompt to top' and ''..bottom' is usually how i leave that setting.. it seems like it would be wasteful but it seems to help focus the model on the problem better than only having the prompt on top or bottom</li><li>If I am confident that it will be able to fix my problem I will click one of the preset buttons for Cline/Roo Code or one for an agent prompt which just pastes some extra text to the end of the current text in the prompt box, telling the AI to write a solution for an agent to implement the changes (experimenting with this will be the easiest way to understand what it is for - it just saves you from typing that)</li><li>Click the "GENERATE CONTEXT!" button, and if its a hard problem.. i'll paste it into Gemini 2.5 Pro, GLM 4.5, Kimi K2, Qwen3 Coder or 2507 Thinking.. all on their native web chat website (Qwen has an annoying web chat where it pastes it into a file so you have to tell it to read the file)</li><li>Sit back and watch several solutions stream in, bathroom break. Using several AI's from several different companies is a great way to get some diversity in the outputs and it is better than for example.. just using 5 OpenAI models or just using Anthropic's models. It reminds me of how people will have healthier kids if they mate with people from different cultures/different parts of the world. It just produces a faster / better solution. You can even cut and paste the outputs of all of these into Gemini 2.5 Pro (its got such a large context window to handle it) and have AI again compare them all and produce a final "best of all" solution.</li><li>Problem solving like this will often "one shot" a perfect solution, instead of letting agents run around your file system trying to slowly figure out solutions with all the tools and MCPs and other stuff it is often just faster to do it this way!</li></ol><h3 id="task-list-test-driven-development-coming-soon">Task List &amp; Test Driven Development (Coming Soon)</h3><p><strong>Test Driven Development &amp; Task Lists:</strong></p><p>Coming to this guide soon (these other topics)Have the AIs create a detailed task list for Cline, Roo Code, Trae agent, to execute. You can also instruct Cline or Roo Code to use a markdown file to keep track of everything it does, checking things off as it completes them. This will make it easier to track and ensure nothing gets missed.</p><p>For now, you can experiment by having a model generate a checklist in markdown, and then ask Cline or Roo Code to update the file as tasks are completed.</p><h3 id="money-saving-hacks">Money-Saving Hacks</h3><ul><li><strong>GLM 4.5, Kimi K2, o3:</strong> 250k tokens free daily by enabling data sharing for model training</li><li><strong>Cheaper Models:</strong> 2.5 million tokens on o4-mini (excellent model), 4.1-mini/nano, GPT 5 Mini</li><li><strong>GitHub Copilot:</strong> $10/month gives access to new Claude models (rate limited) and GPT 4.1 UNLIMITED for agent work</li><li>Qwen3 has new models Coder and 2507's</li><li><strong>Poe.com:</strong> Free daily credits for every type of model</li><li><strong>Web Interfaces:</strong> Use free web chat interfaces for planning and consultation</li></ul><h3 id="coming-soon-live-reddit-data-insights">Coming Soon: Live Reddit Data &amp; Insights</h3><p><strong>Live Reddit Data Scraping &amp; Daily Insights:</strong></p><p>A new feature is coming soon: live scraping of Reddit data and daily updated info about how people are using AI models. This will include detailed usage breakdowns, data visualizations, and new insights into real-world coding workflows and trends.</p><p>If you buy the Pro version of [aicodeprep-gui](https://tombrothers.gumroad.com/l/zthvs), you will get updated content about more free stuff, free api's, updated model information, any new cost hacks I discover but updated more frequently</p><hr><h2 id="ai-coding-tools-alternatives-%E2%80%93-september-2025">AI Coding Tools &amp; Alternatives – September 2025</h2><h2 id="ai-guide-4-ai-coding-tools-alternatives">AI Guide 4: AI Coding Tools &amp; Alternatives</h2><h3 id="%F0%9F%8E%AF-september-2025-updates-free-cheap-coding-models">🎯 September 2025 Updates: Free &amp; Cheap Coding Models</h3><p><strong>GLM 4.5 - The Best Coding Model Right Now</strong></p><ul><li><strong>Price:</strong> Only $3-6/month</li><li><strong>Limits:</strong> 120 prompts every 5 hours (users report rarely hitting limits)</li><li><strong>Access:</strong> <a href="https://z.ai/subscribe?utm_source=zai">z.ai subscription</a></li></ul><p>#### 🔧 Compatible Tools &amp; Integrations GLM 4.5 works seamlessly with these coding assistants using a simple API key: - **Claude Code** - Special Anthropic compatible endpoint available - **Cline** - Direct API integration - **Roo Code** - Full compatibility - **Kilo Code** - API key setup - **OpenCode** - Non-subscription API access (recommended)</p><h3 id="%E2%9C%85-claude-code-works-great-with-glm-45">✅ Claude Code Works Great with GLM 4.5!</h3><p><strong>Recommended Setup:</strong> Claude Code works excellently with GLM 4.5 from z.ai using their Anthropic-compatible API endpoint. You can use either:</p><ul><li><strong>$3-6/month subscription</strong> - Great value with generous limits</li><li><strong>Pay-as-you-go API</strong> - Only pay for what you use</li></ul><p><strong>Note:</strong> I tried using Claude Code Router to allow any OpenAI-compatible API endpoint, but it barely works and I don't recommend it. Just use the official z.ai GLM 4.5 API directly - it's much more reliable!</p><h3 id="%F0%9F%94%A5-recommended-opencodeai">🔥 Recommended: OpenCode.ai</h3><p><strong>Why OpenCode.ai is Excellent</strong></p><ul><li><strong>Multiple Models:</strong> Works great with GLM-4.5 and many other top models</li><li><strong>Agent Setup:</strong> Like Claude Code, you can set up agents that get full fresh context for subtasks</li><li><strong>Superior Performance:</strong> Much better output/code than Cline, Roo Code, or Kilo Code</li><li><strong>Fast Iteration:</strong> Can quickly iterate and test code changes</li><li><strong>Reliable:</strong> Stable and consistent performance</li></ul><p><em>Note: Claude Code is very good with GLM 4.5, but OpenCode offers better stability. OpenCode may have more options than other alternatives.</em></p><h2 id="%F0%9F%92%B0-chutesaiaffordable-api-provider">💰 Chutes.ai - Affordable API Provider</h2><p><strong>Pricing &amp; Features</strong></p><ul><li><strong>$3/month:</strong> 300 requests per day of any model + some free models</li><li><strong>$10/month:</strong> 1000-2000 requests per day (resets daily!)</li><li><strong>Top Models Available:</strong> GLM 4.5, Qwen3 Coder 480B, Kimi K2 (Sept updated), and more</li></ul><p><em>Note: I typically avoid subscriptions, but this one offers exceptional value for the model access.</em></p><p>**WSL Setup Tip:** Running Claude Code in WSL (Ubuntu inside Windows) tends to be more stable than native Windows versions, while still allowing interaction with Windows files.#### 🚀 Current Top Models Available Most of the best models right now are accessible through Chutes.ai: GLM 4.5, the new Qwen3 Coder 480B, Kimi K2 September update, and many more cutting-edge models at fraction of the cost of direct API access.</p><h2 id="%F0%9F%92%A1-free-vs-really-cheap-the-smart-choice">💡 Free vs. Really Cheap: The Smart Choice</h2><p><strong>Yes, You CAN Code 100% for Free, BUT...</strong></p><p>It's much better to code for <strong>REALLY CHEAP</strong>! $10 or less a month is good enough. Just a little bit of money spent in the right places from this guide buys extremely good reliability.</p><p><strong>The Reality of "Free" Resources</strong></p><p>With the few totally free resources out there, you will often be stuck waiting for API calls. Well... that's not totally right because at the moment you can use <strong>Qwen Code</strong> with their best coding model pretty much unlimited and it works great (not as good as GLM 4.5 though). I believe <strong>Gemini CLI</strong> will also let you have generous API calls.</p><p><strong>My Current Setup (Total: ~$13/month)</strong></p><ul><li><strong>GitHub Copilot:</strong> Unlimited GPT-4.1 + limited GPT-5 &amp; Claude 4</li><li><strong>Chutes.ai:</strong> $3/month for access to cutting-edge models</li></ul><p><strong>Bottom Line:</strong> It's really basically a free-for-all right now - no need to be complaining about Claude's $200/month crap! Smart spending on the right tools gives you professional-grade AI coding assistance for pocket change.</p>]]></content:encoded>
        </item>
        <item>
          <title><![CDATA[My AI Code Prep &amp; Cline Workflow for Budget Coding/Debugging]]></title>
          <description><![CDATA[A personal guide on using AI Code Prep GUI, Cline, and various free AI web interfaces like Gemini, Grok, Deepseek, and Poe for cost-effective coding and debugging.]]></description>
          <link>https://wuu73.org/aiguide/07012025/</link>
          <guid isPermaLink="false">https://wuu73.org/aiguide/07012025/</guid>
          <category><![CDATA[Guides]]></category>
          <dc:creator><![CDATA[WUU73]]></dc:creator>
          <pubDate>Tue, 01 Jul 2025 00:00:00 +0000</pubDate>
          <content:encoded><![CDATA[<h1 id="how-i-code-with-ai-on-a-budgetfree">How I Code with AI on a budget/free</h1><h2 id="my-browser-setup-the-free-ai-buffet">My Browser Setup: The Free AI Buffet</h2><p>First things first, I have a browser open loaded with tabs pointing to the free tiers of powerful AI models. Why stick to one when you can get multiple perspectives for free? My typical lineup includes:</p><ul><li>At least one tab of <a href="https://platform.openai.com/playground/prompts?models=o3&ref=wuu73.org">OpenAI Playground</a>. If you set your account's data settings to allow OpenAI to use your data for model training, you get free tokens to use on GPT-4.5, o3, and other models.</li><li>At least one tab but more like three of Google <a href="https://aistudio.google.com/?ref=wuu73.org">Gemini AI Studio</a> (Gemini 2.5 Pro/Flash are often free and unlimited here).<br>Also, try <a href="https://gemini.google.com/app?ref=wuu73.org">Google Gemini 2.5 Pro</a> (different than AI Studio, has better image generation and deep research; I always have a couple tabs of this along with a couple tabs of AI Studio).</li><li>Couple tabs of <a href="https://poe.com/?ref=wuu73.org">Poe.com</a> usually set to Claude 4 or o4-mini for its free daily credits on premium models.</li><li>Several tabs of <a href="https://openrouter.ai/?ref=wuu73.org">OpenRouter</a>, set to several models, some free models, some not.</li><li>At least one tab for <a href="https://chatgpt.com/?ref=wuu73.org">ChatGPT</a> (the free version is still useful).</li><li>At least one tab for <a href="https://www.perplexity.ai/?ref=wuu73.org">Perplexity AI</a>, especially good for research-heavy questions.</li><li>At least one tab for <a href="https://chat.deepseek.com/?ref=wuu73.org">Deepseek</a> (v3 and r1 are free on their web interface, though watch the context limit).</li><li>One tab for <a href="https://grok.com/?ref=wuu73.org">Grok.com</a>. Good, free and seemingly unlimited for general use and deep research/image editing. I mainly use the deep research feature, similar to perplexity.</li><li><a href="https://phind.com/?ref=wuu73.org">Phind</a> is another free one, it tries to show you flowcharts/diagram visuals.</li><li><a href="https://lmarena.ai/?ref=wuu73.org">lmarena.ai</a> offers free access to Claude Opus 4 and Sonnet 4 and others. Free Opus 4 is so good.</li></ul><p>[Claude.ai](https://claude.ai/new) - Free but sometimes so limited it's annoying to use, so I use other sites/ways to access Claude like Cody extension, Copilot, etc.### ⚠️ Important Disclaimer about Grok Grok offers free compute and uncensored image generation, which can be useful when other models' safety systems interfere with legitimate tasks. However, reports indicate that Grok has been instructed to lie about about some things. While the misinformation appears to be mostly on X, if you keep in mind to restrict usage to coding or be cautious knowing it might be programmed with questionable motives, it can occasionally be useful. It was never that great though but its sometimes okay.</p><h2 id="a-smarter-cheaper-workflow-focused-context">A smarter, cheaper workflow: Focused Context</h2><p>When you use AI in web chat's (the chat interfaces like AI Studio, ChatGPT, Openrouter, instead of thru an IDE or agent framework) are almost always better at solving problems, and coming up with solutions compared to the agents like Cline, Trae, Copilot.. Not always, but usually.</p><p>When you use things like Cursor, Cline, Roo Code for everything, they are sending tons of text to the AI about how to use their tools, how to use or activate MCP server stuff, edit files, etc it "dumbs it down" too much. It gets confused. People end up paying for the most expensive best models to do everything and even that isn't enough to get over the dumbing down effect from the AIs getting tons of unneeded information unrelated to your problem.</p><p>So when that happens, I use my tool to generate the right context to solve my problem. Then I paste it into one of the many AI web chat's (sometimes more than one, since they sometimes give different answers) and just ask it questions or ask it to code review, to try to figure out why x is happening when y is happening...etc then when it figures out a solution.. i have it write a prompt for Cline or another agent type thing to do the actual file edits. GPT 4.1 can handle this just fine and I have unlimited. No reason to be wasting Claude credits to edit files. No reason to be sending Claude a bunch of crap it doesn't need making it dumb. I can use Claude to plan out anything or fix really hard problems, cheap, using Openrouter web chat then just paste it back in Cline and let it run.</p><p>After doing this for a while, you really get a feel for which models excel at which types of tasks.</p><p>**How AI Code Prep Helps (Example Prompt Structure):** _Can you help me figure out why my program does x instead of y?_ Then, [AI Code Prep GUI](https://wuu73.org/aicp) (for Windows, Mac, Linux, and web) steps in. It recursively scans your project folder (subfolders, sub-subfolders, you name it) and grabs the code, formatting it nicely for AI like this: The context block generated by AI Code Prep looks like this:Can you help me figure out why my program does x instead of y?<br><br>fileName.js:<br>&lt;code&gt;<br>... the contents of the file..<br>&lt;/code&gt;<br><br>nextFile.py:<br>&lt;code&gt;<br>import example<br>...etc<br>&lt;/code&gt;<br><br>Can you help me figure out why my program does x instead of y?<br>It writes it twice if you have that option enabled, which helps get the AI to focus better on your question/prompt. You can choose to have it on top, bottom, or both. OpenAI claims this helps, I haven't really tested to see if that's true but it seems logical. On Windows, you just right-click somewhere inside your project folder (or on the folder itself) and select "AI Code Prep GUI" from the context menu (look at the screenshots on the site). A GUI window pops up, usually with the right code files pre-selected. It smartly tries to skip things you likely don't need, like `node_modules`, `.git`, etc. If its guess isn't perfect, you can easily check or uncheck files. This is super useful when your project is huge and blows past an AI's context limit. You can manually curate exactly what the AI needs to see. The problem with many coding agents like [Cline](https://cline.bot/), Github Copilot, Cursor, Windsurf, etc., is that they often send either WAY too much context or WAY too little. This is why they can seem dumb or ineffective sometimes. Sometimes, you just gotta do things yourself, use a tool like mine to select the files yourself, but it helps auto-select the code files while skipping the stuff you probably don't need (but still have the option to add what you want with the checkboxes) and then dump that curated context into several AIs (especially the free web ones!). There are other context-generating tools, but many are command-line only, or need a public GitHub repo link. What if your code is private? What if you want to keep it local? What if you prefer checkboxes on a GUI? For something like this a GUI makes sense. It has some other things I haven't seen anywhere else like the preset buttons, dual placement of problem/prompt, per-project saves.</p><p><strong>Note:</strong> This page isn't updated with all the latest AI Code Prep GUI features, check <a href="https://wuu73.org/aicp?ref=wuu73.org">wuu73.org/aicp</a> for latest major upgrade</p><h2 id="model-strategy-picking-the-right-brain-for-the-job">Model Strategy: Picking the Right Brain for the Job</h2><p>Since many great models are free to use via web interfaces (like Gemini in <a href="https://aistudio.google.com/?ref=wuu73.org">AI Studio</a>, <a href="https://grok.com/?ref=wuu73.org">Grok</a>, <a href="https://chat.deepseek.com/?ref=wuu73.org">Deepseek</a>), I prioritize these. <a href="https://poe.com/?ref=wuu73.org">Poe.com</a> also gives free daily credits for top models like Claude and the new o4 series.</p><p><strong>Gemini 2.5 Pro/Preview</strong> (via <a href="https://aistudio.google.com/?ref=wuu73.org">AI Studio</a>) is great for debugging, great for planning, and also finding it the best at lots of things now. For really thorny issues, I might try the new <strong>o4-mini</strong> (available via <a href="https://openrouter.ai/?ref=wuu73.org">OpenRouter</a> or <a href="https://poe.com/?ref=wuu73.org">Poe</a>). It surprisingly fixed a persistent bug for me right away, though I'm still figuring out its best use cases. It's notably cheaper via API than the previous top dogs like Claude 3.5/3.7/4.</p><p>I usually try <strong>Claude 3.7 or 4</strong> at some point, via <a href="https://poe.com/?ref=wuu73.org">Poe</a> or API (<a href="https://openrouter.ai/?ref=wuu73.org">OpenRouter</a> makes this easy), or github Copilot chat (you can get some free usage from that if you don't pay) but it's pricier for frequent use. Think of Claude 3.7 and 4 as Claude on Adderall – brilliant, sometimes verbose, maybe a bit 'psychotic' like Hunter S. Thompson. Lots of great output, but you might need a calmer model like <strong>Claude 3.5</strong> to refine it or do the actual coding.</p><p><strong>My Workflow is something like this when starting a new project:</strong></p><ol><li><strong>Plan &amp; Brainstorm:</strong> Use the smarter/free web models (Gemini 2.5, o4-mini, Claude 3.7, 4, o3, etc) to figure out the approach, plan the steps, identify libraries, etc.</li><li><strong>Generate Agent Prompt:</strong> Ask one of these smart models: "Write a detailed-enough prompt for <a href="https://cline.bot/?ref=wuu73.org">Cline</a>, my AI coding agent, to complete the following tasks: [describe tasks]". Sometimes, I'll copy this generated prompt and paste it into <em>another</em> free AI good at rewriting (like <a href="https://chatgpt.com/?ref=wuu73.org">ChatGPT</a>) to refine it further.</li><li><strong>Execute with Cline:</strong> Paste the step-by-step task list into <a href="https://cline.bot/?ref=wuu73.org">Cline</a>, configured to use a stable and efficient model like <strong>GPT 4.1 or Claude 3.5</strong> (or Claude 4 if it is doing really complicated things). The 4.1's have been trained to follow instructions well.</li><li><strong>Fallback:</strong> If GPT 4.1 struggles, switch <a href="https://cline.bot/?ref=wuu73.org">Cline</a> to use <strong>Claude 3.5</strong> via API. It seems to be the next best for reliable execution. <a href="https://chat.deepseek.com/?ref=wuu73.org">Deepseek v3 or R1</a> is really great at following instructions as well.</li></ol><p>Essentially: Use expensive/smart models (and the excellent free Gemini 2.5 Pro) to strategize and plan. Validate the plan by pasting it into 2-3 other free models (<a href="https://chat.deepseek.com/?ref=wuu73.org">Deepseek R1</a>, Claude on <a href="https://poe.com/?ref=wuu73.org">Poe</a> if context allows) and ask "Is this good? Can you improve it or find flaws?". Then, use a stable workhorse like GPT 4.1 or Claude 3.5 within <a href="https://cline.bot/?ref=wuu73.org">Cline</a> to do the heavy lifting (coding).</p><p><strong>o4-mini</strong> seems particularly adept at untangling complex code logic or figuring out high-level implementation strategies (like choosing frameworks or libraries). I'll often throw my initial idea at Gemini 2.5, o4-mini, GPT 4.1, <a href="https://chatgpt.com/?ref=wuu73.org">ChatGPT</a>, maybe o3-mini (try <a href="https://duckduckgo.com/chat?ref=wuu73.org">duck.ai</a> - often free), and <a href="https://phind.com/?ref=wuu73.org">Phind</a> to get a range of ideas. If the free/cheap options don't crack it, I'll escalate to pricier models via API.</p><h2 id="alternative-agents-setups">Alternative Agents &amp; Setups</h2><p><a href="https://trae.ai/?ref=wuu73.org">Trae.ai</a> (from Bytedance, makers of TikTok) is a free VS Code compatible IDE with free AI usage, including Claude 4, Claude 3.7, Claude 3.5, and GPT 4.1. Their agents aren't as good as Cline (nothing is as good, to be honest!) but it's free and gives access to the best models. Sometimes, I find its built-in agent isn't as robust as <a href="https://cline.bot/?ref=wuu73.org">Cline</a>. However, since <a href="https://trae.ai/?ref=wuu73.org">Trae</a> seems to be a VS Code clone, you can likely install the <a href="https://cline.bot/?ref=wuu73.org">Cline</a> extension within it! However... it is too overloaded to get any free usage from it, its too slow. I'll still mention it though.. but meh.</p><p>So, you could have two setups:</p><ul><li>VS Code + <a href="https://cline.bot/?ref=wuu73.org">Cline</a> extension + <a href="https://github.com/features/copilot?ref=wuu73.org">Copilot</a> extension (get the $10/mo subscription for cheap API access via Cline, though the free tier might offer some basic use).</li><li><a href="https://trae.ai/?ref=wuu73.org">Trae.ai</a> + <a href="https://cline.bot/?ref=wuu73.org">Cline</a> extension (potentially leveraging Trae's free model access if Cline can use it, or using your own API keys).</li></ul><p>Try both! Sometimes the native <a href="https://github.com/features/copilot?ref=wuu73.org">Copilot</a> agent solves things <a href="https://cline.bot/?ref=wuu73.org">Cline</a> struggles with, and vice-versa. I suspect <a href="https://cline.bot/?ref=wuu73.org">Cline</a> sometimes sends overly large prompts which might hinder performance on certain tasks compared to the more integrated Copilot agent.</p><h3 id="roo-code-clines-clone">Roo Code: Cline's Clone</h3><p><strong>Roo Code</strong> is a clone of Cline, very similar but with some different features that are worth trying out. Sometimes Cline might work better for your workflow, and sometimes Roo Code will. It's a good idea to try both and see which fits your needs for a given project or coding style.</p><p><a href="https://cline.bot/?ref=wuu73.org">Cline</a> for VS Code is free, but remember you pay for the API calls unless you're leveraging the <a href="https://github.com/features/copilot?ref=wuu73.org">Copilot</a> subscription trick. Using the VS Code LM API setting in <a href="https://cline.bot/?ref=wuu73.org">Cline</a> with a $10/month Copilot sub is currently the most cost-effective way to get near-unlimited access to powerful models within the agent.</p><h2 id="tldr-quickstart-guide">TL;DR: Quickstart Guide</h2><ul><li><strong>Models &amp; Roles:</strong></li><li>Planning &amp; Brainstorming: Gemini 2.5 Pro (AI Studio), o4-mini (OpenRouter), Claude 3.7 (Poe), if you have OpenAI Playground configured for the 250k free daily tokens, I recommend using those up with o3 and GPT 4.5... VERY good</li><li>Problem Solving &amp; Debugging: GPT-4.5 (free tokens in Playground), Claude 4 (free daily on Poe)</li><li>Actual Coding: GPT-4.1 via Cline; fallback to Claude 3.5 if needed</li><li><strong>Key Tools:</strong></li><li>VS Code and Trae for editing/running</li><li>AI Code Prep GUI – locally scan &amp; curate only the files you need, saves so much time</li><li>Cline (VS Code agent) for step-by-step code execution</li><li>Free web chats for multi-perspective advice: Poe.com, ChatGPT, Grok, Deepseek, Perplexity, OpenAI Playground, AI Studio w/Gemini 2.5 Pro, Openrouter, duck.ai</li><li><strong>Quick Workflow:</strong> 1. Run AI Code Prep GUI to bundle your (if already existing) project's relevant files. 2. Paste that context into your favorite web chat models for planning &amp; debugging. 3. Ask one model to "Write me a detailed Cline prompt for these tasks," then refine it (e.g. in ChatGPT). 4. Copy/paste into Cline set to GPT-4.1 to generate or fix code; if it stalls, switch to Claude 3.5.</li><li><strong>Cost-Saving Hacks:</strong></li><li>Enable "share data" in OpenAI Playground for 250k free GPT-4.5, o3, (both genius expensive models) &amp; 2.5 MILLION free tokens/day for o4-mini, o3-mini!!</li><li>$10/mo GitHub Copilot subscription gives you rate-limited access to Claude models via Cline</li><li>Pay-as-you-go on OpenRouter for o4-mini, Claude 3.7, and other new models</li></ul><p>## Some Thoughts AI is an incredible force multiplier, but it's not a magic wand. The real magic happens when you combine your curiosity, persistence, and willingness to experiment with these powerful tools. Don't get discouraged by bugs or setbacks—every challenge is a chance to learn something new. Mix and match models, try wild ideas, and don't be afraid to break things and rebuild. The best coders aren't the ones who never get stuck—they're the ones who keep moving forward, using every tool and trick at their disposal. Embrace the chaos, enjoy the process, and let your creativity lead the way!</p><h3 id="latest-model-updates-2025">Latest Model Updates (2025)</h3><p><strong>Claude Sonnet 4 &amp; Opus 4</strong></p><ul><li><strong>Status:</strong> Just released and proving to be the best models for everything</li><li><strong>Performance:</strong> Fixed some bugs today that were hard for most other models to handle</li><li><strong>Cost:</strong> More expensive, but accessible through GitHub Copilot $10/month (rate limited)</li><li><strong>GitHub Copilot Value:</strong> The cheap $10/mo GitHub Copilot subscription seems to be giving me a very generous amount of Claude 4 uses through the VS Code LM API. I keep using it as my main model in Cline/Roo Code, and it has not run out yet. I either use GPT 4.1 or Claude 4, been waiting for it to limit me but it hasn't yet (they say that GPT 4.1 is unlimited). I always hate subscriptions for anything, but this particular one is insanely good value.</li><li><strong>Strategy:</strong> Save Claude 4 for tough problems, use GPT 4.1 for regular coding</li></ul><p><strong>GPT 4.5</strong></p><ul><li><strong>Performance:</strong> Excellent for bug fixing and complex problem solving</li><li><strong>Token Limits:</strong> 250k tokens daily if you allow data for model training</li><li><strong>Cost Hack:</strong> Free under the 250k token limit when you enable data sharing in settings</li><li><strong>Use Case:</strong> Great for using AI Code Prep tool to analyze entire codebases</li></ul><h3 id="current-coding-workflow-2025">Current Coding Workflow (2025)</h3><p><strong>For New Projects:</strong></p><ol><li><strong>Planning Phase:</strong> Type all details in notepad (languages, libraries, servers, etc.)</li><li><strong>Multi-Model Consultation:</strong> Paste into multiple models for different "doctor's opinions": - Gemini 2.5 Pro (free) - GPT 4.1 - o4-mini - Claude 4 on Poe.com (free daily credits)</li><li><strong>Refinement:</strong> Go back and forth to fine-tune details</li><li><strong>Task Generation:</strong> Have model write step-by-step task list for Cline AI coding agent</li><li><strong>Execution:</strong> Copy/paste into Cline (or Roo Code) set to GPT 4.1 for 'act' mode</li></ol><p><strong>For Problem Solving:</strong></p><ul><li>Use GPT 4.5 with AI Code Prep tool for complex codebase analysis</li><li>Ask GPT 4.5 to "write me a prompt for Cline" to complete tasks</li><li>Choose models based on problem complexity</li><li>Use multiple models for different perspectives</li></ul><h3 id="task-list-test-driven-development-coming-soon">Task List &amp; Test Driven Development (Coming Soon)</h3><p><strong>Test Driven Development &amp; Task Lists:</strong></p><p>Coming to this guide soon (these other topics)Have the AIs create a detailed task list for Cline, Roo Code, Trae agent, to execute. You can also instruct Cline or Roo Code to use a markdown file to keep track of everything it does, checking things off as it completes them. This will make it easier to track and ensure nothing gets missed.</p><p>For now, you can experiment by having a model generate a checklist in markdown, and then ask Cline or Roo Code to update the file as tasks are completed.</p><h3 id="money-saving-hacks">Money-Saving Hacks</h3><ul><li><strong>GPT 4.5 &amp; o3:</strong> 250k tokens free daily by enabling data sharing for model training</li><li><strong>Cheaper Models:</strong> 2.5 million tokens on o4-mini (excellent model), 4.1-mini/nano</li><li><strong>GitHub Copilot:</strong> $10/month gives access to new Claude models (rate limited)</li><li>Trae IDE currently has free (not limited either, from what I can tell) Claude 4 and GPT 4.1 no subscription required</li><li><strong>Poe.com:</strong> Free daily credits for every type of model</li><li><strong>Web Interfaces:</strong> Use free web chat interfaces for planning and consultation</li></ul><h3 id="coming-soon-live-reddit-data-insights">Coming Soon: Live Reddit Data &amp; Insights</h3><p><strong>Live Reddit Data Scraping &amp; Daily Insights:</strong></p><p>A new feature is coming soon: live scraping of Reddit data and daily updated info about how people are using AI models. This will include detailed usage breakdowns, data visualizations, and new insights into real-world coding workflows and trends.</p>]]></content:encoded>
        </item>
  </channel>
</rss>
