Welcome back to Agentic Coding Weekly.
Recently built this small macos app, FainLane, to solve my personal annoyance of pasting screenshots when running CLI coding agents on a remote VM. It just makes Ctrl+V work for the remote agent. Check it out if you have a similar remote machine workflow.
Here are the updates on agentic coding tools, models, and workflows for the week of Sep 20 - 26, 2026.
1. Tooling and Model Updates
Claude Opus 5.5, GPT 6 Sol and Luna, Mimo 2.6 Pro and Flash, and Grok 4.7 are the newly released models last week.
Opus 5.5 fixes the jargon-filled claude speak and is priced 20% cheaper than previous opus models, $5 / $25 -> $4 / $20. Pricing for both GPT 6 Sol and Luna reduced by 50% compared to 5.6, $4 / $20 -> $2 / $10 for Sol and $0.2 / $1.2 -> $0.1 / $0.5 for Luna. MiMo 2.6 Pro is 1.02T / 42B MoE and beats GPT 5.6 Sol and the massive 2.4T Kimi K3 on most benchmarks.
In terms of tool updates, Claude Code now uses a small, fixed allowance pulled from weekly limit to stop gracefully when 5-hour limit is hit mid-task.
Here's the current state of coding benchmarks:
Model | DeepSWE 1.1 | Frontier Code 1.1 Main | Terminal-Bench 4.0 | Pricing |
|---|---|---|---|---|
Claude Opus 5.5 | - | 54.6% | 66.4% | $4 / $20 |
MiMo 2.6 Pro | 71.9% | - | 34.9% | $0.435 / $0.87 |
GPT 6 Sol | 68.8% | 49.3% | 31.2% | $2 / $10 |
Grok 4.7 | 71% | 47.6% | 37.6% | $2 / $6 |
DeepSeek V4.1 Flash | 74% | - | 31.2% | $0.15 / $0.6 |
GPT 6 Astra | 74% | 53.3% | 57.9% | $10 / $50 |
Fable 5.1 | - | 50.9%. | 55.8% | $10 / $50 |
Opus 5 | 74% | 53.4% | 52.3% | $5 / $25 |
GPT 5.6 Sol | 73% | 47.5% | 37.3% | $4 / $20 |
Kimi K3 | 69% | 44.2% | - | $3 / $15 |
2. Open Source Corner
whiteboard - desktop app where you build with coding agents on a canvas instead of an IDE, kinda impressive but hard to explain, you have to see the demo to understand if this is for you or not
reladraw - create diagrams where you decide how to arrange the diagram, without manually drawing in something like draw io. I've used mermaid and d2 a lot, and more often than not, the layout engine has frustrated me. So this one is a bit interesting
ollaya - ollama for open-source Jev-style decision models
3. Reading List
Plan mode is dead - separately, claude code is planning to sunset plan mode feature and funnily antigravity just added a plan mode. Plan mode and planning are different though. I haven't been using plan mode for at least half an year cause I run all coding agents in yolo mode on a remote machine and do the planning in the yolo mode.
Maintaining quality in the age of Agentic Engineering - from r/ExperiencedDevs
Opus 5.5 is good at explainer videos - from r/ClaudeAI
Attention is all you have - bonus, not about agentic coding
One more bonus, here’s a comment from Simon Willison on MCPs:
Sure, there's almost no reason to use MCPs if you are running a full-blown terminal agent (Claude Code, Codex, Meta Muse, OpenClaw etc) with unfettered internet access - just let it call APIs directly.
If you want to operate something that's less YOLO than that, you'll find yourself wanting:
Control over exactly which external services it can access
A way to handle authentication that doesn't allow the agent to directly access API keys
A sensible UI to allow users to connect and authenticate further services
Strong audit logging for what's going on
MCP makes all of that so much easier to provide. Thinking MCP is obsolete because full coding agents don't need it misses out on all of the other things we might want to build.
That’s it for this week. OpenAI Dev Day is this Tuesday, Sep 29, so expect some new stuff from OpenAI and maybe other labs as well. I’ll be back next Monday with the latest agentic coding updates.
— Prashant
