Patterns, frameworks, and tools from GitHub and RSS feeds
We thought about this carefully before choosing hyperbole; but it is warranted.
<p><strong><a href="https://www.anthropic.com/news/claude-opus-5">Introducing Claude Opus 5</a></strong></p> I've been offline <a href="https://en.wikipedia.org/wiki/Elkhorn_Slough">kayaking ...
Opus 5 is a step change improvement for the Opus tier powering long-running agents while delivering improvements in coding and professional work.
Our most agentic Sonnet yet, with top-tier intelligence for coding and everyday professional work.
Total Anthropic victory!
Anthropic's latest AI model is better at coding, sustaining tasks for longer and creating high-quality professional work.
The new Opus model comes with a tool called Dynamic Workflows, for coordinating swarms of subagents.
<p>Tuesday was <a href="https://x.com/ade_oshineye/status/2082129440943866149">Stateless MCP day</a> - the rollout of MCP 2.0, or <a href="https://blog.modelcontextprotocol.io/posts/2026-07-28...
Anthropicβs Claude Sonnet 5 brings stronger agentic capabilities, lower pricing, and improved safety, positioning the model as a cheaper alternative to Opus, GPT-5.5, and Gemini Pro.
The release candidate for the next Model Context Protocol (MCP) specification is now available: a stateless protocol core, the Extensions framework, Tasks, MCP Apps, authorization hardening, and a ...
Our latest model, Claude Opus 4.7, is now generally available. Opus 4.7 is a notable improvement on Opus 4.6 in advanced software engineering, with particular gains on the most difficult tasks.
We removed over 80% of Claude Code's system prompt for more advanced models. How to apply the lessons we learned to your own context engineering in Claude Code and with your own agents.
Our latest model, Claude Opus 4.8, is an upgrade to our Opus class of models, with stronger performance across coding, agentic tasks, and professional work, and the consistency to handle long-runni...
Dreaming, outcomes, and multiagent orchestration are now available in Claude Managed Agents. Build agents that learn, meet a quality bar, and work in parallel.
Weβre upgrading our smartest model. Across agentic coding, computer use, tool use, search, and finance, Opus 4.6 is an industry-leading model, often by wide margin.
Claude Opus 4.6 is here. There is a lot to say.
Claude Opus 5 is trying to be the best of both worlds.
<p><strong><a href="https://platform.claude.com/docs/en/about-claude/models/whats-new-sonnet-5">What's new in Claude Sonnet 5</a></strong></p> Claude Sonnet 5 came out <a href="https://w...
<blockquote cite="https://twitter.com/karpathy/status/2026731645169185220"><p>It is hard to communicate how much programming has changed due to AI in the last 2 months: not gradually and over ...
The new SOTA model asserts its dominance.
<p><em><a href="https://simonwillison.net/guides/agentic-engineering-patterns/">Agentic Engineering Patterns</a> ></em></p> <p>I use the term <strong>agentic engineering</strong> to des...
ain't nobody beats Anthropic at distilling Fable!
Only six weeks after Opus 4.7, we have Opus 4.8.
The accidental "open sourcing" of Claude Code brings a ton of insights.
Early results from MirrorCode benchmark with METR: AI agents can complete weeks-long coding tasks, including reimplementing a 16,000-line codebase.
<p>Earlier this month I hosted a fireside chat session at the <a href="https://www.ai.engineer/worldsfair/2026">AI Engineer World's Fair</a> with Cat Wu and Thariq Shihipar from Anthropic's Cl...
In our first episode of 2026, swyx sits down with the cofounders of Artificial Analysis to discuss the state of LLM Evals and Benchmarks, and the key trends and drivers of LLM progress for the year.
<p><strong><a href="https://www.promptarmor.com/resources/claude-cowork-exfiltrates-files">Claude Cowork Exfiltrates Files</a></strong></p> Claude Cowork defaults to allowing outbound HTTP tr...
Coding agents cross a meaningful threshold with Opus 4.5.
<p><strong><a href="https://www.dbreunig.com/2026/01/08/a-software-library-with-no-code.html">A Software Library with No Code</a></strong></p> Provocative experiment from Drew Breunig, who de...
<blockquote cite="https://steve-yegge.medium.com/software-survival-3-0-97a2a6255f7b"><p>Getting agents using Beads requires much less prompting, because Beads now has 4 months of βDesire Paths...
<p>New from Fly.io today: <a href="https://sprites.dev">Sprites.dev</a>. Here's their <a href="https://fly.io/blog/code-and-let-live/">blog post</a> and <a href="https://www.youtube.com/watch?...
How I use coding agents, and what I think they mean
A Claude Code skill for autonomous skill extraction and continuous learning. Have Claude Code get smarter as it works. - blader/Claudeception
Built into the Claude Desktop app, Cowork lets users designate a specific folder where Claude can read or modify files, with further instructions given through the standard chat interface.
Introducing Agent Readiness Factory can now evaluate how well your codebase supports autonomous development. Run /readin...
Coordinate multiple Claude Code instances working together as a team, with shared tasks, inter-agent messaging, and centralized management.
<p><strong><a href="https://www.kimi.com/blog/kimi-k2-5.html">Kimi K2.5: Visual Agentic Intelligence</a></strong></p> Kimi K2 landed <a href="https://simonwillison.net/2025/Jul/11/kimi-k2/">i...
<p><strong><a href="https://www.anthropic.com/engineering/how-we-contain-claude">How we contain Claude across products</a></strong></p> A complaint I often have about sandboxing products is t...
<p><strong><a href="https://z.ai/blog/glm-5">GLM-5: From Vibe Coding to Agentic Engineering</a></strong></p> This is a <em>huge</em> new MIT-licensed model: 754B parameters and <a href="https...
<blockquote cite="https://www.robinsloan.com/winter-garden/agi-is-here/"><p><strong>AGI is here</strong>!βWhen exactly it arrived, weβll never know; whether it was one companyβs Pro or another...
Open Standards for Rich generative UI is all you need.
<blockquote cite="https://twitter.com/bcherny/status/2022762422302576970"><p>Someone has to prompt the Claudes, talk to customers, coordinate with other teams, decide what to build next. Engin...
Anthropic’s new Cowork tool, Anthropic Raising $10 Billion at $350 Billion Value, Deep Delta Learning
<blockquote cite="https://twitter.com/dhh/status/2012543705161326941"><p><em>[On agents using CLI tools in place of REST APIs]</em> To save on context window, yes, but moreso to improve accura...
<p>Last month I <a href="https://simonwillison.net/2025/Dec/15/porting-justhtml/">wrote about porting JustHTML from Python to JavaScript</a> using Codex CLI and GPT-5.2 in a few hours while al...
<p><strong><a href="https://mitchellh.com/writing/my-ai-adoption-journey">Mitchell Hashimoto: My AI Adoption Journey</a></strong></p> Some really good and unconventional tips in here for gett...
Anthropic is developing a new Customize section for Claude, centralizing Skills, Connectors, and upcoming Commands for Claude Code.
<p>It genuinely feels to me like GPT-5.2 and Opus 4.5 in November represent an inflection point - one of those moments where the models get incrementally better in a way that tips across an in...
<p><strong><a href="https://steve-yegge.medium.com/the-ai-vampire-eda6e4f07163">The AI Vampire</a></strong></p> Steve Yegge's take on agent fatigue, and its relationship to burnout.</p> <bloc...