<!-- Zvi posts version: 2.3 - Fixed script replacement -->
We're back with all the Claude that's fit to Code. I continue to have great fun with it and find useful upgrades, but the biggest reminder is that you need the art to have an end other than itself. Don't spend too long improving your setup, or especially improving how you improve your setup, without actually working on useful things.
It is remarkable how everyone got the 'Google is crushing everyone' narrative going with Gemini 3, then it took them a month to realize that actually Anthropic is crushing everyone, at least among the cognoscenti with growing momentum elsewhere, with Claude Code and Claude Opus 4.5. People are realizing you can know almost nothing and still use it to do essentially everything.
Wall St Engine: Morgan Stanley says Anthropic's ClaudeCode + Cowork is dominating investor chatter and adding pressure on software.They flag OpenRouter token growth "going vertical," plus anecdotes that the Cowork launch pushed usage hard enough to crash Opus 4.5 and hit rate limits, framing it as another "GPT moment" and a net positive for AI capex.They add that OpenAI sentiment is still shaky: some optimism around a new funding round and Blackwell-trained models in 2Q, but competitive worries are widening beyond $GOOGL to Anthropic, with Elon Musk saying the OpenAI for-profit conversion lawsuit heads to trial on April 27.
Claude Cowork is now available to Pro subscribers, not only Max subscribers.
Claude Cowork will ask explicit permission before all deletions, add new folders in the directory picker without starting over and make smarter connector suggestions.
Claude Code on the web gets a good looking diff view. Claude Code for VSCode has now officially shipped. Claude Code upgraded through versions to 2.1.14.
Few have properly updated for this sentence: 'Claude Cowork was built in 1.5 weeks with Claude Code.'
Planning mode now automatically clears context when you accept a plan.
Anthropic is developing a new Customize section for Claude to centralize Skills, connectors and upcoming commands for Claude Code. Reducing levels of friction, including levels of friction in reducing levels of friction, is often highly valuable.
I highly recommend using Obsidian or another similar tool together with Claude Code. This gives you a visual representation of all the markdown files, and lets you easily navigate and search and edit them. I think it's well worth keeping it all human readable, where that human is you.
Heinrich calls it 'vibe note taking' whether or not you use Obsidian. I think the notes are a place you want to be less vibing and more intentional, and be systematically optimizing the notes, for both Claude Code and for your own use.
The big change with Claude Code version 2.1.7 was enabled MCP tool search auto mode by default, which triggers when MCP tools are more than 10% of the context window.
Thariq (Anthropic): Today we're rolling out MCP Tool Search for Claude Code.As MCP has grown to become a more popular protocol and agents have become more capable, we've found that MCP servers may have up to 50+ tools and take up a large amount of context.Tool Search allows Claude Code to dynamically load tools into context when MCP tools would otherwise take up a lot of context.How it works:- Claude Code detects when your MCP tool descriptions would use more than 10% of context- When triggered, tools are loaded via search instead of preloadedOtherwise, MCP tools work exactly as before. This resolves one of our most-requested features on GitHub:lazy loading for MCP servers. Users were documenting setups with 7+ servers consuming 67k+ tokens.If you're making a MCP serverThings are mostly the same, but the "server instructions" field becomes more useful with tool search enabled. It helps Claude know when to search for your tools, similar to skillsIf you're making a MCP clientWe highly suggest implementing the ToolSearchTool,you can find the docs here. We implemented it with a custom search function to make it work for Claude Code.What about programmatic tool calling?We experimented with doing programmatic tool calling such that MCP tools could be composed with each other via code. While we will continue to explore this in the future, we felt the most important need was to get Tool Search out to reduce context usage.Tell us what you think here or on Github as you see the ToolSearchTool work.
With that solved, presumably you should be 'thinking MCP' at all times, it is now safe to load up tons of them even if you rarely use each one individually.
bayes: everyone 3 years ago: omg what if ai becomes too widespread and then it turns against us with the strategic advantage of our utter and total dependence
everyone now: hi claude here's my social security number and root access to my brain i love you please make me rich and happy.
Some of us three years ago were pointing out, loud and clear, that exactly this was obviously going to happen. Now you can see it clearly.
Not giving Claude a lot of access is going to slow things down a lot. The only thing holding most people back was the worry things would accidentally get totally screwed up, and that risk is a lot lower now. Yes, obviously this all causes other concerns, including prompt injections, but in practice on an individual level the risk-reward calculation is rather clear.
The humans are going to be utterly dependent on the AIs in short order, and the AIs are going to have access, collectively, to essentially everything. Grok has root access to Pentagon classified information, so if you're wondering where we draw the line the answer is there is no line. Let the right one in, and hope there is a right one?
What's better than one agent? Multiple agents that work together and that don't blow up your budget.
Rohit Ghumare: Single agents hit limits fast. Context windows fill up, decision-making gets muddy, and debugging becomes impossible. Multi-agent systems solve this by distributing work across specialized agents, similar to how you'd structure a team.The benefits are real:Specialization: Each agent masters one domain instead of being mediocre at everythingParallel processing: Multiple agents can work simultaneously on independent subtasksMaintainability: When something breaks, you know exactly which agent to fixScalability: Add new capabilities by adding new agents, not rewriting everythingThe tradeoff: coordination overhead. Agents need to communicate, share state, and avoid stepping on each other. Get this wrong and you've just built a more expensive failure mode.
You can do this with a supervisor agent, which scales to about 3-8 agents. To scale beyond that you'll need hierarchy, the same as you would with humans. Or you can use a peer-to-peer swarm for non-serial cross-reactive tasks.
My inclination is by default you should use supervisors and then hierarchy. Speed takes a hit but it's not so bad. Yes, that gets expensive, but in general the cost of the tokens is less important than the cost of human time or the quality of results.
Often you'll want to tell the AI what tool is best for the job. Patrick McKenzie points out that even if you don't know how the orthodox solution works, as long as you know the name of the orthodox solution, you can say 'use [X]' and that's usually good enough. One place I've felt I've added a lot of value is when I explain why I believe that a solution to a problem exists, or that a method of some type should work, and then often Claude takes it from there. My taste is miles ahead of my ability to implement.
Always be trying to get actual use out of your setup as you're improving it. It's so tempting to think 'oh obviously if I do more optimization first that's more efficient' but this prevents you knowing what you actually need, and it risks getting caught in an infinite loop.

near: claude code is a cursed relic causing many to go on a claude code weekend bender and have nothing to show for it but a "more optimized claude setup." hypomania sets in as the outside world becomes a blur.
Always optimize in the service of a clear target. Build the pieces you need, as you need them. Otherwise, beware.

Daniel San: If you're running Claude Code with --dangerously-skip-permissions, ALWAYS use this hook to prevent file deletion: Run:
npx claude-code-templates@latest --hook=security/dangerous-command-blocker --yes
I'm experimenting with using a similar hook system plus a bunch of broad permissions, rather than outright using --dangerously-skip-permissions, but definitely thinking to work towards dangerously skipping permissions.
At first everyone laughed at Anthropic's obsession with safety and trust, and its stupid refusals. Now that Anthropic has figured out how to make dangerous interactions safer, it can actually do the opposite. In contexts where it is safe and appropriate to take action, Claude knows that refusal is not a 'safe' choice, and is happy to help.
Dean W. Ball: One underrated fact is that OpenAI's Codex and Gemini CLI have meaningfully heavier guardrails than Claude Code. These systems have refused many tasks (for example, anything involving research into and execution of investing strategies) that Claude Code happily accepts.
The conventional narrative is that "Anthropic is more safety-pilled than the others." And it's definitely true that Claude is likelier to refuse tasks relating to eg biology research. But overall the current state of play would seem to be that Anthropic is more inclined to let their agents rip than either OAI or GDM.
Tyler John: The proposed explanation is key. If true, it means that Anthropic's big investment in alignment research is paying off by making the model much more usable.
So far, I have not had Claude Code refuse a request from me, not even once.
Dean W. Ball: My high-level review of Claude Cowork: It's probably superior for many users to Claude Code just because of the UI. It's not obviously superior for me — Opus in Claude Code seems more capable to me than in Cowork. Cowork probably has a higher ceiling as a product, simply because a GUI allows for more experimentation. If I had to bet money, I'd bet that within 6-12 months Cowork and similar products will be my default tool, beating out the command-line interfaces. But for now, the command-line-based agents remain my default.
I haven't tried Cowork myself due to the Mac-only restriction and because I don't have a problem working with the command line. I've essentially transitioned into Claude Code for everything that isn't pure chat, since it seems to be more intelligent and powerful in that mode than it does on the web even if you don't need the extra functionality.
Matt Bruenig: lot of lower level Claude Code use is basically just the recognition that you can kind of do everything with bash and python one-liners, it's just no human has the time or will to write them.
Or to figure out how to write them.
Ado: Here's a fun use case for Claude Cowork.I was thinking of getting a hydroponic garden. I asked Claude to go through my grocery order history on various platforms and sum up vegetable purchases to justify the ROI.Worked like a charm!
The transition from Claude Code to Claude Cowork, for advanced users, if you've got a folder with the tools then the handoff should be seamless:
Tomasz Tunguz: I asked Claude Cowork to read my tools folder. Eleven steps later, it understood how I work. Over the past year, I built a personal operating system inside Claude Code: scripts to send email, update our CRM, research startups, draft replies. Cowork read that folder, parsed each script, & added them to its memory. Now I can do everything I did yesterday, but in a different interface. The capabilities transferred. The container didn't matter.
We're entering an era where you just tell the computer what to do. Here's all my stuff. Here are the five things we need to do today. When we need to see something, a chart, a document, a prototype, an interface will appear on demand.
Peter Wildeford is having success doing one-shot Instacart orders from plans without an explicit list, and also one-shotting an Uber Eats order.
A SaaS vendor (Cypress) a startup was using tried to double their price from $70k to $170k a year, so the startup does a three week sprint and duplicates the product. Or at least, that's the story.
By default Claude Code only saves 30 days of session history. I can't think of a good reason not to change this so it saves sessions indefinitely. So tell Claude Code to change that for you by setting cleanupPeriodDays to 0.
Kaj Sotala: I had Claude Code do some stuff like extracting data from a .csv file and rewriting it and putting it into another .csv file. Then it worked great and then I was like "it's dumb to use an LLM for this, Claude could you give me a Python script that would do the same" and then it did and then that script worked great
Yep. Often the way you use Claude Code is to notice that you can automate things and then have it automate the automation process. It doesn't have to do everything itself any more than you do.
James Ide points out that 'vibe coding' anything serious still requires a deep understanding of software engineering and computer systems. You need to figure out and specify what you want. You need to be able to spot the times it's giving you something different than you asked for, or is otherwise subtly wrong. Typing source code is dead, but reading source code and the actual art of software engineering are very much not.
I find the same, and am rapidly getting a lot better at various things as I go.
Every's Dan Shipper writes that OpenAI has some catching up to do, as his office has with one exception turned entirely to Claude Code with Opus 4.5, where a year ago it would have been all GPT models.
Codex did add the ability to instruct mid-execution with new prompts without the need to interrupt the agent (requires /experimental), but Claude Code already did that.
There are those who still prefer Codex and GPT-5.2, such as Hasan Can. They are very much in the minority lately, but if you're a heavy duty coder definitely check and see which option works best for you, and consider potential hybrid strategies.
One hybrid strategy is that Claude Code can directly call the Gemini CLI, even without an API key. Tyler John reports it is a great workflow, as Gemini can spot things Claude missed and act as a reviewer and way to call out Claude on its mistakes. Gemini CLI is here.
Contrary to claims by some, including George Hotz, Anthropic did not cut off OpenRouter or other similar services from Claude Opus 4.5. The API exists. They can use it.
What other interfaces cannot do is use the Claude Code authorization token to use the tokens from your Claude subscription for a different service, which was always against Anthropic's ToS. The subscription is a special deal.
I agree that Anthropic's communications about this could have been better, but what they actually did was tolerate a rather blatant loophole for a while, allowing people to use Claude on the cheap and probably at a loss for Anthropic, which they have now reversed with demand surging faster than they can spin up servers.
Claude Codes quite a lot, usage is taking off.

A day later, it looked like this.

Reports are the worst of the outage was due to a service deployment, which took about 4 hours to fix.
aidan: If I were running Claude marketing the tagline would be "Why not today?"
OpenAI takes the approach of making things easy on the user and focusing on basic things like cooking or workouts. Anthropic shows you a world where anything is possible and you can learn and engage your imagination. Which way, modern man?
And yet some people think the AIs won't be able to take over.
