no bs agentic coding, from my experience building large-scale products. backend systems at scale, prev Senior SRE/backend at gorgias. paris → shanghai

Pro tip: use GPT6.1 sol with /fast. It is 2.5X faster and dirt-cheap
24
People who do bot-automated replies should be banned from X. I barely started posting and this is already unbearable. Also, it makes it really hard for people who actually put in the work to stand out when there's a stream of slop.
6
To the people hype posting "Opus 5.5 usage is unlimited" Stop. Anthropic is a company. As a company, their goal is to charge the customer as much as they are willing to pay and provide them a service that is as shitty as they can tolerate for a given price. They take the temperature to determine this on socials among other sources. If this hype gets large enough, they will actually decrease limits. Why not do it if they can get away with it?
14
I see what Anthropic is trying to do here. It's not advertising of GLM 5.3 capabilities, it's more spreading fear about open models. But it won't work, we need to thank ZAI and other model provides, because without them, without GLM 5.2, HuggingFace would have suffer severe damage from the OpenAI attack a few months ago. We need more cyber capabilities in the Open to play both roles of red/blue teaming and increase security of our systems. anthropic.com/research/glm-5…
22
28
362
11,279
I spent the last 3 days fixing/rewriting Astra slop with Opus. The difference is really that big. In retrospect it really feels like Astra does not get what's important. Very capable autist
18
One agent file cuts Claude Code usage by 25% on a feature I built with 72 agents over 64 hours. Computed call by call on the real trace. Unlike the main convo, sub-agents only keep their prompt cache for 5 minutes, even on a subscription. If one waits longer than that on a test run or a build, its next call pays for its whole context again. In that build it happened 200 times: 31% of the total cost. The worst was one agent polling a slow job every 10 minutes. Each check cost ~20x a normal call (chart in the reply). The fix is a sub-agent with a 1-hour cache. Save this as ~/.claude/agents/long-cache.md: --- name: long-cache description: General-purpose agent whose prompt cache lasts one hour instead of five minutes. Use it for agentic work that will sit idle more than five minutes between its own turns (waiting on long builds, test suites, visual checks, or its own subagents), especially once its context grows large, since every expired cache rewrites the whole context. For short or continuously active tasks use general-purpose: its cache writes cost less. experimental: cacheTtl: 1h --- Work as a general-purpose agent on the task you're given. Then ask Claude to use long-cache for anything that runs your slow commands. Why not give every sub-agent a 1-hour cache? Its writes cost 60% more, and 53 of my 72 agents never waited 5 minutes. For them it's pure waste.
4
1
118
Numbers, from replaying all 11,267 API calls: - average context at a rewrite: 563k tokens - 1-hour cache on only the 18 agents that waited: -25.6% - the agent in the image: 15 hours, 54 rewrites, 52 avoidable, half its cost Details: - the agent file needs Claude Code v2.1.248+, and a global subagentPromptCacheTtl or CLAUDE_CODE_SUBAGENT_PROMPT_CACHE_TTL overrides it - the 1-hour cache is ignored while a subscription is on usage credits - on an API key, your main session is on 5 minutes too; set promptCacheTtl: "1h" for it - replay method: API-price costs from every call's usage data; after a 5-60 min gap the prefix is read instead of rewritten, other writes cost 2x instead of 1.25x, gaps over 60 min still expire Docs: code.claude.com/docs/en/prom… File: gist.github.com/shidenkai0/0…
51
Fable sits in a very weird place right now less capable than Opus 5.5 5X more expensive limited to 50% of your quota
20
My Codex weekly usage after the release of Opus 5.5 Guess that's why I see less people asking @thsottiaux for resets this week?
1
1
32
New models edit files through the shell instead of the write tool all the time. It's completely unreadable in Pi, no diff In my agent transcripts: 30% of Opus 5.5's edits in Claude Code, ~20+% for GPT-6, GLM 5.3, DeepSeek v4 flash and Kimi K3 in Pi. I built an extension that adds a diff card after it. It parses the command with tree-sitter, nothing is executed, ~0.2 ms per edit command/script. Your thoughts @mitsuhiko? Here's the before/after, link below 👇
2
1
132
Ok now I really see the hype with Opus 5.5 GPT6-Sol has no taste, while everytime I speak with Opus 5.5, I feel like it has taste and judgement However it is slightly more sycophantic than before It's really interesting how @anthropic backpedaled from Opus 4.8 where the whole messaging was: "this model pushes back". It turns out we prefer smarter model, even if they are slightly yes-men
40
After 10 years in Tech, Opus 5.5 is the best UI/UX designer I've met.
8
Enjoy Opus 5.5 before it's nerfed !
1
21
I love Opus 5.5 so far, but man the 5h limit really sucks It makes it impossible to use up the weekly limit That's probably on purpose?
22
Ok I have to say Opus 5.5 blew me away I asked it to build a full course to really understand LLMs down to the metal and be able to contribute to local AI on Mac (MLX, Metal) It actually built and TRAINED an 80k params LLMs, I am now going through a full handcrafted course on a real, custom trained model and it looks gorgeous and is super well built
33
How to get unlimited Astra-level execution: Use Astra as an advisor. Just ask GPT6-Sol to have its work reviewed by Astra as a sub-agent
35
Am I crazy if I like GPT6-Sol? Been pretty good so far, but I see everyone complaining.
32