engineering @qoder_ai_ide, created umijs, dvajs, mako and neovate code.

浙江, 中华人民共和国
100 秒看懂个人 agent 四方对比 🎉 grok bot、muse、cue、dots,50 天里四家做了同一个东西。 -- built with opus 5.5
1
2
19
3,456
旅行间隙给自己写了个浏览器,今天开源,叫 Tiller。 Chromium 内核,Swift 原生壳,侧边栏里跑一个 coding agent 来操作网页。 常见做法是让 agent 从外面遥控浏览器。我想反过来,让 agent 坐在浏览器里,用浏览器自带的工具。开 tab、读页面、点按钮、输入文字,就这几个。同一套工具也通过 MCP 和 tiller 命令行给外部脚本用。 9 月 28 号第一个 commit,今天开源,32 个 commit,Rust 加 Swift 大约 1.3 万行,期间开了 37 个 Claude Code session。 agent 默认只拿到浏览器工具,读写文件和跑命令都关着。网页内容会试图指挥 agent,这个问题我还没有好答案,所以先关掉。 Chrome 的 cookie、密码、历史和扩展能导进来。扩展能跑,但 Chrome 的 tab 和 window API 看不到 Tiller 的 tab,这块还差不少。 自用为主,你的情况可能不一样。 github.com/sorrycc/tiller
15
6
76
14,030
Opus 5.5 用了一天,攒了 24 条 tips,分五块。 一、怎么用 1. 整个任务直接扔给它。说清什么算做完,什么时候要回来找你。 2. 别写「think carefully」,它本来就会先想。思考模式也关不掉了。 3. 长任务跑完,问它「还缺什么才能继续」,别急着自己接手。 4. 长任务跑一半停下来汇报是前沿模型的通病。Anthropic 给了一段专门的提示词压住这个毛病。 5. 默认 medium 就够。智能跟 Fable 5.1 差不多但更快,输出 token 比官方建议还少两成。 6. 别开 max。有人两次测试都在思考阶段撞上 128K 输出上限,一个结果没出,每次 2.56 美元、20 分钟。Fable 5.1 的 max 反而没事。 7. 简单活 medium,难活切 high。Claude Code 2.1.280 之后,会话中途切强度不再打掉缓存。 二、把旧提示词清一遍 8. 跑 /claude-api prompt-audit。skills、CLAUDE.md、AGENTS.md 里那些给旧模型写的保姆规则,现在只会拖累 5.5。 9. 该删的典型句式:全大写 IMPORTANT、一刀切的 be concise、always ask、always summarize、be thorough / don't stop。 10. 每次模型升级都重跑一遍,拿几个代表任务验证,只留有效果的改动。 11. 反模式不只在 CLAUDE.md,工具描述里也有。有人把 12 个工具砍到 4 个,效果比改任何提示词都大。 12. 让 audit 解释每一条删除理由。有些看着奇怪的行,是绕某个 bug 的 workaround。 三、账单 13. 几乎所有事都能用 Opus 5.5 干,Fable 那 50% 的上限可以不管了。只有「给我点惊喜」那种模糊的创意原型,我还是回去找 Fable。 14. 多 agent 编排里,coordinator 和 worker 都换成 5.5,额度掉得明显慢了。 15. 缓存读取从 0.5 降到 0.2 美元。编程会话缓存命中率一般 95% 以上,实际省的比标价多。 16. 三家里只有 Anthropic 不对长上下文加价,到 1M 单价不变。OpenAI 过 272K 翻倍,Grok 过 200K 翻倍。 17. banked reset 不会挪动每周重置时间。留到周期后半段额度见底再用。 四、顾问模式 18. /advisor fable 把 advisor 设成 Fable。Fable 出方案,5.5 执行,做完让 advisor 验收,按反馈返工到通过为止。配 /goal 更顺。 19. Opus 和 Fable 的真正差距在找开关。Opus 被 auto-mode 拦住会让你手动粘命令,Fable 会去 settings.json 里翻出 autoMode.allow。所以 Opus 当主力,Fable 别关。 五、它变了什么 20. 只在该长的时候长。「Claude 腔」和破折号没了,结论放最前面。 21. 「honest mistake」循环没了。以前是先下结论、走错、道歉、再花几轮修,现在一次做对。 22. 会自己遵守 skill,不用你提醒。 23. 3D、视觉、Blender、设计品味都明显强了。视频、Minecraft、3D 网页一次成型。 24. 网络安全和生物任务会被自动转给 Opus 4.8。2026-08-31 后注册的 API 账号拿不到推理过程。输出带文本水印。
24
16
202
28,281
그록봇 스타일 케릭터 만들어 주는 프롬프트 공유 grokbot-icon-studio.serio-ai… 파딱이 아니라 긴 텍스트 업로드가 안되어 아예 웹앱 형태로 배포합니다. 다음 사이트에서 복사 버튼을 누르고 사용하는 이미지 생성 Ai에 붙여넣기해서 활용해 주세요
Made with AI
1
13
4,205
Cute.
그록봇 스타일 케릭터 만들어 주는 프롬프트 공유 grokbot-icon-studio.serio-ai… 파딱이 아니라 긴 텍스트 업로드가 안되어 아예 웹앱 형태로 배포합니다. 다음 사이트에서 복사 버튼을 누르고 사용하는 이미지 생성 Ai에 붙여넣기해서 활용해 주세요
Made with AI
2
6
105
47,139
接下来几天在云栖大会 Qoder 展台看摊,逛云栖的同学可以来找我唠会
10
2
35
8,026
9 月 23 日前晒一个用 Qoder 做的项目,有机会拿 20K Sonus credits 。
More ideas than your daily 2K Sonus credits can handle? Show us what you’re building with Qoder. 🏆 3 standout builders will win 20K Sonus credits each. 🔥Builds that turn heads may unlock surprise credits! To enter: 1. Follow @qoder_ai_ide 2. Post a Qoder-made project, demo, workflow, or WIP 3. Tag @qoder_ai_ide + #Qoder #BuildWithQoder Entries close Sep 23, 24:00 Beijing time (16:00 UTC).
3
7
3,957
Jev 玩马里奥,单次决策延迟 229ms,1-1 都没过,对实时性要求高的游戏还是不够。
14
67
23,119
拿一个小项目试了下,源码减少 30.9%。
recommended reading. I want a massive simplification set of PRs. or a single monolithic PR. I want LOC to drop dramatically. Minimum 30% overall. I want god files broken up. I want simplification across the board. I want unification of helpers and methods that can be reused. I want less if-if-if-if-if-if-else routing. I want code legibility up. I want interpretability of the codebase and how things connect to each other up. I want elegance. I want superfluous excess bloat code cleaned up and removed. I want it all done fully. No excuses. No waiting for my decisions. Get it all done, and present me a PR or set of PRs when done.
1
5
32
20,933
Fable looked at the edit tool, said no thanks.
2
1,595
Claude Code 团队自己,已经有 70~80% 的工作不在终端里做了。 看完 Thariq、Sid、Robert 三个人的 22 分钟对谈,记了几条。 1. 大部分活在 Slack 里的 Claude Tag 上干。TUI 和桌面端只有两种时候才开,精修某个东西,或者"想微观管理我的那些 Claude"。给的也不再是任务,是目标。 2. 以前底层技术的保质期以年计,现在是两个月。Sonnet 3.5 时代模型给它五件事只做三件就放弃,于是加了 to-do list,一下就好了。一年后这个功能没了,因为不需要了。Sid 的结论是对自己造的东西不要有感情。 3. harness 里大部分功能,说白了是在给当前模型的失败模式打补丁。模型变强就主动删。但任务尺度同时在变大,又得造形态完全不同的新工具。这两句得一起记,只记前一句就是为删而删。 4. AskUserQuestion 这个工具磨了很久才让模型学会调用。结果作者现在自己都不怎么用了,直接让 Claude 出一个带 mockup 的 artifact 来问他问题。 5. Code review 变了。人类 reviewer 挑三个小 nitpick,其实是在证明"我读了"。这活交给 Claude。人只管 Claude 不知道的事,这个 API 为什么长这样,服务边界为什么画在这。 6. workflows 是从 code review 长出来的。fan out 找 bug,每个 bug 再从几个角度对抗性验证,过滤掉误报只留真的。Robert 信任 workflows 的理由很简单,编排层是代码,for 循环不会漏掉任何一项。 7. Claude Tag 第一次把界面和 transcript 彻底分开。Slack 里看到的每条消息都是它主动调工具发的,内心独白看不到。他们一开始很慌,后来发现不盯着也能拿到好结果,才意识到模型真的到这一步了。 8. 怀念什么?Sid 怀念性能优化,"现在 Claude 比我强得多"。Robert 曾经花一整天手写 CSS 复刻 Mac OS 10.4 的 Aqua 按钮,叠了好几层 radial gradient。他说再也不会手工做这种事了。听着有点唏嘘。
I talked to Sid & Robert about building Claude Code: how much things have changed, how hard it's been to keep up with model capabilities and also what we miss about software engineering before AI. piped.video/watch?v=S-sYlFiG…
42
38
327
97,236
Interesting prompt.
On macOS and wanna see something interesting? Ask your agent: "look at ~/Library/Application Support/Knowledge/knowledgeC.db and tell me some interesting facts" I had no idea this stuff was all getting logged and it can infer a lot about your activity.
1
3
15
21,052
recommended reading. I want a massive simplification set of PRs. or a single monolithic PR. I want LOC to drop dramatically. Minimum 30% overall. I want god files broken up. I want simplification across the board. I want unification of helpers and methods that can be reused. I want less if-if-if-if-if-if-else routing. I want code legibility up. I want interpretability of the codebase and how things connect to each other up. I want elegance. I want superfluous excess bloat code cleaned up and removed. I want it all done fully. No excuses. No waiting for my decisions. Get it all done, and present me a PR or set of PRs when done.
New blog post: We had a million lines of Python to clean up. On September 2nd @Teknium asked Hermes Agent to do it. 1,393 subagents and nineteen hours later, the codebase was 34.4% smaller, saving us nearly $2m in engineering hours. nousresearch.com/refactoring…
5
12
191
59,880
gpu-doodle:一个跑在浏览器里的涂鸦识别小模型。 - 边画边猜,第一笔画个圆它会在月亮、时钟、太阳、苹果之间摇摆,第二笔下去就定了 - 100 个 Quick, Draw! 类别,测试集 top-1 89.4% - 整个包(int6 权重 + 手写 WGSL 内核)Brotli 后 26 KB,3.2 万参数 - 有 WebGPU 就走 GPU,没有就走 CPU,任何现代浏览器都能玩 - 自带 20 秒限时的"你画我猜"模式,中英文界面跟随系统 训练数据是 Google 的 Quick, Draw!(CC BY 4.0),做法参考了 arikchakma 的 gpu-time。 Demo:sorrycc.github.io/gpu-doodle… 仓库:github.com/sorrycc/gpu-doodl…
11
2,363
Qoder duck stop motion x 100 with ChatGPT Images 2.5, each one 36 frames and 3 seconds, in 100 different visual styles: sorrycc.github.io/qduck-stop…
1
7
2,136
Fable 5.1 tips for saving money: 1. Don't run xhigh by default. Stay on high. On the Artificial Analysis index, xhigh burns about 2x the output tokens of high. $5.98 vs $3.91 per task, and 2 more points on the score. Those numbers drift, so check them before you copy me. 2. Opus by default, then /advisor fable when you're stuck. Cheap model does the typing, expensive model reviews. 3. Or flip it. Fable by default for design, then /model and press s to toggle to Opus for implementation. 4. Or use fable xhigh to design, then /effort and press s to toggle down to high to implement. Changing thinking effort doesn't invalidate the prompt cache. 5. /subtask for anything that spews output. Only the result comes back and the main session stays clean.
1
1
28
6,520
deepseek 4.1 flash is sooooo fast that for anything non-complex i dont need a bigger model anymore.
3
5
2,902