agent civilizations algal.computer + xcb.sh + gobstopper.sh + peopleblade.com + many more prev @stripe 2015-23 @venmo

platonic space
Pinned Tweet
Don't forget: 4. Property based tests 5. Model checking (TLA+/Kani/Quint/Apalache) 6. Deductive proofs (Lean/Verus) 7. Linearizability, Jepsen-adjacent stuff
IMO the only tests you should have in the age of coding agents are, in order of priority: 1. Full E2E tests. Nothing mocked out at all. Can even be something that runs in prod on test accounts via Playwright or equivalent. 2. Integration tests. Agents make more mistakes as things bleed between data/API boundaries where schemas can drift. 3. Golden tests. This helps ground the code with real examples of data, and can be used as regressions against edge cases. Unit tests are bloat 99% of the time if you’re using a frontier model since they’re smart enough now to get the implementation (and subsequent iterations) right on the first shot, so they only add bloat at this point.
1
22
2,029
this is why i have 17 Codex subscriptions. i'm sorry. keep up > Subscription businesses often count on low utilization and cancellation rates which keeps the prices lower, subsidizing the optimizers and value extractors.
The introduction of the infinite optimization machine won’t create new supply of scarce assets (physical resources, attention, etc.) but will substantially/infinitely create new demand for them.
Article

The future of ruthless optimizers and knowing a guy.

In normal life (pre agents) not everyone is a ruthless optimizer. Most of the time, most people are content to take the path of least resistance while a few freaks value-max their way through life.

3
6
1,136
guys i have a problem. i only have usage left on 3 of my 30 AI subs and there are 3 whole days left in the week i've used 90% of my allotted AI rations when there's still 40% of the week left. fml i hate being compute poor. i want intelligence too cheap to meter. i want astra coming out of my ears and every orifice
11
1
17
2,254
whoa
Can complex multi-step reasoning emerge purely from cells that only talk to immediate neighbors? Happy to share our paper “Reasoning with Neural Cellular Automata (NCA)”, from our team at Google, Paradigms of Intelligence 🧵👇
5
775
i'm on my last 6 (out of 17) Codex Pro 20X (RIP) subs. i'm switching to Claude. these 6 Codex accounts are my black sheep, they always get rate limited (probably shadowbanned) good thing i have my trusty sword Excalibur (xcb) to slay those rate limits and get through the end of the week as i become compute poor. on the weekend i am a beggar scrounging around for free tokens (SWE-2 on @DevinAI) > use xcb to resume my codex sessions on this machine i have like 11 of them that have been active in the last 3 hours or so, the problem is my remaining subs are getting rate limited so we need to like round robin the subs so that we distribute the work as evenly as possible and auto continue when we hit a rate limit like a 429 we should back off but try again pretty soon like give it a minute
7
1
15
1,308
completely absurd. the subscriptions are an amazing deal and no individual is going to pay API prices i'm a proud owner of 17 Codex + 5 Claude subscriptions and i'm having a great time people spend money on frivolous hobbies all the time! why does everyone hate on tokenmaxxers
Replying to @thdxr
also some individuals will pay per token once their limits are hit which also helps the overall system work the irony of these people skirting this is they're always claiming some god tier level of productivity but they still can't afford to pay API prices so what exactly are they doing? is it a good allocation of resources?
14
1
48
9,622
i don't know how to tell you this but your cracked lead engineer "working from home" right now is definitely making music, building 10 different side projects, getting high, or worse i have at least 30 agents running at a time and i still spend most of my time dicking around
20
4
155
6,641
There are two paths as a software engineer in 2026: 1. Learn the domain, write the spec, wait for agents to build the thing, repeat 2. Build the system that builds all future things. Use your newfound freedom to take long walks and brood about the future
5
2
51
2,506
i'm reimagining the DAW one-shotted in a couple hours with a bunch of GPT 6.1 Sol Ultra in my software factory original track "Valhalla" – handcrafted using my human brain and hands (though i'm cooking on some agentic music composition tools, more on that later) sound on 🔊
1
19
1,384
friends of friends are saying they're glad i'm leaning into my cyberpsychosis because hraness dot com is a trove of useful ideas and tools for your agents to visit
6
26
1,367
peak AI psychosis is setting up 25 hillclimbing agents every night before you go to bed so you can wake up and smell all the optimizations they're so excited to tell you about after you make coffee every morning is christmas morning in the singularity
51
40
774
40,014
"...complex, confusing, and largely uncontrollable." That was true of the people building the first computers, and it's true of us now. really great read!
This is the first blog I've written in 8 months. I'm back. Intelligence is free. Good luck. evis.dev/posts/cheap_intelli… ( I have a bunch more that I've been working on that I'll publish over the coming weeks ranging from deeply technical to macro.) Check it out!
10
1,100
italian food is the chinese food of europe french food is the japanese food of europe
1
16
1,007
wtf
OPENAI OFFICIALLY RELEASED THE PRO $500/M PLAN WITH 25X USAGE OF PLUS We are now effectively paying 2.5x more dollars to only get 1.25x more usage $500/m is now 25x usage $200/m is now 10x usage (was 20x) $100/m is now 5x usage That extra 5x usage from 20x to 25x is now effectively $300 a month
4
2
34
4,377
Replying to @thsottiaux
derp i guess i can't read
91
personal agents like Instinct, @Muse and Grok @bot are so 2026 basic, mainstream, totally washed 2027 is personal agent swarms you don't need to wait for 15k tok/s inference from @cerebras. i regularly 20x that speed in the last 10hrs i've spent 6B tokens peak 387k tok/s
My conversation with Noah Shinn (@noahrshinn), founder of Instinct. Noah is building a personal AI assistant. It's still invite only, has spent nothing on marketing, and is growing roughly 10% A DAY. This is his first long conversation about the company. We discuss: - Why Instinct doesn't have an app - Buying compute months ahead of exponential demand - How users learn to trust it with a credit card - Safety and security - Agents coordinating with other people's agents - Instinct's business model - Apps built on consumer inertia - and more Enjoy! Timestamps: 0:00 Intro 4:11 What people are using AI agents for 15:07 Rethinking travel, reservations, and the internet 22:43 Trust, privacy, and personal data 27:50 The business model behind Instinct 38:04 How existing businesses will adapt 47:55 Designing a personal assistant people love 53:15 Growth, compute, and competing with Big Tech 1:11:44 What’s next for Instinct and personal AI
13
1
61
11,060
typical afternoon in the software factory. it's almost 420 so i'm rolling a j while my 25 agents are cooking 12x Claude Opus 5.5 ultracode 8x Codex Astra Ultra 5x @DevinAI SWE-2 ps you can buy 1 oz of top shelf for $100 in puerto rico listening to Skee Mask - Resort
20
2
53
4,307