Good afternoon from Vienna People of Pi ☀️ Some slightly different Sunday meditations today in which @mitsuhiko talks about some shared frustrations with agentic software engineering. Remember that you're not alone in finding it harder than it looks!

Sep 20, 2026 · 1:59 PM UTC

31
45
491
55,153
The thread that was mentioned:
What when it comes to AI in software engineering are you struggling with the most?
7
7,886
Sort replies: Relevant Recent Liked
Replying to @pidotdev @mitsuhiko
Getting strong Uncle Bob Morning Bathrobe Rant vibes xD
2
12
663
Replying to @pidotdev @mitsuhiko
Sunday meditation, Monday agent debugging :)
346
Replying to @pidotdev @mitsuhiko
3:20-4:05 is unfortunately something that I observed myself as well. And it really saddens me deeply...
1
202
Replying to @pidotdev @mitsuhiko
Spent four hours undoing what took it thirty seconds to break.
1
356
Replying to @pidotdev @mitsuhiko
Even for someone like me, who doesn't read code (can't) this resonates strongly as working on some of my stuff has developed this weird rythm that some changes just work and others (not much different in shape) seem to fail endlessly before they work eventually.
62
Replying to @pidotdev @mitsuhiko
Thank you @mitsuhiko for this! I’m feeling the same way and trying to have productive conversations with some people is just impossible in this current timeline. I don’t think it’s a “skill issue” but you can’t seem to talk reason with someone who already believes this.
1
534
Replying to @pidotdev @mitsuhiko
Spot on. Since Astra landed, I noticed a deterioration in my production, not advancing. it felt not reliable with too much fluctuations in output quality. Perhaps my prompting had to adjust to this model. I have switched to GLM5.3 Max & Flash locally, they are reliable and stable
1
272
Replying to @pidotdev @mitsuhiko
"inherently non-deterministic" Remember in the normal coding times you could not reproduce bugs ? The horror.
337
Replying to @pidotdev @mitsuhiko
Let's tackle this. Do you think you could produce identical projects as perfect code and as slop in a couple formats? I wouldn't dare call my code good, so I can't produce that, but I would have some ideas on what to look into with that.
1
191
Replying to @pidotdev @mitsuhiko
We have the tools for QA since 50 years. Static code analysers, formatters, test coverage, profilers, logic resolvers. They are very fast. We can measure cognitive complexity. Requirements/specs, architecture, design are still a (truth maintenance) problem. 1/
1
217
Replying to @pidotdev @mitsuhiko
It's true and most people feel this way, but if you mentioned that to your manager, you'd get fired... and replaced with the AI first engineer that DGAF about how the code looks. I mean, I'm AI first, but there's just some criteria the code has to meet, other than "just work".
35
Replying to @pidotdev @mitsuhiko
I hear trust, agents and tons of PR and I 'member @poteto x.lingyaoai.com/poteto/status/21020504… wud luv to see armin/mario x lauren talk
here's how i shipped 2,500 PRs last month to production this was originally supposed to be for Cursor Compile in London. i couldn't make it since i was livestreaming for Grok @Bot Galaxy so i'm making it available for free here on X! watch it on 2x speed, i talk slowly
22
Replying to @pidotdev @mitsuhiko
Thank you for this! I’ve been feeling this a lot recently too. It’s cool that we’re able to contribute code so much faster (especially new engineers or on new projects). But we’re sacrificing understanding for speed which will eventually have an inverse effect when we can’t understand anything anymore. It’s awesome that I was able to contribute a performance fix to Zensical which is written in Rust, a language which I don’t know well. I understood what I could, but they are responsible for what goes into the repo ultimately. It unblocked me and they accepted the quality, great! But we MUST NOT apply the same bar to projects where we are the core maintainers. However, that’s what’s happening unfortunately. Quick though is to have some global or repo-level skill that asks questions throughout a session or towards the end to probe for understanding. But I agree, there needs to be some deeper process or cultural change.
210
Replying to @pidotdev @mitsuhiko
The AI tortoise will still win the race...... busy on the inside! Quietly munching through the code on the outside!
145
Replying to @pidotdev @mitsuhiko
not sure how important it will be in the future that the code is 'nice' from a human perspective
56
Replying to @pidotdev @mitsuhiko
the context window management alone breaks half the assumptions
1
261
Replying to @pidotdev @mitsuhiko
models getting rewarded on success has been reinforced to such extent that their output doesn't take into consideration that humans are in the loop. the harness has to step in here to guide the output through feedback at different stages.
116
Replying to @pidotdev @mitsuhiko
the main selling point of ai/agentic-engineering was - you don't have to look at code anymore.
1
34
Replying to @pidotdev @mitsuhiko
I share the experience that the new models feel like a racing car with no traction control. They are powerful and hard to steer. What remedies have you been trying in your workflow? I’ve been trying more design-contract prototypes to get a clearer boundary before they build too much.
1
229
Replying to @pidotdev @mitsuhiko
Let's do something.
115
Replying to @pidotdev @mitsuhiko
My thoughts exactly
61
Replying to @pidotdev
@mitsuhiko so first, as the creator of Pi, you get me addicted to something more addictive than any drug. And then you make this video that made me spend an entire evening reflecting on my own journey that's very similar to what you're saying. x.lingyaoai.com/pyronaur/status/210212… On the increasingly losing understanding part: this has been my love & hate relationship with AI since day 1, and even after many months of failing, I still think it's a solvable problem, provided enough people actually want it to be solved. I know I do, but, I wonder if I am a dying breed?
61
Replying to @pidotdev @mitsuhiko
where is the bathrobe
58
Replying to @pidotdev @mitsuhiko
same here. stopped using Astra a week ago and work immediately improved. Maintaining a consistent level of service and model behavior is like the ground shift under our feet. Always back and forth from highly productive to a crawl for reasons unrelated to my workflow.
1
123
Replying to @pidotdev @mitsuhiko
"so many times [working with coding agents] is a magical experience. but they're also deeply frustrating. ...sometimes they produce an amazing result, and then [sometimes] you're spending two weeks trying to massage one pull-request into a shape that you like, with no feeling of forward progress. and it seems to be random" 100% this
48