Prev engineer at faang, cto of vc-backed startup. Building: shortcutchat.com - one prompt, multiple AI answers listingquick.com - chrome ext. listing manager

Utah, USA
Pinned Tweet
Replying to @sundarpichai
5
29
953
54,754
There are roughly 1.7 million accountants in the U.S. according to the U.S. Bureau of Labor Statistics bls.gov/cps/cpsaat11b.htm Hard to imagine a scenario where that number doesn't shrink. What do they pivot to in the near term? Especially when similar white-collar professions start being outperformed by AI as well.
New research: we hired 12 CPAs to complete four realistic accounting tasks, then had AI models from the past few years complete the same work. As of May 2025, accountants outperformed AI. Now they lose out to today's frontier models, which are near perfect on these tasks. 🧵
1
59
"The future is already here – it's just not evenly distributed"
We live in a 𝕏 bubble. Only 2% of us households pay for AI. Comparatively: - 25% of households pay for SiriusXM - 55% of pay for cloud storage - 91% pay for at least one streaming service You are so early
1
30
Google after dropping Gemini 4 Argon
Introducing Gemini 4 Argon – our new frontier model. It’s built for complex workflows across coding, enterprise knowledge work, and cybersecurity defense – rolling out today to a set of trusted testers through our Fairwind Program.
1
157
Trying to keep up with all the new model releases
you're not crazy between anthropic and openai, a new model used to come out every ~10 weeks now it's every ~11 days
2
131
The machines really are made in the image of man 😆
Ask an AI model to pick the better answer, and it'll usually pick its own. We compared 34,580 verdicts from 12 models with human votes across 1,460 battles on Arena. The results show that AI judges have their own taste, and it's unlike ours. They: - Favor their own answers. On average, a model picked its own answer 58% of the time. People picked that same answer 34% of the time. GPT-6 Astra picked itself 88% of the time. - Rarely call a draw. People called a tie or "both bad" in 32% of battles. GPT-5.6 Sol picked a winner 96% of the time. - Side with each other over people. They agreed with other AIs 79% of the time and with people 57% of the time. Every judge did, by a margin of 18 to 27 percentage points. Full results in the article from @DawidGalarowicz below.
2
133
Hard to argue with sadly
everything is fake - posts on here, ghostwritten - statistics, misleading - benchmarks, bullshit - reviews, paid for nearly every company is doing this. the frenzy over ai is showing me so few founders have any backbone
69
Sonnet 5 came out on June 30th of this year. The rate of progress is getting scary.
A thread of early experiments with Claude Sonnet 5.5. A fall foliage simulator by @_re_pete, made with Sonnet 5 vs Sonnet 5.5.
92
Claude Sonnet 5.5 vs Claude Opus 5.5: - Sonnet 5.5 is faster and cheaper - Opus 5.5 still performs better on most benchmarks
Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.
119
Another day, another new frontier model

ALT Animated GIF

Claude Sonnet 5.5 is now available:
56

ALT michael fassbender perfection GIF

most companies adding AI to their products:
1
155
Looks like software engineers aren't out of a job quite yet. Still need humans to help fix other humans' bugs 😁
We are aware that codex is down and are working hard to bring back normal service.
1
150
AI is taking libraries like three.js to the moon
Claude Opus 5.5 has been out for a few days. Some of our favorite things people have explored and discovered with it so far:
1
154
Surprising how different it looks when you order by token volume instead of token spend. A lot of folks seem to like @deepseek_ai
Spend of OpenAI vs Anthropic vs Open (Vercel AI Gateway, last 2 months) • Anthropic still #1 in spend, but went 69% → 40% • OpenAI: 10% → 24% in spend • GPT-6 Astra + GPT 5.6 Sol are ripping • OpenAI now leads in tokens # • Kimi K3 + DeepSeek took ~half of Anthropic's loss • Opus 5.5 is up to 10% of spend in 2 days • OpenAI is 62% of image generations Watch here: token-race.vercel.app
114
Crazy to see Claude Fable 5.1 this far down the list on @zapier's AutomationBench after all the releases yesterday. Things are moving fast.
1
4
145
At the very least, GPT-6 Astra, Sol, Luna make it easier to tell which model is which on a chart 😁
Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
1
226