Building software & AI | Tech Lead & Advisor | Tinkering with 3D printing, electronics & hardware

France, Thailand, World
Feeling rewarded to have “kicked ass” of my agents today, reviewing a very slow chat app, breaking down every step of it and improving it 15x one step at a time. Hand coding is dead, long live to coding.
25
Stephane Busso retweeted
Cool eval. Simply ask an LLM “Land or Water?” and give it a latitude and longitude coordinate as text. Ask 16,200 times, plot as image. The models know. From compressing the internet.
Replying to @celestepoasts
results for all claudes
465
919
16,654
1,034,619
Stephane Busso retweeted
Discover the nicest easter egg in Pi 1.0 online using Wasmer! wasmer.sh/?example=pi
1
6
43
3,674
Stephane Busso retweeted
all software is becoming malleable, so it was important for us to support this as a first class citizen in Claude Code, I hope more and more software becomes extensible like this! here are some of what I'm most excited about with mods:
You can now mod Claude Code: - Change how it behaves - Customize the UI - Swap in your own features Write one with a few lines of TypeScript, or have Claude build it for you. Mods ship inside plugins, so you install them with /plugin in the CLI or desktop app. A few examples:
140
110
2,782
363,046
Stephane Busso retweeted
MLX-Serve 26.10.1 is out: speed across the board. Exactly faster. Qwen3.8 27B with its drafter +66% on an M5 Ultra, +28% on an M4 Max, +37% on an M1 Pro. Qwen Flash Next: +11% decode on M4 Max, +51% prefill on M5 Ultra. (up to ~5400 PP!!!!) Gemma 4, Qwen, LFM2.5, Spark, Bonsai 2 all faster, output byte-identical to 26.9.6 (across 18 models). New: * GGUF models run on our own MLX engine, written in Zig and Metal, instead of the llama.cpp fallback. (experimental) * 🍣 Sushi-format Flash Next packs run directly. * Qwen-Image 2.1 edits pictures from instructions. * Drafting is on by default for every model that supports it. Huge thanks to everyone who filed, tested and fixed this one: @STRML_, @sbusso, @wutang_superfan , lborloz, LXD-8 @alinselea zeeshanhaque21 @Lojza3D, h9q2cyxvgm-ui @kennethrdegraff, @CowboyCoderHQ , codysk, @Beamsters1, @CerebralCoding_, @AjAbsaki , and more.. THANKS ! Please comment if I did not include your name, for visibility, I know people mostly by their GH handle. Special thanks to @ashxhart and @Spangler3000 for pushing ! Benchmarked and tested across 5 machines: htmlhost.jax.workers.dev/ren… GH Release Download + Changelog: github.com/ddalcu/mlx-serve/… Website: mlxserve.com
45
42
387
28,740
Stephane Busso retweeted
Claude desktop app pro tip: The advisor functionality is now supported. This means you can use sonnet 5.5 for balanced speed and intelligence, and when it's facing a tough decision, it can call Opus for help. Probably an under rated combo that will both speed things up and save your limits.
8
7
95
8,688
Programming Languages are a Human Interface made to communicate instructions to a computer. They are sub efficient as every layer is adding an abstraction. Coding with Agents should fundamentally remove some layers and communicate closer to the metal.
1
4
117
Those numbers are getting pretty insane for local AI. It’s amazing how much performance has been gained in just a few weeks, and given the speed of progress, we may be in for a long run.
2
85
Improvement on Nemotron 3.5 Lightning 30B-A3B with mlx-serve, M5 Max: • decode 37 → 205 tok/s (245 with MTP) • prompt reading 420 → ~3,900 tok/s (9x) • 16k-token prompt: first token 41s → 9s That starts to be serious numbers for a local AI 🎉
1
1
106
Lightning is not an everyday chat model. It is a fast, small-active (3B of 30B) base for specific tasks: fine-tune it for one job (extraction, routing, tool calls, a domain like finance) and it runs that job at 200+ tok/s on a Mac.
34
Stephane Busso retweeted
Built by Opus 5.5 (Ultracode) Created by its own creativity. No images, no video, no audio files. Every frame and every note is generated by code, live, from the word you type. Live here: mia-ai.net/experiments/let-t… How it works and the full prompt below.
85
74
1,120
84,258
Stephane Busso retweeted
26.9.6 is out ! github.com/ddalcu/mlx-serve/… * M5 Ultra optimizations (some apply to all M class) * Laya typed decisions (local-JEV!) * Fixed handling of media and preserve_thinking * UI Updates * Renamed to MLX-Serve (no more MLX Core) * Bonsai-2 Optimizations Sound On 🤘
20
22
269
56,570
Stephane Busso retweeted
Had an "interview" for a blockchain project last week. Camera was on, we're chatting, and the guy tells me to clone a GitHub repo and run it locally before we go further into the technical round. I said sure, but first can you do me a favor hold up 3 fingers in front of your face for me real quick. He froze. Didn't move. Just sat there for a few seconds before the call cut off and he blocked me. That's when I knew. A real interviewer doesn't glitch out over a random ask like that. A deepfake/AI overlay does. These "run this repo" scams are getting scary common in crypto and dev hiring right now. The setup is always the same: - flattering DM - real-sounding project - rushed timeline - a "quick step" before the call that's really just remote access or a credential stealer in disguise. If someone wants you to run code or install something before you've even had a real conversation, that's the whole scam. Trust the instinct. Stay safe out there.
325
3,515
19,472
1,802,086
.@DDH has dropped a contrarian keynote at Rails Conf, but he made a point, as he often does. We like it or not, we are all becoming project managers and makers, leaving programming to agents. Coding by hand is over. Embrace and enjoy the change, and adjust how we work.
2
75
I got what you meant, @ddalcu, about the iPhone 18 Pro being a beast, and mlx-serve is really taking great advantage of it. It is amazing what the app is shipping on a phone ... My only complaint about the app is that it is easy to mistake it for X :)
1
3
226
Stephane Busso retweeted
shocking, who would've guessed
JUST IN: OpenAI and Anthropic reportedly exaggerated AI security threats to push the government to protect their market position
20
47
753
37,136
Stephane Busso retweeted
An OpenAI co-founder spent 3 years quietly building Jev, only for someone to open source a better performing model in just 3 days. ​This is exactly why starting an AI company right now carries way too much risk.
Laya é um open source do Jev que já veio acima dele. tá cada vez mais rápido huggingface.co/convaiinnovat…
125
183
2,513
308,590
Stephane Busso retweeted
I built a Photoshop-like image editor. It’s called Compositor. robbietilton.com/compositor I originally built it for myself to get off my Adobe subscription, but decided to release it for free and make it open source. It has all the essential tools I need for compositing, with none of the BS. I know this workflow is kind of archaic in 2026, but I’m still using it until AI can actually get pixel-perfect on some of the details. The entire app is 12MB. Photoshop is 6,455MB on my machine.
1,088
2,110
19,351
2,275,228
There are 2 kinds of people: those who see Jev as the future and those who don't understand why there is so much hype.
1
67
Stephane Busso retweeted
What happens to the 10Y once the Fed starts hiking? We modeled every tightening cycle since 1963. With rates on the move like this we looked at what the 10Y does around the start of a hiking cycle. The chart shows the average across every cycle since the early 1960s, with the range of the individual cycles behind it. A year after the first hike the 10Y is up around 100bp on average. More telling, only one cycle, 2004, saw yields fall at any point in that first year. On the historical odds, yields are more likely than not to rise over the coming year.
1
2
94