Now that the AI spam has died down, the bounty program is back! We have four open bounties. If you want a job here, this is the way. If you are only willing to put in an AI amount of effort, please don't, it won't work.
TinyStories-15M on a $60 Google Coral: all 6 transformer layers on the chip, 176 tok/s at batch 1.
An AI agent got the chip and zero docs. It wrote the driver, decoded the instruction set and built a @__tinygrad__ backend. tinygrad is small enough for an agent to understand.
Does some university want to host me (George) and offer a winter semester Compilers for Machine Learning course?
12+ weeks, hands on, build a compiler. Syllabus in next tweet. Needs a student body capable of doing this work and leadership capable of recognizing how SOTA this is.
Syllabus here. Would be for winter 2027. On odd weeks the students will write tests, even weeks they will make their compiler pass them. A few written tests too, no skating by with AI. Similar to CMU's 15-411, one of my favorite classes of all time. gist.github.com/geohot/47685…
It's tinygrad hosted Qwen3.8-27B in Pi Coding Agent playing Pokémon Red on a tinybox red by looking at the screen and pressing buttons. Will leave it running all night, how far do you think it will get?
Here's a trace of AMD's AITER gemm on MI350X. There's tons of MFU wasted during the non overlapped global memory writeback. The top 4 rows are the MFMA engines, green means running, black means idle.
As far as I know, Pokemon Red still hasn't been beaten from pixels! Trying MiMo-V2.6-Pro, it's so nice to have unlimited local tokens for dumb experiments.
Harness here. Astra is very good out of the box. Other models not so much. Playing with local tinygrad hosted Qwen3.8, maybe try a finetune. github.com/geohot/pi_plays_p…
We have MLPerf qualifying gpt-oss-20b training in 119 minutes on MI350X. tinygrad is likely now the fastest LLM training framework in the world. @nvidia want to give us a contract to beat your published MLPerf times? I'm confident we can do it.
Update: they want a meeting, I offered an e-mail. When you think about it, $2M for a complete alternative stack to CUDA is such an absurdly good deal. Who thinks they'll accept?
Apparently all of these “AI is escaping containment” stories are actually just AI researchers not knowing jack squat about the absolutely most basic security practices.
.@cerebras should add pricing and a buy it now button for the CS-4. I don't want a stupid cloud. If they add buy it now for their hardware, I'm bullish on the company. If they don't, it's clear they aren't close to price competitive.