Video of tensorfold running Qwen3.8-Flash-Next-MLX-oQ8-MTP high on M3 Ultra ~80 t/s! It's a great experience, less errors and better quality. Where possible I suggest to use the highest possible quants!
18
3
80
7,479
Replying to @ivanfioravanti
I am going to start optimising TF for 8bit

Oct 3, 2026 · 10:15 AM UTC

4
24
617
Sort replies: Relevant Recent Liked
Replying to @ashxhart
AMAZING!
2
295
Nice! Just started on issue #188 this morning! Testing with some 8 bit Qwen models.
3
LFG!!
1
1
8
Freeing the machine for you!
1
24
Many thanks!
13