nitter
Ivan Fioravanti
@ivanfioravanti
10h
Video of tensorfold running Qwen3.8-Flash-Next-MLX-oQ8-MTP high on M3 Ultra ~80 t/s! It's a great experience, less errors and better quality. Where possible I suggest to use the highest possible quants!
Enable hls playback
18
3
80
7,479
Ash Hart
@ashxhart
9h
Replying to
@ivanfioravanti
I am going to start optimising TF for 8bit
Oct 3, 2026 · 10:15 AM UTC
4
24
617
Sort replies:
Relevant
Recent
Liked
Ivan Fioravanti
@ivanfioravanti
8h
Replying to
@ashxhart
AMAZING!
2
295
BHCC Spark Lab
@BHCC2025
4h
Replying to
@ashxhart
@ivanfioravanti
Nice! Just started on issue #188 this morning! Testing with some 8 bit Qwen models.
3
Naz
@nothingto12
4h
Replying to
@ashxhart
@ivanfioravanti
LFG!!
1
1
8
Ivan Fioravanti
@ivanfioravanti
4h
Freeing the machine for you!
1
24
Antho G
@antho__g
5h
Replying to
@ashxhart
@ivanfioravanti
Many thanks!
13