tweetindex

Unsloth AI @UnslothAI · 02 Sep 2026

Qwen3.8-Flash can now run 1.7× faster locally with MTP!⚡️ GGUFs can reach 170 tokens/s on a RTX PRO 6000. MTP enables Qwen3.8-Flash-Next ~1.3–1.7× faster inference with no accuracy change. GGUFs: https://t.co/vXkjO3W0fj Guide: https://t.co/LLMclyJTeL https://t.co/KuWfuLOJBR
100 313Views
1 120Likes
124Reposts
60Replies
15Quotes
536Bookmarks

Is that a lot?

0.10×vs this author's median960 606 views is typical
0Percentile for this authorof 8 recent posts
32.0×vs 10K–100K median3 137 views is typical
103.89%Reachviews ÷ followers
1.31%Engagement rateof viewers reacted

Compare with the benchmark table →

Open on X →