Engineering notes
Qwen3.8 hits 81.4 tokens/s on an RTX 3090
Local Qwen3.8-27B peaks at 81.4 tokens/s with a few perf adjustments.
Muse Glimmer is a flop
Muse Glimmer is a flop: chatty, prone to hallucinations, bad at tool calling, slower than Qwen and only half of Qwen 3.6's context.
Subscribe
New posts by email, roughly twice a month.
✓ Subscribed. Check your inbox to confirm.
RSS