Engineering notes

Qwen3.8 hits 81.4 tokens/s on an RTX 3090

Local Qwen3.8-27B peaks at 81.4 tokens/s with a few perf adjustments.

Muse Glimmer is a flop

Muse Glimmer is a flop: chatty, prone to hallucinations, bad at tool calling, slower than Qwen and only half of Qwen 3.6's context.