Github

代码库

One llama.cpp flag unlocks +33-39% decode speed for Qwen3.8-27B on consumer GPUs. The MTP head already ships inside your GGUF. Recipe, paired benchmarks, probe tool.
Python
Time-boxed passwordless sudo for agents. Grants expire on a systemd timer, validated with visudo before install, audited with sudoreplay.
Shell