Popular repositories Loading
-
qwen38-mtp
qwen38-mtp PublicOne llama.cpp flag unlocks +33-39% decode speed for Qwen3.8-27B on consumer GPUs. The MTP head already ships inside your GGUF. Recipe, paired benchmarks, probe tool.
-
octopus-invaders
octopus-invaders PublicA pixel art space shooter built entirely by a 9B AI model on a single RTX 3060. Zero hand-written code.
-
hermes-agent
hermes-agent PublicForked from NousResearch/hermes-agent
The agent that grows with you
-
bonsai2-small-gpu
bonsai2-small-gpu Publicrun ternary bonsai 2 27b well on the gpus people own: serve lines per vram tier, a 1.5x decode kernel for the prismml fork, the qwen 3.8 mtp head grafted back, sweeps by pr
-
dgx-spark-ling
dgx-spark-ling PublicServe the official Ling-3.0-flash INT4 on one NVIDIA DGX Spark, 38.7 tok/s, faster than the community GGUF. Recipe, watchdog, benchmarks.
-
nopasswd-sudo
nopasswd-sudo PublicTime-boxed passwordless sudo for agents. Grants expire on a systemd timer, validated with visudo before install, audited with sudoreplay.
Shell 26
If the problem persists, check the GitHub status page or contact support.


