MDRSS · MARKDOWN SNAPSHOT
0.0/10

buun-llama-cpp

A research and development fork of llama.cpp, providing unique KV cache codecs, inference techniques, and bleeding edge features. Why pay 3-bit or 4-bit quality for a context length you only sometimes reach?

#llm-engineering / card #1231★ 0◌ 0snapshot 2026-08-04