github.comMarch 31, 2026GitHub - ubergarm/llama.cpp at ug/port-sweep-benchLLM inference in C/C++. Contribute to ubergarm/llama.cpp development by creating an account on GitHub.
github.comMay 21, 2026GitHub - AtomicBot-ai/atomic-llama-cpp-turboquant: llama.cpp fork with TurboQuant WHT-rotated KV cache & weight compression + Gemma 4 MTP and Qwen 3.6 NextN speculative decoding (+30-50% throughput).llama.cpp fork with TurboQuant WHT-rotated KV cache & weight compression + Gemma 4 MTP and Qwen 3.6 NextN speculative decoding (+30-50% throughput). -...
github.comAugust 30, 2026GitHub - quimmedes/cafe-llama.cpp: A experimental fork of llama.cpp with LLM inference in C/C++A experimental fork of llama.cpp with LLM inference in C/C++ - quimmedes/cafe-llama.cpp