
GitHub - QuinsZouls/llama-cpp-turboquant at llama-next
LLM inference in C/C++. Contribute to QuinsZouls/llama-cpp-turboquant development by creating an account on GitHub.
Developers and AI enthusiasts seeking efficient local LLM deployment on consumer-grade AMD and Apple hardware