GGUF
GGUF is a model file format for loading and running large language models with llama.cpp.
1 article
Last mentioned GGUF was introduced by the llama.cpp developers in 2023 to replace the earlier GGML format. A single file holds a model's weights, tokenizer and configuration. It supports many quantization types and is widely used to run local LLMs on personal computers.
This entry is based on AIPOST articles and widely known facts. If something is wrong, please send us a correction request.