GGUF

GGUF is a model file format for loading and running large language models with llama.cpp.

1 article
Last mentioned

GGUF was introduced by the llama.cpp developers in 2023 to replace the earlier GGML format. A single file holds a model's weights, tokenizer and configuration. It supports many quantization types and is widely used to run local LLMs on personal computers.

This entry is based on AIPOST articles and widely known facts. If something is wrong, please send us a correction request.

Articles covering this entry

The model is available on Hugging Face in GGUF, the file format used by the llama.cpp inference engine, in two builds:


© 2026 AIPOST. All rights reserved.

AIPOST is an AI publication covering practical AI, AI security, performance, startups, health, ethics and industry news. No account is needed, and our privacy policy explains how we handle personal information.