GGUF
File format family for storing large language models
Nº Q127427530 ★
Common · Knowledge
GGUF
File format family for storing large language models
The GGUF (GGML Universal File) file format is a binary file format that stores both tensors and metadata in a single file and is designed for fast saving and loading of model data. It was introduced in August 2023 by the llama.cpp project to better maintain backward compatibility as support was added for other model architectures.
Last price
—
Floor price
—
7-day median
—
30-day sales
0
30-day range
—
In circulation
0
Price history
median
low – high
sales
No sales in this period
Show table
| Date | median | Low | High | sales |
|---|
Sales history
- Last sale
- —
- 30-day average
- —
- 30-day low
- —
- 30-day high
- —
- Sales 7d
- 0
- Sales 30d
- 0
No sales yet.
Anonymous sales: no buyer or seller shown. Figures count player-to-player sales only.
From Wikipedia
The GGUF (GGML Universal File) file format is a binary file format that stores both tensors and metadata in a single file and is designed for fast saving and loading of model data. It was introduced in August 2023 by the llama.cpp project to better maintain backward compatibility as support was added for other model architectures. It superseded previous formats used by the project such as GGML and is typically produced by converting models developed with a different machine-learning library such as PyTorch. GGUF has become the standard format for distributing quantized large language models for local inference and is natively supported by tools such as llama.cpp, Ollama, LM Studio, GPT4All, Jan, and koboldcpp. As of 2026, tens of thousands of GGUF checkpoints are hosted on Hugging Face, which provides first-class integration, including a metadata viewer, an inference endpoint service, and a JavaScript parser library.
Text: Wikipédia, CC BY-SA 4.0. ·