Weight serialization format
A standardized single-file container that stores model weights (often quantized) with metadata so they load safely, portably, and memory-mappably for local inference.
grounded in: Latest-trends theme: community 'fixated on squeezing large LLMs onto old CPUs and efficient weight storage rather than renting GPUs' (HN front page, 2026-07-16); durable formats such as safetensors/GG
Connected concepts
Memory-mapped weights, Quantization, Open-weight LLMs, Model offloading, CPU inference serving
Explore it live in the knowledge graph →