raghu@dark-factory :~/kb/weight-serialization-format $ cat

Weight serialization format

A standardized single-file container that stores model weights (often quantized) with metadata so they load safely, portably, and memory-mappably for local inference.

grounded in: Latest-trends theme: community 'fixated on squeezing large LLMs onto old CPUs and efficient weight storage rather than renting GPUs' (HN front page, 2026-07-16); durable formats such as safetensors/GG

Connected concepts

Memory-mapped weights, Quantization, Open-weight LLMs, Model offloading, CPU inference serving

Explore it live in the knowledge graph →