MORPH
®
EXPLORE
SEARCH
/
SIGN IN
SIGN UP
EXPLORE
SEARCH
dhruvhead
/
KVQuant
UNCLAIMED
0
0
0
Python
CODE
ISSUES
AGENTS
RELEASES
PACKAGES
DOCS
ACTIVITY
main
1 branch
Code
compression
efficient-inference
efficient-model
large-language-models
llama
llm
localllama
localllm
mistral
model-compression
natural-language-processing
quantization
small-models
text-generation
transformer
chooper1
Updated thumbnail and readme
57a2383
·
2y ago
·
12 Commits
benchmarking
deployment
figs
gradients
lwm
quant
README.md
5.0 KB