Support Bamboo LM sparse and dense inference (#171)
* support inference for sparse mistral arch * patch convert script for the new model type * fix: runtime constants for new model kind * renaming: bamboo arch * fix: model type * bugfix: wrong arch when converting bamboo HF models * chore: remove debugging info * add script for dense model conversion (for Bamboo LM) * support dense inference of Bamboo LM * Add news in README.md * Add readme for BambooLM and dense inference * update todos
H
Holden X committed
b25cb86b0ba9b72af8bf6900a869f7238e876e1d
Parent: 1374ab0
Committed by GitHub <noreply@github.com>
on 3/28/2024, 8:53:03 AM