SIGN IN SIGN UP

Support Bamboo LM sparse and dense inference (#171)

* support inference for sparse mistral arch

* patch convert script for the new model type

* fix: runtime constants for new model kind

* renaming: bamboo arch

* fix: model type

* bugfix: wrong arch when converting bamboo HF models

* chore: remove debugging info

* add script for dense model conversion (for Bamboo LM)

* support dense inference of Bamboo LM

* Add news in README.md

* Add readme for BambooLM and dense inference

* update todos
H
Holden X committed
b25cb86b0ba9b72af8bf6900a869f7238e876e1d
Parent: 1374ab0
Committed by GitHub <noreply@github.com> on 3/28/2024, 8:53:03 AM