llama : use n_embd_gqa instead of n_embd to handle llama-2 70B (#2433)
R
Rand Xie committed
65cdf34bdc469fa86248e667a5880992684ef114
Parent: edcc7ae
Committed by GitHub <noreply@github.com>
on 7/28/2023, 8:42:53 AM