Hugging Face Blog

Fast Inference on Large Language Models: BLOOMZ on Habana Gaudi2 Accelerator

arXiv AI
Jul 28

Bayesian-LoRA: Probabilistic Low-Rank Adaptation of Large Language Models

arXiv:2601. 21003v3 Announce Type: replace Abstract: Large Language Models usually put more emphasis on accuracy and therefore, will guess even when not certain about the prediction, which is especially severe when fine-tuned on small datasets due to the inherent tendency toward miscalibration.

By Moule Lin, Shuhao Guan, Andrea Patane, David Gregg, Goetz Botterweck