← Back to all news
Hugging Face Blog September 10, 2020

Block Sparse Matrices for Smaller and Faster Language Models

Read the original on Hugging Face Blog →

The Flow has not summarised this story yet — read it at Hugging Face Blog.

  • llms

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

Hugging Face Blog
Oct 3, 2022

Very Large Language Models and How to Evaluate Them

llms
More like this →
OpenAI Blog
Jul 28, 2022

Efficient training of language models to fill in the middle

llms
More like this →
Hugging Face Blog
Oct 26, 2021

Large Language Models: A New Moore's Law?

llms
More like this →
arXiv AI
Jul 24

Break Through the Compression Bottleneck: From Theory to Practice

arXiv:2607. 20434v1 Announce Type: cross Abstract: As the parameter size of language models continues to grow, effective model compression is required to reduce their computational and memory overhead.

By Xiusheng Huang, Lu Wang, Yequan Wang, Jun Zhao, Kang Liu
llmsefficiency
More like this →
arXiv Machine Learning
Jun 2

Riemannian Gradient Descent for Low-Rank Architectures

arXiv:2606. 02328v1 Announce Type: new Abstract: We explore Riemannian optimization techniques for rank-factored matrix parameters, targeting contemporary deep learning applications.

By Nicholas Knight
llms
More like this →
OpenAI Blog
Jul 7, 2021

Evaluating large language models trained on code

llms
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.0.0 · bb4ee0e