← Back to all news
Hugging Face Trending Papers October 6, 2026

Align, Then Correct: Training-Free Two-Stage Low-Rank Compensation for Extremely Quantized Large Language Models

Read the original on Hugging Face Trending Papers →

The Flow has not summarised this story yet — read it at Hugging Face Trending Papers.

  • llms
  • efficiency

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
1d ago

Align, Then Correct: Training-Free Two-Stage Low-Rank Compensation for Extremely Quantized Large Language Models

arXiv:2610.08164v1 Announce Type: new Abstract: Low-rank quantization error compensation (LQEC) recovers the accuracy lost under aggressive weight quantization by attaching a closed-form rank-$r$ ada...

By Seobin Song, Geonho Lee, Janghwan Lee, Jungwook Choi
llmsefficiency
More like this →
arXiv Machine Learning
Jun 2

ProjQ: Project-and-Quantize for Adapter-Aware LLM Compression

arXiv:2606. 00494v1 Announce Type: new Abstract: Post-Training Quantization (PTQ) and Low-Rank Adaptation (LoRA) constitute the standard pipeline for efficient Large Language Model (LLM) deployment.

By Wneya Yu, Chao Zhang, Li Wang, Samson Lasaulce, Merouane Debbah
llmsfine-tuningefficiency
More like this →
Hugging Face Trending Papers
Aug 14

QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction

As large language model inference shifts toward lower precision, post-training quantization (PTQ) becomes increasingly brittle, making quantization-aware training (QAT) essential for preserving model...

llmsefficiency
More like this →
arXiv Machine Learning
Aug 17

QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction

arXiv:2608. 13966v1 Announce Type: new Abstract: As large language model inference shifts toward lower precision, post-training quantization (PTQ) becomes increasingly brittle, making quantization-aware training (QAT) essential for preserving model quality.

By Vincent Counathe, Ben Athiwaratkun, Christopher De Sa, Tianyi Zhang
llmsefficiency
More like this →
arXiv AI
Aug 17

QuaSAR: Quantization Compensation via Stable Activation-Aware Rank Truncation

arXiv:2608. 14149v1 Announce Type: new Abstract: Recent training-free post-training quantization methods restore model accuracy through closed-form residual compensation.

By Lin-Fa Lee, Yi-Yu Chang, Kuo-Hei Yeh
fine-tuningefficiency
More like this →
arXiv Machine Learning
Jun 2

GPTQ-intrinsic LoRA: A Near-optimal Algorithm for Low-precision Quantization with Low-rank Adaptation

arXiv:2606. 01412v1 Announce Type: new Abstract: Post-training quantization is widely used for compressing large neural networks, but aggressive low-bit quantization can significantly degrade model quality.

By Shihao Zhang, Rayan Saab
llmsfine-tuningefficiency
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea