← Back to all news
arXiv Computer Vision September 10, 2026 By Tanvir Muntakim Tonoy, Sajjad Ghiasvand, Mahnoosh Alizadeh, Ramtin Pedarsani

Low-Rank Prompt Learning for Vision-Language Models with Fixed-Token Bases

Read the original on arXiv Computer Vision →

The Flow has not summarised this story yet — read it at arXiv Computer Vision.

  • llms
  • rag
  • multimodal
  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv AI
Aug 11

Rethinking Factor Sharing in Federated LoRA: A Rank-Aware Adaptive Approach

arXiv:2608. 09742v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) represents large language model (LLM) updates with two compact matrix factors, i.

By Xinyi Xu, Bingnan Xiao, Shuang Qin, Gang Feng, Tony Q. S. Quek
llmsfine-tuning
More like this →
arXiv Machine Learning
Aug 4

Not the Dimension, the Norm: What Matters in Gradient-Free Weight Perturbation of Language Models

arXiv:2608. 01624v1 Announce Type: cross Abstract: Adapting a language model to a task no longer requires training all of its weights, and a line of parameter-efficient methods has driven the trainable count from billions down to a handful of scalars.

By Taeyeong Kim, Ahhyun Kim, TaeHyeon Kim, Unggi Lee
llmsbenchmarks
More like this →
arXiv Machine Learning
Sep 10

MpSub: A Momentum $p$-Dimensional Subspace Trust-Region Method for Derivative-Free Fine-Tuning of Large Language Models

arXiv:2609.07666v1 Announce Type: new Abstract: Full-parameter fine-tuning of large language models has substantial memory costs because backpropagation stores activations and gradients. Zeroth-order...

By Yuyang Wang, Haoyu Yao, Pengcheng Xie
llmsfine-tuning
More like this →
arXiv AI
Aug 13

Post-Training with Policy Gradients: Optimality and the Base Model Barrier

arXiv:2603. 06957v2 Announce Type: replace-cross Abstract: We study post-training linear autoregressive models with outcome and process rewards.

By Alireza Mousavi-Hosseini, Murat A. Erdogdu
reinforcement-learning
More like this →
arXiv Machine Learning
Jun 15

Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation

arXiv:2606. 13657v2 Announce Type: replace Abstract: On-policy distillation (\textsc{OPD}) has recently become a prominent post-training recipe by combining two desirable ingredients: on-policy student trajectories and dense teacher supervision.

By Guo Yu, Wenlin Liu, Yulan Hu, Hao-Xuan Ma, Jun-Peng Jiang, Han-Jia Ye
llmsefficiencymultimodalbenchmarks
More like this →
arXiv Machine Learning
Jun 29

NormGuard: Reward-Preserving Norm Constraints in Flow-Matching Reinforcement Learning

arXiv:2606. 27771v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training improves the reward alignment of flow-based generators, but often degrades perceptual quality in ways that are not captured by the reward proxy.

By Tianlin Pan, Lianyu Pang, Cheng Da, Huan Yang, Changqian Yu, Kun Gai, Wenhan Luo
reinforcement-learningfine-tuningsafety
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea