arXiv AI

FineSID: Scalable and Efficient Semantic Identifier Learning for Generative Recommendation

FineSID introduces a new quantization framework for semantic identifier learning in generative recommendation systems. By replacing the traditional Top‑1 hard assignment with a soft, differentiable approach, it distributes gradient updates across all codewords, leading to balanced codebook optimization and reduced identifier collisions. Experiments on public benchmarks show that FineSID improves codebook utilization and recommendation accuracy without relying on complex initialization strategies.

arXiv AI
Aug 28

Refine-POI: Reinforcement Fine-Tuned Large Language Models for Next Point-of-Interest Recommendation

Refine-POI introduces a reinforcement fine-tuned framework for next point-of-interest recommendation that tackles two key issues: topology-blind indexing of semantic IDs and the limitation of supervised fine-tuning to top‑1 predictions. It uses a hierarchical self‑organizing map to generate topology‑aware semantic IDs and a policy‑gradient approach to produce top‑k recommendation lists. Experiments on three real‑world datasets show that Refine‑POI outperforms state‑of‑the‑art baselines, combining LLM reasoning with accurate, explainable recommendations.

By Peibo Li, Shuang Ao, Hao Xue, Yang Song, Maarten de Rijke, Johan Barth\'elemy, Tomasz Bednarz, Flora D. Salim
arXiv Machine Learning
Aug 24

From a Static Multi-Level Small Semantic Codebook to a Dynamic Single-Level Large Semantic Codebook for Generative Recommendation

The paper proposes replacing a static multi‑level small semantic codebook with a dynamic single‑level large semantic codebook for generative recommendation. It introduces an exposure‑aware update mechanism and an offline evaluation framework, achieving significant improvements in recall, NDCG, decoding efficiency, and online consumption metrics on public datasets and production traffic.

By Tianlu Xie, Xin Ku, Mingjie Sun, Yunhao Sha, Lixiang Wang, Peng Wang, Yiyu Wang, Wenjin Wu, Zhaojie Liu, Peng Jiang, Wenwu Ou
arXiv AI
Jun 26

The Best of the Two Worlds: Harmonizing Semantic and Hash IDs for Sequential Recommendation

arXiv:2512. 10388v3 Announce Type: replace-cross Abstract: Conventional Sequential Recommender Systems (SRS) typically assign unique hash IDs (HID) to construct item embeddings, which mainly capture collaborative signals from historical user-item interactions.

By Ziwei Liu, Yejing Wang, Wanyu Wang, Wang Zejian, Qidong Liu, Zijian Zhang, Chong Chen, Wei Huang, Xiangyu Zhao
arXiv Machine Learning
Aug 11

Preserving Item Semantics for Free: Rethinking Token Initialization in LLM-Based Generative Recommendation

arXiv:2608. 07816v1 Announce Type: cross Abstract: Recent advances in generative recommendation (GR) leverage large language models (LLMs) as recommender backbones, enabling LLMs to directly generate recommendations conditioned on item-interaction histories.

By Donald Loveland, Liam Collins, Bhuvesh Kumar, Danai Koutra, Neil Shah