arXiv Machine Learning
Jul 8

Strategic Bargaining in Multi-Buyer Markets: Reinforcement Learning from Verifiable Rewards for LLM Negotiations

arXiv:2607. 05863v1 Announce Type: new Abstract: Negotiation is a fundamental strategic interaction in management science, characterized by agents attempting to reach agreements while protecting private information, such as reservation costs and hidden valuations.

By Shuze Daniel Liu, Claire Chen, Jiabao Sean Xiao, Xin Chen, David Simchi-Levi
arXiv AI
Jun 16

STRIDE: Strategic Trajectory Reasoning via Discriminative Estimation for Verifiable Reinforcement Learning

arXiv:2606. 15866v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become an effective post-training paradigm for improving the reasoning abilities of large language models.

By Qinjian Zhao, Zhihao Dou, Dinggen Zhang, Xiangyu Li, Chaoda Song, Zhongwei Wan, Xinpeng Li, Yanyan Zhang, Kaijie Chen, Qingtao Pan, Chengcheng Feng, Zhiqiang Gao, Xiaoyu Xia
arXiv AI
Sep 4

Data Market Design through Deep Learning

The paper tackles the data market design problem, which seeks signaling schemes that maximize revenue for an information seller. It applies deep learning to learn these schemes, addressing both obedience and incentive constraints, and demonstrates that the framework can replicate known theoretical solutions, extend to more complex scenarios, and suggest new optimal designs. The study builds on prior auction‑design work and introduces a novel approach for revenue‑optimal data markets.

By Sai Srivatsa Ravindranath, Yanchen Jiang, David C. Parkes