arXiv AI By Mark Walsh

Support sufficiency as action-sufficient compression: a single-cycle rate-regret formulation

Read the original on arXiv AI →

arXiv:2606. 09858v1 Announce Type: cross Abstract: Robust decision-making requires compression.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Machine Learning
Aug 14

Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry

arXiv:2608. 12753v1 Announce Type: new Abstract: We study decentralized multi-player reinforcement learning in episodic tabular Markov decision processes (MDPs) under three forms of information asymmetry: (A) unobserved actions with common rewards, (B) observed actions with independent rewards, and (C) unobserved actions with independent rewards.

By Larissa Xu, King Bi, William Chang
arXiv Machine Learning
4d ago

SCAMP: Sparse-anchor Control is One Small Projection

SCAMP introduces a training‑free, damped Gauss‑Newton method that adjusts only the sparse anchor points in a frozen differentiable decoder, keeping the rest of the state unchanged. By operating solely in the space of the anchors’ Jacobian rows, it solves a system whose size matches the number of anchor constraints rather than the full state, enabling efficient control across diverse text‑to‑motion generators. Applied to seven existing generators, SCAMP achieves anchor errors that match or surpass all released control methods and can close anchors on hosts that originally lacked them.

By Pengcheng Fang, Tengjiao Sun, Xiaoyu Zhan, Yanwen Guo, Hansung Kim, Xiaohao Cai, Dongjie Fu