← Back to all news
Hugging Face Trending Papers September 21, 2026

Reinforcement Learning under State and Outcome Uncertainty: A Foundational Distributional Perspective

Read the original on Hugging Face Trending Papers →

The Flow has not summarised this story yet — read it at Hugging Face Trending Papers.

  • agents
  • reinforcement-learning

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
Sep 22

Reinforcement Learning under State and Outcome Uncertainty: A Foundational Distributional Perspective

arXiv:2609.24103v1 Announce Type: new Abstract: In many real-world planning tasks, agents must tackle uncertainty about the environment's state and variability in the outcomes of any chosen policy. W...

By Larry Preuett, Qiuyi Zhang, Muhammad Aurangzeb Ahmad
agentsreinforcement-learning
More like this →
arXiv Machine Learning
Aug 4

Analytic Planning under Uncertainty with Moment Closure

arXiv:2608. 02519v1 Announce Type: new Abstract: Effective model-based reinforcement learning in stochastic environments requires planning that accounts for predictive uncertainty.

By Shishir Sharma, Doina Precup
reinforcement-learning
More like this →
Hugging Face Trending Papers
Aug 3

Analytic Planning under Uncertainty with Moment Closure

Effective model-based reinforcement learning in stochastic environments requires planning that accounts for predictive uncertainty. Propagating full state distributions analytically offers a principled way to do this, but has traditionally required restrictive policy or reward structures to remain tractable.

reinforcement-learning
More like this →
arXiv AI
Jun 11

Planning under Distribution Shifts with Causal POMDPs

arXiv:2602. 23545v2 Announce Type: replace Abstract: In the real world, planning is often challenged by distribution shifts.

By Matteo Ceriscioli, Karthika Mohan
reinforcement-learning
More like this →
arXiv AI
Jul 1

Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity

arXiv:2602. 03778v2 Announce Type: replace-cross Abstract: Tail-end risk measures such as static conditional value-at-risk (CVaR) are used in safety-critical applications to prevent rare, yet catastrophic events.

By Aneri Muni, Vincent Taboga, Esther Derman, Pierre-Luc Bacon, Erick Delage
reinforcement-learningsafety
More like this →
arXiv AI
Jul 14

Efficient Q-Learning and Actor-Critic Methods for Robust Average-Reward Reinforcement Learning

arXiv:2506. 07040v4 Announce Type: replace-cross Abstract: We study model-free methods for distributionally robust infinite-horizon average-reward Markov decision processes (MDPs).

By Yang Xu, Swetha Ganesh, Vaneet Aggarwal
reinforcement-learning
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea