← Back to all news
Hugging Face Trending Papers July 28, 2026

dtControl2+$\varepsilon$: Trading Optimality for Explainability in MDPs via Decision Trees

Read the original on Hugging Face Trending Papers →

Over the past decade, decision trees have been used to represent controllers (a. k.

Summary generated by The Flow from the publisher's feed. The full article lives at Hugging Face Trending Papers.

  • reinforcement-learning
  • benchmarks

Related stories

arXiv AI
Jul 29

dtControl2+$\varepsilon$: Trading Optimality for Explainability in MDPs via Decision Trees

arXiv:2607. 25925v1 Announce Type: new Abstract: Over the past decade, decision trees have been used to represent controllers (a.

By Tereza Kinsk\'a, Jan K\v{r}et\'insk\'y, Tobias Meggendorfer, Sabine Rieder, Maximilian Weininger
reinforcement-learningbenchmarks
More like this →
arXiv Machine Learning
Jun 4

Explainably Safe Reinforcement Learning

arXiv:2606. 04634v1 Announce Type: new Abstract: Trust in a decision-making system requires both safety guarantees and the ability to interpret and understand its behavior.

By Sabine Rieder, Stefan Pranger, Debraj Chakraborty, Jan K\v{r}et\'insk\'y, Bettina K\"onighofer
reinforcement-learningsafety
More like this →
arXiv AI
Aug 10

Interpretable reinforcement learning with decision-tree pruning

arXiv:2608. 07151v1 Announce Type: cross Abstract: Reinforcement learning policies are difficult to inspect, but interpreting them is a prerequisite for trustworthiness.

By Mark Leon Ringer, Michel Tokic
reinforcement-learningefficiencybenchmarkssafety
More like this →
arXiv Machine Learning
Aug 3

Beyond Black-Box Advice: Learning-Augmented Algorithms for MDPs with Q-Value Predictions

arXiv:2307. 10524v3 Announce Type: replace Abstract: We study the tradeoff between consistency and robustness in the context of a single-trajectory time-varying Markov Decision Process (MDP) with untrusted machine-learned advice.

By Tongxin Li, Yiheng Lin, Shaolei Ren, Adam Wierman
reinforcement-learning
More like this →
arXiv AI
Jun 10

Bellman-Taylor Score Decoding for Markov Decision Processes with State-Dependent Feasible Action Sets

arXiv:2606. 10979v1 Announce Type: new Abstract: Many Markov decision processes (MDPs) in operations research have feasible actions that are state dependent and defined implicitly by various operational constraints.

By Yi Chen (Lucy), Rushuai Yang (Lucy), Qiang Chen (Lucy), Dongyan (Lucy), Huo
reinforcement-learningbenchmarks
More like this →
arXiv AI
Aug 12

SPOTting the Future: Lookahead Explanations for Deep Reinforcement Learning

arXiv:2608. 09967v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) agents achieve strong performance in complex environments, yet their decision-making processes remain difficult to interpret.

By Tamar Gozlan, Claudia V. Goldman
agentsreinforcement-learning
More like this →