引用本概念的论文(4)
- 分位数分布强化学习的统计效率与推断 Statistical Efficiency and Inference of Quantile Distributional Reinforcement Learning arXiv 2607.08444
- 审计分布式强化学习的风险声明 Auditing the Risk Claims of Distributional Reinforcement Learning arXiv 2607.11607
- Cramér 几何下的分布型软 Bellman 算子 Distributional Soft Bellman Operator under the Cramér Geometry arXiv 2607.17897
- 基于状态感知探索的双流强化学习 Dual-Flow Reinforcement Learning with State-Aware Exploration arXiv 2606.29820