遇见数据集

Data for: 3401653

收藏
Mendeley Data2019-06-10 更新2026-04-09 收录
官方服务:

资源简介:

The Courchevel environment is hereby published to ease the development of streaming machine learning algorithms. In first solving this problem, a rapid reinforcement learning algorithm was invented. A simple transformation is added to the Bellman equation, a principal pillar of AI, particularly for solving Markov Decision Problems. By adding stochasticity to Bellman, sustained Reward-Per-Episode gains of an order of magnitude are validated, for environments where the reward function is structurally anticipated to be multi-modal. Courchevel as a decision problem, a first solution, and the Biased Bellman innovation are revealed -- with accompanying data. For ease of discussion, Courchevel's dynamics are described in military terms.

创建时间:
2019-06-10
二维码
社区交流群
二维码
科研交流群
商业服务