Views provided by UsageCounts
A complete redesign - non-backward compatible. Enabling multi-agent support. New features - PIP package Benchmarks Hierarchical Reinforcement Learning (demonstrated by Hierarchical Actor-Critic) Tutorials Shared memory (e.g. Replay Buffer) between workers Tests (unit-tests, reward-based tests, trace-based tests) Using Coach as a library (see example here) New Environments - Toy Environments (Exploration Chain, BitFlip) DeepMind PySC2 support (Starcraft 2) DeepMind Control Suite New Algorithms - Hindsight Experience Replay Prioritized Experience Replay Hierarchical Actor-Critic UCB with Q-Ensembles
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
| views | 5 |

Views provided by UsageCounts