ExperienceBufferLength in Reinforcement Learning Toolbox

19 次查看（过去 30 天）

qun wang 2021-11-15

0
链接

此问题的直接链接

https://ww2.mathworks.cn/matlabcentral/answers/1587044-experiencebufferlength-in-reinforcement-learning-toolbox

评论： Francisco Serra 2024-5-2

Hello, everyone,

I found a problem with the 'ExperienceBufferLength' property in 'rlDDPGAgentOptions' when specifying options for rl agents.

Usually this property is set as 1e6 in the examples of the Help documentation, such as here.

In this example, every episode has 600 (60/0.1) steps. Does the agent start to train when the experience buffer is filled up with the experiences (S,A,R,S'). If so, it would take at least 1667 (1000000/600 ) episodes before the agent starts to improve.

So I want to know how to determine this value.

0 个评论
显示 -2更早的评论隐藏 -2更早的评论

请先登录，再进行评论。

请先登录，再回答此问题。

采纳的回答

Ari Biswas 2021-11-17

0
链接

此回答的直接链接

https://ww2.mathworks.cn/matlabcentral/answers/1587044-experiencebufferlength-in-reinforcement-learning-toolbox#answer_833904

The agent will train until at least one minibatch can be sampled from the buffer. If your mini batch size is 64, then the first learn step will occur after the buffer has stored 64 experiences. The experience buffer is circular, i.e., it removes older experiences when full. The size of the buffer is hence important. You may lose important experiences if the buffer size is too small.