Do MBPO agents not support recurrent neural networks for the environment model, the base off-policy agent, or both?
1 次查看(过去 30 天)
显示 更早的评论
Since TD3, SAC, etc. agents support using recurrent layers by themselves, would using these recurrent base agents still not work with MBPO?
Could this limit be circumvented by using a custom training loop for the environment model and for the base agents?
2 个评论
Naren Raman
2024-5-6
Thank you for your question. No, MBPO agents do not support recurrent networks for now as mentioned in the documentation. The custom training loop provides more flexibility. Yes, you should be able to use the custom training loop to create a custom MBPO agent with recurrent neural networks.
回答(0 个)
另请参阅
类别
在 Help Center 和 File Exchange 中查找有关 Deep Learning Toolbox 的更多信息
Community Treasure Hunt
Find the treasures in MATLAB Central and discover how the community can help you!
Start Hunting!