Oscillation of Episode Q0 during DDPG training

4 vues (au cours des 30 derniers jours)

Afficher commentaires plus anciens

Heesu Kim le 6 Avr 2021

0
Lien

Utiliser le lien direct vers cette question

https://fr.mathworks.com/matlabcentral/answers/794607-oscillation-of-episode-q0-during-ddpg-training

Commenté : Heesu Kim le 6 Avr 2021

How do I interpret this kind of Episode Q0 oscillation?

The oscillation shows a pattern like up and down and the range also increases quite regularly.

According to other docs, they're saying the Q0 is supposed to approach actual discounted future reward as long as the critic network is designed properly.

Is this kind of Q0 oscillation just evidence that my critic network is not well-designed?

Is there any solution to work it out?

I'm not sure this question is acceptable to this community because I think it's more or less a theoretical issue.

1 commentaire
Afficher -1 commentaires plus anciensMasquer -1 commentaires plus anciens

Heesu Kim le 6 Avr 2021

As a side note, I'm using DDPG + LSTM model that RL toolbox provides

Connectez-vous pour commenter.

Connectez-vous pour répondre à cette question.

Réponses (0)

Connectez-vous pour répondre à cette question.

Catégories

AI and Statistics Deep Learning Toolbox Applications Autonomous and Control Systems Reinforcement Learning

En savoir plus sur Reinforcement Learning dans Help Center et File Exchange

Produits

Reinforcement Learning Toolbox

Version

R2021a

Community Treasure Hunt

Find the treasures in MATLAB Central and discover how the community can help you!

Start Hunting!

Translated by

Oscillation of Episode Q0 during DDPG training

1 commentaire
Afficher -1 commentaires plus anciensMasquer -1 commentaires plus anciens

Réponses (0)

Voir également

Catégories

Tags

Produits

Version

Community Treasure Hunt

Oscillation of Episode Q0 during DDPG training

1 commentaire Afficher -1 commentaires plus anciensMasquer -1 commentaires plus anciens

Réponses (0)

Voir également

Catégories

Tags

Produits

Version

Community Treasure Hunt

1 commentaire
Afficher -1 commentaires plus anciensMasquer -1 commentaires plus anciens