2022
DOI: 10.48550/arxiv.2205.08176
|View full text |Cite
Preprint
|
Sign up to set email alerts
|

On the Convergence of Policy in Unregularized Policy Mirror Descent

Abstract: In this short note, we give the convergence analysis of the policy in recent famous policy mirror descent (PMD) [5,12,8,6,11]. We mainly consider the unregularized setting following [11] with generalized Bregman divergence. The difference is that we directly give the convergence rates of policy under generalized Bregman divergence. Our results are inspired by the convergence of value function in previous works and are an extension study of policy mirror descent. Though some results have already appeared in pre… Show more

Help me understand this report

Search citation statements

Order By: Relevance

Paper Sections

Select...

Citation Types

0
0
0

Publication Types

Select...

Relationship

0
0

Authors

Journals

citations
Cited by 0 publications
references
References 6 publications
0
0
0
Order By: Relevance

No citations

Set email alert for when this publication receives citations?