TD Algorithm for the Variance of Return and Mean-Variance Reinforcement Learning

Makoto Sato, Hajime Kimura, Shibenobu Kobayashi

Open source

DOI
10.1527/tjsai.16.353
Published
2001
Container
Transactions of the Japanese Society for Artificial Intelligence
Publisher
Japanese Society for Artificial Intelligence
Open access
unknown

Credibility signals

uncertain Score 64/100 under policy 1.0.0. This is a metadata assessment, not a judgment of the paper's conclusions.

Show all credibility signals

Cite this work

BibTeX

@article{allodium:10.1527/tjsai.16.353,
  title = {TD Algorithm for the Variance of Return and Mean-Variance Reinforcement Learning},
  author = {Makoto Sato and Hajime Kimura and Shibenobu Kobayashi},
  year = {2001},
  journal = {Transactions of the Japanese Society for Artificial Intelligence},
  doi = {10.1527/tjsai.16.353},
  url = {https://doi.org/10.1527/tjsai.16.353}
}

RIS

TY  - JOUR
TI  - TD Algorithm for the Variance of Return and Mean-Variance Reinforcement Learning
AU  - Makoto Sato
AU  - Hajime Kimura
AU  - Shibenobu Kobayashi
PY  - 2001
JO  - Transactions of the Japanese Society for Artificial Intelligence
DO  - 10.1527/tjsai.16.353
UR  - https://doi.org/10.1527/tjsai.16.353
ER  - 

APA

Sato, M., Kimura, H., & Kobayashi, S. (2001). TD Algorithm for the Variance of Return and Mean-Variance Reinforcement Learning. Transactions of the Japanese Society for Artificial Intelligence. https://doi.org/10.1527/tjsai.16.353

Source records