Hilbert Space Embeddings of POMDPs

Nishiyama, Yu; Boularias, Abdeslam; Gretton, Arthur; Fukumizu, Kenji

Computer Science > Machine Learning

arXiv:1210.4887 (cs)

[Submitted on 16 Oct 2012]

Title:Hilbert Space Embeddings of POMDPs

Authors:Yu Nishiyama, Abdeslam Boularias, Arthur Gretton, Kenji Fukumizu

View PDF

Abstract:A nonparametric approach for policy learning for POMDPs is proposed. The approach represents distributions over the states, observations, and actions as embeddings in feature spaces, which are reproducing kernel Hilbert spaces. Distributions over states given the observations are obtained by applying the kernel Bayes' rule to these distribution embeddings. Policies and value functions are defined on the feature space over states, which leads to a feature space expression for the Bellman equation. Value iteration may then be used to estimate the optimal value function and associated policy. Experimental results confirm that the correct policy is learned using the feature space representation.

Comments:	Appears in Proceedings of the Twenty-Eighth Conference on Uncertainty in Artificial Intelligence (UAI2012)
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Machine Learning (stat.ML)
Report number:	UAI-P-2012-PG-644-653
Cite as:	arXiv:1210.4887 [cs.LG]
	(or arXiv:1210.4887v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1210.4887

Submission history

From: Yu Nishiyama [view email] [via AUAI proxy]
[v1] Tue, 16 Oct 2012 17:46:07 UTC (465 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.AI

< prev | next >

new | recent | 2012-10

Change to browse by:

cs
cs.LG
stat
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

Yu Nishiyama
Abdeslam Boularias
Arthur Gretton
Kenji Fukumizu

export BibTeX citation

Computer Science > Machine Learning

Title:Hilbert Space Embeddings of POMDPs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Hilbert Space Embeddings of POMDPs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators