Doesn't suit? No problem! You can return items for up to 30 days
You won't go wrong with a gift voucher. The gift recipient can choose anything from our offer.
Up to 30 days for returns
Presents sequential decision theory from a novel algorithmic information theory perspective. This book introduces the two different ideas and removes the limitations by unifying them to one parameter-free theory of an optimal reinforcement learning agent embedded in an unknown environment.
Hi! I'm Libroamiko, your book advisor.
How can I help you?