Sequential Independently Published Sequential Decision Making for AI: Interaction, Learning, Memory, and Control Paperback

Sequential Independently Published Sequential Decision Making for AI: Interaction, Learning, Memory, and Control Paperback

Alle 2 prijzen en aanbieders

Meest populaire keuze
Laagste prijs
Amazon.be· Bekende aanbieder
€ 36,73
3 tot 4 dagenGratis verzending
Check de website voor de levertijd | Gratis bezorgd > €20,-
Bekijk product
Bekijk product
Laagste prijs
Amazon.be Marketplace· Marketplace
€ 36,73
3 tot 4 dagenGratis verzending
Check de website voor de levertijd | Gratis bezorgd > €20,-
Bekijk
Bekijk product

Specificaties

Belangrijkste kenmerken
EAN
9798907070332

Productomschrijving

What changes when a learning system does more than predict-when it acts, changes what it will observe next, remembers a history, and must control a process over time?

Sequential Decision Making for AI develops the mathematical language for that setting. Beginning with decision theory, Markov decision processes, bandits, and dynamic programming, it builds toward reinforcement learning under exploration, function approximation, offline data, partial observability, planning, constraints, and multi-agent interaction. The unifying theme is interaction: once actions influence future data, the objects that govern learning change.

The book separates familiar ideas from the assumptions that make them valid. Representability is distinguished from learnability; Bellman closure from mere function-class membership; offline sample size from policy coverage; belief-state sufficiency from minimal memory; and statistical information from computational access. Upper bounds, lower bounds, and algorithmic guarantees are tied to the observation model, interaction protocol, horizon, and resource being counted.

The same framework is connected to modern foundation-model post-training and agent design: behavioural cloning, preference and reward modelling, KL-regularised optimisation, offline evaluation, context and recurrent memory, test-time planning, tool use, constrained control, and multi-agent interaction. These are interpreted through theorem-native quantities such as interaction rounds, independent samples, effective horizon, coverage, planning depth, memory bits, and computation.

Written for graduate students and researchers in machine learning, statistics, applied mathematics, control, and AI, this volume is not an algorithm catalogue. It is a resource-aware guide to what sequential-learning theorems establish, what they leave open, and what additional assumptions are required before they become claims about real AI agents.

Er zijn nog geen reviews geschreven

Vraag 1 van 4

Heb jij dit product in bezit en wil je graag je mening geven? Start dan hieronder met het schrijven van je review. Afhankelijk van de details duurt het schrijven van een review gemiddeld tussen de 3 en 10 minuten. Met jouw mening help je andere bezoekers een betere keuze te maken én maak je iedere maand kans op €250,-! Klik hier voor de actievoorwaarden.

Welk cijfer geef jij dit product?