Sciweavers

23703 search results - page 403 / 4741
» Learning from Demonstration
Sort
View
SIGMOD
2007
ACM
197views Database» more  SIGMOD 2007»
16 years 7 months ago
Automated and on-demand provisioning of virtual machines for database applications
Utility computing delivers compute and storage resources to applications as an `on-demand utility', much like electricity, from a distributed collection of computing resource...
Piyush Shivam, Azbayar Demberel, Pradeep Gunda, Da...
ICML
2004
IEEE
16 years 8 months ago
Apprenticeship learning via inverse reinforcement learning
We consider learning in a Markov decision process where we are not explicitly given a reward function, but where instead we can observe an expert demonstrating the task that we wa...
Pieter Abbeel, Andrew Y. Ng
ICML
2008
IEEE
16 years 8 months ago
Democratic approximation of lexicographic preference models
Previous algorithms for learning lexicographic preference models (LPMs) produce a "best guess" LPM that is consistent with the observations. Our approach is more democra...
Fusun Yaman, Thomas J. Walsh, Michael L. Littman, ...
ICML
2005
IEEE
16 years 8 months ago
Exploration and apprenticeship learning in reinforcement learning
We consider reinforcement learning in systems with unknown dynamics. Algorithms such as E3 (Kearns and Singh, 2002) learn near-optimal policies by using "exploration policies...
Pieter Abbeel, Andrew Y. Ng
ECTEL
2006
Springer
15 years 11 months ago
CELEBRATE's Lessons
The CELEBRATEproject developed and successfully demonstrated a federated learning object brokerage system architecture and made available to schools over 1350 learning objects prod...
David Massart