Sciweavers

4035 search results - page 511 / 807
» Useless Actions Are Useful
Sort
View
AROBOTS
1999
104views more  AROBOTS 1999»
15 years 7 months ago
Reinforcement Learning Soccer Teams with Incomplete World Models
We use reinforcement learning (RL) to compute strategies for multiagent soccer teams. RL may pro t signi cantly from world models (WMs) estimating state transition probabilities an...
Marco Wiering, Rafal Salustowicz, Jürgen Schm...
FAC
1998
111views more  FAC 1998»
15 years 7 months ago
A Formal Axiomatization for Alphabet Reasoning with Parametrized Processes
In the process-algebraic veri cation of systems with three or more components put in parallel, alphabet axioms are considered to be very useful. These are rules that exploit the i...
Henri Korver, M. P. A. Sellink
PRESENCE
2002
113views more  PRESENCE 2002»
15 years 7 months ago
"It/I": A Theater Play Featuring an Autonomous Computer Character
"It / I" is a two-character theater play where the human character I (played by a real actor) is taunted and played with by an autonomous computer character It on a comp...
Claudio S. Pinhanez, Aaron F. Bobick
TSMC
2002
98views more  TSMC 2002»
15 years 7 months ago
The STAR automaton: expediency and optimality properties
Abstract--We present the STack ARchitecture (STAR) automaton. It is a fixed structure, multiaction, reward-penalty learning automaton, characterized by a star-shaped state transiti...
Anastasios A. Economides, Athanasios Kehagias
CORR
2010
Springer
132views Education» more  CORR 2010»
15 years 7 months ago
Calibration and Internal no-Regret with Partial Monitoring
Calibrated strategies can be obtained by performing strategies that have no internal regret in some auxiliary game. Such strategies can be constructed explicitly with the use of B...
Vianney Perchet