Sciweavers

12194 search results - page 419 / 2439
» cans 2010
Sort
View
NECO
2010
97views more  NECO 2010»
15 years 6 months ago
Derivatives of Logarithmic Stationary Distributions for Policy Gradient Reinforcement Learning
Most conventional Policy Gradient Reinforcement Learning (PGRL) algorithms neglect (or do not explicitly make use of) a term in the average reward gradient with respect to the pol...
Tetsuro Morimura, Eiji Uchibe, Junichiro Yoshimoto...
OL
2010
185views more  OL 2010»
15 years 6 months ago
On the Hamming distance in combinatorial optimization problems on hypergraph matchings
In this note we consider the properties of the Hamming distance in combinatorial optimization problems on hypergraph matchings, also known as multidimensional assignment problems....
Alla R. Kammerdiner, Pavlo A. Krokhmal, Panos M. P...
P2P
2010
IEEE
280views Communications» more  P2P 2010»
15 years 6 months ago
Performance Evaluation of Peer-to-Peer Gaming Overlays
—In this demo we present a performance evaluation testbed for peer-to-peer gaming overlays. It consists of a 3D first person shooter game that is designed to run in a simulated ...
Max Lehn, Tonio Triebel, Cchritof Leng, Alejandro ...
PTS
2010
138views Hardware» more  PTS 2010»
15 years 6 months ago
Alternating Simulation and IOCO
We propose a symbolic framework called guarded labeled assignment systems or GLASs and show how GLASs can be used as a foundation for symbolic analysis of various aspects of forma...
Margus Veanes, Nikolaj Bjørner
PVLDB
2010
120views more  PVLDB 2010»
15 years 6 months ago
Cloudy: A Modular Cloud Storage System
This demonstration presents Cloudy, a modular cloud storage system. Cloudy provides a highly flexible architecture for distributed data storage and is designed to operate with mu...
Donald Kossmann, Tim Kraska, Simon Loesing, Stepha...