(4 events) Time zone: Conference local time zone
Show all
Toggle Poster Visibility
Oral
Wed Jul 11 05:00 PM -- 05:20 PM (CEST) @ A1
Efficient Bias-Span-Constrained Exploration-Exploitation in Reinforcement Learning
Oral
Wed Jul 11 05:20 PM -- 05:40 PM (CEST) @ A1
Path Consistency Learning in Tsallis Entropy Regularized MDPs
Oral
Wed Jul 11 05:40 PM -- 05:50 PM (CEST) @ A1
Improved Regret Bounds for Thompson Sampling in Linear Quadratic Control Problems
Oral
Wed Jul 11 05:50 PM -- 06:00 PM (CEST) @ A1
Least-Squares Temporal Difference Learning for the Linear Quadratic Regulator
Successful Page Load