Material Detail

Numerical exploration-exploitation trade-off for large-scale function optimization

This video was recorded at Large-scale Online Learning and Decision Making (LSOLDM) Workshop, Cumberland Lodge 2013. I will show how the "optimism in the face of uncertainty" principle developed in multiarmed bandits can be extended to address large scale decision making problems. Initially motivated by the empirical success of the Monte-Carlo tree search (MCTS) methods popularized in computer-go and further extended to many other optimization problems, I will report elements of theory that characterize the complexity of the underlying search problems and describe efficient algorithms with performance guarantees.

Keywords:: videolectures, ocwc, oec

Disciplines:

Science and Technology / Computer Science

More...

Go to Material

Bookmark / Add to Course ePortfolio

Create a Learning Exercise

Add Accessibility Information

Rate

Add a Comment

Quality

User Rating
Comments
Learning Exercises
Bookmark Collections
Course ePortfolios
Accessibility Info

Report Broken Link
Report as Inappropriate

More about this material

Material Type:: Presentation
Date Added to MERLOT:: February 10, 2015
Date Modified in MERLOT:: February 10, 2015
Author:: Rémi Munos, SequeL lab, INRIA Lille - Nord Europe
Submitter:: The Open Education Consortium
Primary Audience:: College General Ed, College Lower Division, College Upper Division
Technical Format:: Video

Mobile Compatibility:: Not specified at this time
Language:: English
Cost Involved:: No
Source Code Available:: No
Creative Commons:: This work is licensed under a Attribution-NonCommercial-NoDerivs 3.0 United States