Nathan Lambert
Nathan Lambert
@natolambert·8K subscribers·27 videos

[Paper Summary] Objective Mismatch in Model-based Reinforcement Learning

Posted

February 4, 2021

Views

297

Likes

6

Engagement

2.02%

Search the Record

Indexed

Every word spoken in this episode is indexed. Type any phrase to jump straight to the moment it was said.

Type any word or phrase that may have been spoken. Click a result to seek the player to that exact moment.

Try a name, a topic, or a quoted line

YouTube Description

as posted by the channel

Two optimization problems leave model-based RL in a tricky point: you cannot optimize both the model and the controller simultaneously. This video points a direction for a new class of model-based RL algorithms.

Sentinel Indexing in Progress

Metadata and chapters are available. Claim extraction for this episode is pending.

All video content is delivered via YouTube embedded players in accordance with the YouTube Terms of Service. Sentinel provides research tools that promote discovery and accountability across political media.

[Paper Summary] Objective Mismatch in Model-based Reinforcement Learning · Nathan Lambert · Sentinel