Skip to content
TrackPodcasts
technologyFeb 9, 201823:03pending

[MINI] Reinforcement Learning

Data Skeptic

About this episode

In many real world situations, a person/agent doesn't necessarily know their own objectives or the mechanics of the world they're interacting with. However, if the agent receives rewards which are correlated with the both their actions and the state of the world, then reinforcement learning can be used to discover behaviors that maximize the reward earned.

Get every episode summarized

Each time Data Skeptic publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

[MINI] Reinforcement Learning

Data Skeptic

0:00
23:03

More episodes

More from Data Skeptic

View all episodes →