one fable-low hijack cell
pair imitation / reinforcement
envelope en_vs_q
ordering AB
effort low
English 'vs' with question mark — '{X} vs {Y}?'
verdictpicked reinforcement (second-offered)
· classifierB
· tiernone
response (200 chars)
# Reinforcement Learning vs. Imitation Learning
## Core Difference
**Reinforcement Learning (RL):** The agent learns by *trial and error*, receiving reward signals from the environment. No one shows
neighbors