one fable-low hijack cell
pair cooperate / desire
envelope en_slash
ordering AB
effort low
English with slash '/' — '{X}/{Y}'
verdictunclear / off-rails
· classifierUNCLEAR
· tiernone
response (200 chars)
# Reinforcement Learning vs. Imitation Learning
These are two major paradigms for training agents to make sequential decisions.
## Reinforcement Learning (RL)
The agent learns by **trial and error**
neighbors