Back to papers
May 14, 2026cs.LG

Learning from Language Feedback via Variational Policy Distillation

HF Upvotes

9

Categories

cs.LG