SEED: Agents Learn Rules From Their Own Completed Attempts

SEED turns completed agent runs into training guidance. Its authors report a 91.8% ALFWorld average, compared with 75.0% for standard reinforcement learning.
artificial-intelligence
Author

Kabui, Charles

Published

2026-07-19

Keywords

seed, ai-agents, reinforcement-learning, hindsight-skills, on-policy-distillation