SEP 8, 2026 · PREPRINT
Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks
arXiv
This is an unrefereed preprint describing a novel environment-design approach for training LLM agents via RL, with computational experiments on benchmarks but no peer-reviewed publication and no clinical or direct human outcomes.
Study details
DesignComputational benchmarking study with pilot…
InterventionFeedback-Enriched Environments (FEEs): envi…
ComparatorStandard environment settings without feedb…