2026-10-01 · America/Los_Angeles · 论文 · #13
EVOKE: Eliciting World Knowledge in Agents for Transferable Decision-Making
Large language models (LLMs) are increasingly deployed as agents for multi-step decision-making, yet transfer poorly to unseen environments. World-model methods address this by training agents to predict future observations, at the cost of additional training and errors that compound when predictions are used for planning. However, for LLM agents operating in digital environments, much of this world knowledge is already internalized during pretraining, which shifts the problem from acquiring it to eliciting it. We argue that typical post-training provides little pressure for such…
热度 54.2 / 100;排名与评分保留该期记录。