The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks

arXiv:2609.25804v2 Announce Type: new Abstract: LLM agents increasingly work on long-horizon tasks, and the decisions they make along the way, such as which hypothesis to test or which implementation to build on, determine the outcome of the whole run. Making these decisions well is becoming a key…

aiscience

Sources