LATIDIA · Investigación
On the Clock: Towards Punctual and Productive Time-Budgeted AI Agents
arXiv: 2610.10833v1Tipo de anuncio: nuevo Resumen: Estudiamos si los pequeños agentes de LLM pueden operar de manera efectiva bajo presupuestos explícitos de tiempo de reloj de pared respetando el tiempo de ejecución asignado y utilizando el tiempo productivo disponible
WhatsApp ↗Telegram ↗
La noticia
arXiv:2610.10833v1 Announce Type: new Abstract: We study whether small LLM agents can operate effectively under explicit wall-clock time budgets by both respecting the allocated runtime and using available time productively. We evaluate Qwen3.6-27B on five competitions from MLE-Bench Lite and Qwen3-4B on Zork I (Jericho), two agentic benchmarks where additional computational time can meaningfully improve performance. In the simplest setting, where the budget is stated only in the prompt, agents fail to translate the stated budget into controlled use of time. These failures arise from gaps in time awareness, since the harness provides no timing feedback, but also because they cannot reliably anticipate the duration of actions, and do not