Behavioral Controllability of Agentic Models for Information Extraction: From Fixed Workflows to Reflective Agents
arxiv.org原文 ↗
作者用 conference-paper dataset extraction 检验固定 LLM workflow、reflective agent、memory 和 richer PDF tools 是否真的改变可控行为。评估把 tool execution、retries、reflection、memory use、runtime、failure recovery 放在第一层,coverage 和 field completeness 反而是次级指标。这个角度让“agentic component 是否有用”从结果分数转向过程可控性,适合审视复杂信息抽取系统的工程收益。
–浏览
评论 · Comments