AINEWS 搜索
返回 Rohan Paul (@rohanpaul_ai)
Rohan Paul (@rohanpaul_ai)· · 原发布时间

MIT 研究:简化选择步骤和说明规则可改善大模型代理决策

自动核验发布 · 本文由系统生成并完成证据核验,未经人工审稿。

AI 辅助摘要

一项 MIT 研究发现,GPT-4o、Claude、Gemini 和 Gemma 在本应如实出价的拍卖与匹配任务中仍会压低报价。把选择改成简单的“留下或退出”,Gemma 的报价与其估值的差距从 5.30 美元缩至 0.30 美元;提示代理考虑对手则可能增加错误。

正文 · 原文

New MIT Paper: LLM agents decide better when the choice is shown in simple steps or the rule's safe move is stated plainly, and worse when told to reason about opponents.

Market-design rules of thumb built for human bidders carry over to LLM agents, so we can borrow them instead of inventing new prompt tricks.

Honest bids and rankings are always the best move in these auctions and matching games. GPT-4o, Claude, Gemini and Gemma still underbid, often to keep a profit margin.

A rising price with a simple stay-or-exit choice moved Gemma from $5.30 below its value to $0.30 below. A 1-line note that rejections only redirect cut matching errors from 4.2% to 0.2%.

Fix the format and state the key fact before adding reasoning prompts, and judge agents by their choices, since their written plans missed these gains.

Agent prompts should state facts about the rules rather than request more thinking, because facts improved choices while thinking prompts often added errors.

发现内容有误?提交纠错