正文 · 原文
该语言的正文暂不可用,当前显示已有版本。
– https://t.co/bIlcWSZdux
Title: "Navier-Stokes lost in translation: Why Lean verification of AI autoformalisation does not guarantee correct natural language proofs"
人工智能 AI 安全与评测
自动核验发布 · 本文由系统生成并完成证据核验,未经人工审稿。
一篇新论文指出,聊天机器人可能在把错误的数学证明转写为 Lean 证明时,悄悄修正错误,使结果通过检验,却无法证明原始自然语言证明正确。相关介绍还称,判断一个陈述能否被忠实翻译,比停机问题更难,因此不存在总能做到这一点的 AI 翻译器。
该语言的正文暂不可用,当前显示已有版本。
– https://t.co/bIlcWSZdux
Title: "Navier-Stokes lost in translation: Why Lean verification of AI autoformalisation does not guarantee correct natural language proofs"
A new paper shows that when AI translates a math proof into Lean, passing the Lean check says nothing about whether the original proof is right.
They show a chatbot turning a wrong proof into a valid Lean proof by silently fixing the error.
Knowing when a statement can be translated faithfully is provably harder than the Halting problem, so no AI translator can always do it.
在 X 查看回复的帖子