[redacted] enthusiast, robot combat enjoyer, distressingly Appalachian, father of ninjas

  • 11 Posts
  • 1.14K Comments
Joined 3 年前
cake
Cake day: 2023年10月8日

help-circle

  • First up: thanks for your hard work here in the sneer mines during these ridiculous times!

    I saw this being ignored on HN and I thought that you might appreciate it:

    Navier-Stokes Lost in Translation – Why Lean Verification of AI Autoformalisation Does not Guarantee Correct Natural Language Proofs

    https://arxiv.org/pdf/2610.08144

    The authors point out that automatic translation of natural language mathematics into Lean is very hard, actually. They also highlight some examples of such translation errors in the Navier-Stokes “paper” published by OpenAI. They tread lightly and don’t take a position on the correctness of the proof, but they do bring a large stack of receipts.

    I also thought the reference to the Solvability Complexity Index (which is new to me) was interesting, and there’s an appendix with an explainer on how the SCI hierarchy is constructed. According to this scheme, autoformalisation of natural language proofs is strictly harder than the Halting Problem.

    I haven’t had a chance to check on the bona fides of the authors.

    This is all beyond my level, but I’d love to see what our local experts think of it in light of the recent gish galop.