arXiv AI By Donghwan Kim

LLM-as-a-Judge Scores Are Unreliable Optimization Signals in Closed-Loop Table Recognition

Read the original on arXiv AI →

arXiv:2607. 13347v2 Announce Type: replace-cross Abstract: LLM-as-a-judge is widely used to provide feedback and selection signals in closedloop regeneration, but this use remains insufficiently validated.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.