arXiv AI By Riccardo Revalor, Jalees Rehman, Debjit Pal

Can We Trust LLM's Logic? Quantifying Uncertainty, Coherence, and Robustness via a Graph-Based Framework

Read the original on arXiv AI →

arXiv:2607. 08017v1 Announce Type: cross Abstract: Large-Language Models (LLMs) can be prone to flawed and unfaithful reasoning that decoding strategies like Self-Consistency (SC) fail to detect as they evaluate only final-answer agreement while ignoring the logical validity of intermediate steps.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.