arXiv Computer Vision By Yani Guan, Dengpan Dong, Shuang Luo, Zi Wei, Joah Han, Dan Hannah, Yumin Zhang, Qichao Hu, Kang Xu

VERDICT: Agreement Beats Pixel-Space Verification in Real-Document OCSR

Read the original on arXiv Computer Vision →

The paper introduces VERDICT, a method for validating optical chemical structure recognition (OCSR) outputs by leveraging agreement among multiple recognizers rather than pixel‑space re‑rendering. On 263 ACS journal images, agreement achieved an AUROC of 0.916, far surpassing the 0.547 AUROC of re‑rendering similarity. VERDICT was applied to PMC Open Access, yielding over 6,000 high‑precision structure labels, and is integrated into SES AI’s Molecular Universe platform for image‑based molecular search.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv Machine Learning
3d ago

Interpretable-by-Design Descriptor Portfolios Match a 2048-Dimensional Foundation Embedding on Low-Data Molecular Assays

The study evaluates whether a portfolio of compact, semantically named descriptor blocks can match the performance of a 2048‑dimensional CheMeleon embedding in low‑data molecular assays. Using a fixed 11‑dimensional physicochemical base and greedily adding provenance‑screened blocks, the portfolio achieves a mean test AUC of 0.762 across nine ADME/Tox assays, comparable to CheMeleon’s 0.764 and better than Mordred’s 0.756. The results meet a predeclared pooled parity threshold but not all per‑assay thresholds, and further analysis confirms the competitiveness of the auditable representation while highlighting unresolved assay‑level differences.

By Yiqi Yao, Miquel Duran-Frigola
arXiv Machine Learning
Sep 14

What an odour descriptor corpus can and cannot measure: valence, attenuation, and the ceiling of the public record

The study evaluates the reliability of shared odor descriptor words across four public corpora from Pyrfume, finding substantial disagreement (I² = 80 %) and limited agreement on descriptor application (median tetrachoric = 0.795, κ = 0.413). Only a fraction of the achievable variance in odor perception is captured by current models and descriptor sets, with valence emerging as the primary missing component. Even with extensive model capacity and merged corpora, the gap remains, indicating that valence must be measured directly to improve machine olfaction.

By Stylianos Kampakis, Fabio Rovai
arXiv AI
Jul 29

Beyond Predictive Accuracy: A Reliability-Aware Audit of Molecular Representations for Human Olfaction

arXiv:2607. 24848v1 Announce Type: cross Abstract: Pretrained molecular encoders are commonly evaluated through downstream prediction, but predictive accuracy alone does not establish that a learned representation captures reproducible scientific structure, adds information beyond strong conventional baselines, or transfers out of distribution.

By Kai Lun Huang (California State University, Fullerton), Wei Chieh Sun (University of Washington)
arXiv Machine Learning
Jul 31

DS@GT ARC at ImageCLEFmedical 2026: Architectural Diversity for Concept Detection and Foundation-Model Scaling for Caption Prediction in Medical Image Analysis

arXiv:2607. 27763v1 Announce Type: cross Abstract: We describe the DS@GT submissions to the ImageCLEFmedical Caption 2026 challenge, which continues a long-running benchmark on the ROCOv2 dataset with two tracks: Concept Detection (Task 1), assigning UMLS Concept Unique Identifiers (CUIs) to radiology images, and Caption Prediction (Task 2), generating natural-language captions.

By Bowen Wang, Youwen Zhang, Ritesh Mehta