arXiv AI By Frederik Hagelskj{\ae}r

A novel network for classification of cuneiform tablet metadata

Read the original on arXiv AI →

arXiv:2603. 03892v2 Announce Type: replace-cross Abstract: In this paper, we present a network structure for classifying metadata of cuneiform tablets.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computer Vision
Sep 18

Automated Goldsmith's Mark Retrieval in Silverware

The paper introduces an AI-assisted pipeline for retrieving goldsmith marks from silverware images, combining mark localization with metric‑learning fine‑tuning on three backbone models (ResNet‑50, ViT‑S/16, and DINOv2 ViT‑S/14). Systematic tests of cropping strategies show that manual cropping and metric‑learning fine‑tuning yield the best performance, with DINOv2 ViT‑S/14 achieving 62.63% mAP and 73.74% Top‑1 accuracy. The authors release a manually annotated dataset, code, and a public web interface to support reproducibility and adoption in digital humanities.

By Atmik Tiwari, Vincent Christlein, Mark Fichtner, Freya Gohlke, Birgit Sch\"ubel, Theresa Witting, Heike Zech, Mathias Zinnen
Hugging Face Trending Papers
Jun 21

Automated sign detection across the Electronic Babylonian Library: A large-scale dataset and end-to-end cuneiform OCR pipeline

Learning to read cuneiform tablets is an extremely demanding task; consequently, of the roughly half million excavated tablets, only a small fraction has been analysed by Assyriologists. Computer vision offers a promising avenue for decipherment but requires large, densely annotated datasets.

Hugging Face Trending Papers
5d ago

Evaluating Hierarchy-Aware Deep Learning for the Recognition of Tironian Notes

The paper evaluates whether incorporating the hierarchical structure of Tironian notes can improve automatic recognition of this complex Latin shorthand system. Experiments compare flat classifiers (ResNet18, ConvNeXt, Swin, ViT) with hierarchy‑aware models (HD‑CNN and routing approaches) on handwritten and manuscript samples, with and without few‑shot adaptation. Results show that hierarchical models outperform flat ones when no adaptation is applied, but flat models surpass them after few‑shot adaptation, indicating that hierarchy can aid recognition under non‑adapted conditions.

arXiv Machine Learning
Sep 18

A Large-Scale Vision-Language Dataset Derived from Open Scientific Literature to Advance Biomedical Generalist AI

The paper introduces Biomedica, an open-source dataset sourced from PubMed Central that includes over 6 million scientific articles and 24 million image‑text pairs, along with 27 metadata fields and expert human annotations. To facilitate use, the authors provide scalable streaming and search APIs via a web server. They demonstrate the dataset’s value by training embedding models, chat‑style models, and retrieval‑augmented chat agents, all of which outperform previous open systems in their categories.

By Alejandro Lozano, Min Woo Sun, James Burgess, Jeffrey J. Nirschl, Christopher Polzak, Yuhui Zhang, Liangyu Chen, Jeffrey Gu, Ivan Lopez, Josiah Aklilu, Anita Rau, Austin Wolfgang Katzer, Collin Chiu, Orr Zohar, Xiaohan Wang, Alfred Seunghoon Song, Chiang Chia-Chun, Robert Tibshirani, Serena Yeung-Levy