arXiv AI By Momina Ahsan, Sarfraz Ahmad, Ming Shan Hee, Roy Ka-Wei Lee, Preslav Nakov

TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs

Read the original on arXiv AI →

arXiv:2606. 09578v1 Announce Type: new Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) are increasingly evaluated on table reasoning tasks, but the role of table representation remains under-explored.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Aug 28

A Table Is Worth 64 Tokens: Pixel-level Compression for Multi-Table Document Question Answering

The paper investigates pixel-level table compression for document question answering, comparing five vision‑language models across two benchmarks and varying visual‑token budgets. It finds that representing tables as native‑resolution images matches text in performance and efficiency, while highly downscaled images still allow the model to identify relevant tables but lose readability, leading to longer reasoning traces. A two‑step, training‑free method first selects relevant tables from compressed images and then reasons over them at native resolution, saving 41% of tokens and improving accuracy by 7 points over single‑step native‑resolution QA, while using 15% fewer tokens than the most efficient single‑step compressed setup without accuracy loss.

By I\~nigo Alonso, Mirella Lapata
arXiv Machine Learning
Jun 9

IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation

arXiv:2601. 04498v2 Announce Type: replace Abstract: Infographics are composite visual artifacts that combine data visualizations with textual and illustrative elements to communicate information.

By Yinghao Tang, Xueding Liu, Boyuan Zhang, Tingfeng Lan, Yupeng Xie, Jiale Lao, Yiyao Wang, Haoxuan Li, Tingting Gao, Bo Pan, Luoxuan Weng, Xiuqi Huang, Minfeng Zhu, Yingchaojie Feng, Yuyu Luo, Wei Chen
arXiv Machine Learning
Sep 4

Towards Universal Tabular Embeddings: A Benchmark Across Data Tasks

The paper introduces TEmBed, a unified benchmark for evaluating tabular embeddings across four representation levels—cell, row, column, and table—using a diverse set of models. It demonstrates that the best model depends on the specific task and representation level, providing practical guidance for selecting embeddings in real-world applications. The study aims to facilitate the development of more general-purpose tabular representation models.

By Liane Vogel, Kavitha Srinivas, Niharika D'Souza, Sola Shirai, Oktie Hassanzadeh, Horst Samulowitz