arXiv Machine Learning

LORE: Jointly Learning the Intrinsic Dimensionality and Relative Similarity Structure From Ordinal Data

arXiv AI
Jun 9

DiffoR: A Unified Continuous Generative Framework for Universal Ordinal Regression

arXiv:2606. 07599v1 Announce Type: cross Abstract: Ordinal Regression (OR) aims to predict target values with inherent order, underpinning critical applications across diverse domains, from recommender systems to computer vision.

By Hongxu Ma, Lin Wang, Chenghou Jin, Han Zhou, Jie Zhang, Xiaoyu Yang, Chunjie Chen, Jihong Guan, Shuigeng Zhou
arXiv Computer Vision
Aug 27

Towards Fine-Grained Text-to-3D Quality Assessment: A Benchmark and A Two-Stage Rank-Learning Metric

The paper introduces T23D-CompBench, a new benchmark for fine‑grained text‑to‑3D quality assessment that includes 3,600 textured meshes generated from ten state‑of‑the‑art models and 129,600 human ratings. It also proposes Rank2Score, a two‑stage rank‑learning metric that first trains with supervised contrastive regression and curriculum learning, then refines predictions using mean opinion scores to better align with human judgments. Experiments show Rank2Score outperforms existing metrics and can be used as a reward function for generative model optimization.

By Bingyang Cui, Yujie Zhang, Qi Yang, Zhu Li, Yiling Xu
arXiv AI
Sep 4

CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation

CORE improves compositional reasoning in multimodal language models by distilling a cross‑attentive reranker’s fine‑grained judgments into the embedding model. It generates candidate lists across five compositional matching levels and trains with a Rank‑KL objective to replicate the reranker’s ranking. Experiments on COLA, SUGARCREPE++, and NEGBENCH show CORE‑RERANKER‑8B outperforms Jina‑Reranker by 10.7 points, while CORE‑EMBED‑8B achieves the best overall average among evaluated embeddings, with gains also transferring to the MCMR benchmark without harming COCO or Flickr30K retrieval.

By Tingyu Song, Mingxin Li, Yanzhao Zhang, Dingkun Long, Chu Liu, Pengjun Xie, Yilun Zhao, Shu Wu
Hugging Face Trending Papers
Aug 11

E$^3$mo-Bench: A Scalable Benchmark for Multimodal Evoked and Expressed Emotion Understanding via Bayesian Pairwise Alignment

Understanding both expressed and evoked emotions is critical for multimodal large language models (MLLMs) to achieve comprehensive affect-aware interactions. However, existing benchmarks typically examine expressed and evoked emotions in isolation or are constrained to coarse-grained and incomplete affective characterizations.

arXiv Machine Learning
Aug 28

Aitchison Embeddings for Learning Compositional Graph Representations

The paper introduces a compositional graph embedding framework based on Aitchison geometry, where nodes are represented as simplex-valued mixtures over latent archetypal factors. By embedding these mixtures using isometric log-ratio coordinates, the method preserves Aitchison distances while allowing unconstrained optimization in Euclidean space, yielding intrinsically interpretable embeddings. The approach achieves competitive performance on node classification and link prediction tasks and enables principled component restriction through subcompositional coherence, allowing analysis of how archetype groups influence representations and predictions.

By Nikolaos Nakis, Chrysoula Kosma, Panagiotis Promponas, Michail Chatzianastasis, Giannis Nikolentzos
arXiv Machine Learning
Aug 28

The Rashomon Effect for Visualizing High-Dimensional Data

The paper introduces the Rashomon set for dimension reduction, a collection of equally good embeddings that preserve high‑dimensional structure. It proposes PCA‑informed alignment to make axes interpretable, concept‑alignment regularization to incorporate external knowledge, and a method to extract trustworthy nearest‑neighbor relationships across the Rashomon set for refined embeddings. These techniques aim to produce interpretable, robust, and goal‑aligned visualizations by leveraging multiple valid embeddings instead of a single one.

By Yiyang Sun, Haiyang Huang, Gaurav Rajesh Parikh, Cynthia Rudin