arXiv AI

Interaction valence reveals contrasting social networks in dairy cattle

arXiv:2608. 19222v1 Announce Type: new Abstract: Social relationships shape access to resources, exposure to conflict and group stability, yet automated livestock monitoring typically treats behaviour as isolated events.

arXiv Machine Learning
Aug 12

EweAcT: Ewe behaviour aligned to accelerometer data for activity monitoring in extensive grazing systems

arXiv:2608. 09943v1 Announce Type: cross Abstract: Monitoring livestock behaviour under extensive conditions would provide valuable insights to assess animal adaption to environmental perturbations in agroecological systems (e.

By Lucile Riaboff (GenPhySE, INRAE), Ny Aina Andriamampandry (GenPhySE, GenPhySE), Jean-Fran\c{c}ois Bompa (GenPhySE, GenPhySE), Mathias Aletru (GenPhySE, GenPhySE), Christian Durand (UEF), S\'ebastien Douls (UEF), Ga\"etan Bonnafe (UEF), Morgane Costes-Thir\'e (GenPhySE, GenPhySE), Guillaume Delosi\`eres (GenPhySE, GenPhySE), Jean- Marc Mongrelet (GenPhySE, GenPhySE), Enzo Niro (GenPhySE, GenPhySE), N\'emuel Tadi (GenPhySE, GenPhySE), S\'everine Deretz (DEPT GA, UEF, INRAE), Sara Parisot (UEF), Margot Lamarque (UEF), Dominique Hazard (GenPhySE), Emilie Cobo (GenPhySE)
arXiv AI
Jun 4

Do LLMs Hold Their Values? MANTA: A Multi-Turn Adversarial Benchmark for Animal Welfare Reasoning

arXiv:2605. 16301v2 Announce Type: replace-cross Abstract: Evaluating animal welfare reasoning in LLMs remains an open challenge despite rapid deployment in consumer and professional contexts where welfare considerations appear implicitly in everyday queries.

By Isabella Luong, Joyee Chen, Arturs Kanepajs, Jasmine Brazilek, Sankalpa Ghose, David Williams-King, Linh Le, Allen Lu
arXiv Machine Learning
Aug 31

What Do Interaction Representations Actually Measure? Pre-Event Separability in Weakly-Supervised Violence Detection

The paper investigates whether detailed articulated human pose provides more discriminative power than coarse spatial relationships for early violence detection. By fixing the downstream pipeline and comparing five interaction representations—including bounding‑box geometry, handcrafted pose analogues, enriched pose descriptors, and a learned joint encoder—the study finds that pose‑based representations do not outperform coarse geometry. When visual encoders are frozen and evaluated on larger datasets, person‑crop appearance and whole‑frame context outperform geometry, but cropping to interacting people offers no advantage over encoding the entire frame. The authors further demonstrate that pre‑onset frames contain source‑related artifacts (e.g., title cards, watermarks) that contribute significantly to discrimination, suggesting that benchmark performance may reflect these artifacts rather than true event evidence.

By Parishruthi Ganesh
arXiv Computation and Language
Sep 23

Same Chart, Different Story: Bias in Vision-Language Chart Interpretation

The paper introduces ChartBias, a benchmark of 820 real-world charts covering six social attributes, designed to audit bias in vision‑language models (VLMs) that interpret charts. Across 12 VLMs, the study identifies three failure modes—narrative shift, group hallucination, and preference polarity—where models produce different or misleading narratives when the referenced social group changes. A multi‑agent mitigation framework is proposed, separating evidence extraction from group‑conditioned generation and using a counterfactual judge, which reduces narrative shift while maintaining chart‑grounded reasoning.

By Mizanur Rahman, Huan Wu, Arash Asgari, Enamul Hoque Prince, Laleh Seyyed-Kalantari
arXiv AI
Sep 10

What Does Animal Re-Identification Learn? Linear Biological Concepts and Their Origins in Visual Representations

The study investigates whether Vision Transformer (ViT)-based animal re-identification models learn biologically meaningful concepts. Using a DINOv3 backbone fine‑tuned on Western lowland gorilla images, the authors find that sex and age emerge as linear directions in the model’s representations, generalizing to unseen individuals with high AUROC scores. They demonstrate that the sex direction is causally used by the model, that fine‑tuning relocates these concepts within the network, and that the representations reflect a graded biological axis encoded redundantly across the population.

By Robert Nolting, Alexandra Schild, Moritz Weckbecker, Maximilian Schall, Gerard de Melo
arXiv AI
4d ago

VISTA: Value-Informed Event Appraisal for Multimodal Emotion Conflict

VISTA (Value-Informed Semantic Trust Arbitration) is a learned seven-field appraisal interface that conditions modality arbitration on concerns, event relations, and expression conditions while retaining a joint-evidence residual. It uses a log-odds decomposition to separate emotion expectation from cue diagnosticity, allowing appraisal to change how evidence is interpreted. With a shared Qwen2.5-Omni-7B backbone, VISTA achieves 64.5% conflict accuracy on CA-MER, improving on modality gating by 2.5 percentage points on conflict and 0.2 on consistency, and a frozen-backbone probe reaches 0.600 macro CCC for appraisal readout versus 0.505 for emotion-only fine-tuning.

By Jiale Dai, Liuxian Ma, Xiaoke Niu, Wenjing Zhang, Huiying Zhao, Zhaoxiang Liu, Shiguo Lian, Guojie Song
arXiv AI
Sep 16

Pseudo-Label Augmentation for Affect Sensing in Small Collaborative Groups

The study examines pseudo‑label augmentation for affect sensing in small collaborative groups using the GroupAffect‑4 dataset, which includes wearable physiology, eye tracking, personality traits, and post‑task valence, arousal, and dominance (VAD) labels. Various augmentation strategies—no augmentation, Gaussian Process pseudo‑labelling, personality‑aware trust weighting, and joint personality‑plus‑confidence weighting—were evaluated within a shared target‑construction pipeline. Results show that pseudo‑label augmentation improves performance over a labelled‑only baseline in the known‑team setting, with the joint personality‑plus‑confidence variant achieving the highest dominance score, while personality similarity mainly serves as a same‑team filter rather than a calibrated trust signal.

By Meisam Jamshidi Seikavandi, Tanya Ignatenko, Fabricio Batista Narcizo, Paolo Burelli, Jesper B\"unsow Boldt, Andrew Burke Dittberner