arXiv AI By Jakob Suchan, Julius Monsen, Salim Baloch, Mehul Bhatt

Answer Set Programming Energised! End-to-End Neurosymbolic Reasoning and Learning with ASP and Energy Based Models

Read the original on arXiv AI →

arXiv:2607. 08136v1 Announce Type: new Abstract: We present a general neurosymbolic reasoning and learning methodology based on a modular integration of answer set programming with an energy based model substrate.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

Hugging Face Trending Papers
Jul 9

Answer Set Programming Energised! End-to-End Neurosymbolic Reasoning and Learning with ASP and Energy Based Models

We present a general neurosymbolic reasoning and learning methodology based on a modular integration of answer set programming with an energy based model substrate. Key contributions are: (1) supporting joint optimisation in the continuous latent space through explicit ASP-based declarative semantics fully incorporating background knowledge, constraints, non-monotonic inference; and (2) advancing recent works at the interface of answer sets, probabilistic logic, and answer set modulo theories by providing a generalised model and practical platform for ASP-centric robust, end-to-end training for applications in dynamic domains (e.

arXiv AI
Sep 21

Beyond Exact Match: Task-Aware GRPO for Cross-Domain PCBA Visual Question Answering

The paper introduces a multimodal reasoning framework for cross‑domain visual question answering in Printed Circuit Board Assembly (PCBA) inspection, converting diverse data sources into a unified instruction format and generating verified reasoning traces. It proposes Task‑Aware Group Relative Policy Optimization (GRPO) that uses semantic, distance‑aware, and format rewards to improve choice‑based and counting tasks beyond exact‑match supervision. During inference, the system applies semantic consistency correction, self‑consistency voting, and multi‑model arbitration, achieving an overall score of 83.24 on the PCBA Standard‑to‑Real Grand Challenge leaderboard.

By Jia Li, Li Dai, Peng Jia, Zhenzhen Hu, Chee Seng Chan, Bingkun Bao, Richang Hong
arXiv AI
Jun 8

Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization

arXiv:2606. 07000v1 Announce Type: new Abstract: Recent post-training methods, particularly Reinforcement Learning with Verifiable Rewards (RLVR), have significantly enhanced the reasoning ability of Large Vision-Language Models (LVLMs).

By Shizhe Xiang, Ke An, Wenlong Yu, Yue Liu, Jian Luan, Pei Fu, Qilong Wang