arXiv AI

LLM-Assisted Discovery of Typed Semantic Links for Ontology Network Construction

arXiv AI
Sep 10

Building evidence-based knowledge bases from full-text literature for disease-specific biomedical reasoning

EvidenceNet is a disease‑specific dataset that transforms full‑text biomedical literature into structured evidence records and graph representations, preserving study design, provenance, and quantitative support. Using an LLM‑assisted pipeline, it extracts experimentally grounded findings, normalizes entities, scores evidence quality, and links related records via typed semantic relations. The released subsets—EvidenceNet‑HCC and EvidenceNet‑CRC—contain thousands of evidence records and richly connected graphs, with high extraction and relation‑type accuracy, enabling retrieval‑augmented question answering and graph‑based tasks such as link prediction and target prioritization.

By Chang Zong, Jinyu Chen, Sicheng Lv, Si-tu Xue, Huilin Zheng, Jian Wan, Lei Zhang
arXiv AI
Jul 28

Retrieval-Augmented Generation of Ontologies from Relational Databases

arXiv:2506. 01232v2 Announce Type: replace-cross Abstract: Deriving OWL ontologies from relational database schemas supports semantic interoperability and downstream tasks such as knowledge graph population, ontology-based data access, graph-based learning, and automated reasoning.

By Nadeen Fathallah, Mojtaba Nayyeri, Athish A Yogi, Ratan Bahadur Thapa, Hans-Michael Tautenhahn, Anton Schnurpel, Steffen Staab
arXiv AI
Sep 2

Do General NLP Embeddings Capture Ontological Reasoning?

The paper introduces AVA, a framework that tests whether general NLP embeddings can differentiate logic-sensitive relational semantics in ontologies and knowledge graphs. AVA uses 171,007 contrastive triplets from 163 ontologies, each containing an ontology statement, a paraphrase, and a hard negative with contradictory meaning. Evaluation of over 25 embedding models shows significant limitations, with the best model achieving only 0.739 triplet accuracy and 0.135 for hard negatives; fine‑tuning helps but does not transfer well to downstream Semantic Web tasks.

By Hamed Babaei Giglou, Jennifer D'Souza, S\"oren Auer