arXiv AI
Sep 16

Extracting ontology-compliant knowledge from scientific text describing irradiated materials using large language models

arXiv:2609.17291v1 Announce Type: new Abstract: The quest for new materials increasingly relies on predictive models and comprehensive simulations that span scales from atomic to macroscopic levels....

By Marco Luca Sbodio, Marcos Mart\'inez Galindo, Vanessa Lopez, Blanca Biel, Pablo Canca, Pedro Delgado, Jes\'us I. Mendieta-Moreno, Raphael Tack, Maria J. Caturla
arXiv Computation and Language
Sep 16

SciNLP: A Domain-Specific Benchmark for Full-Text Scientific Entity and Relation Extraction in NLP

SciNLP is a new benchmark dataset for full‑text entity and relation extraction in the NLP domain, comprising 60 manually annotated papers with 6,429 entities and 1,649 relations. It is the first dataset to provide full‑text annotations of entities and their relationships specifically for NLP literature. Experiments show that models trained on SciNLP outperform baselines on certain tasks, and the dataset enabled the automatic construction of a fine‑grained knowledge graph with an average node degree of 3.3.

By Decheng Duan, Yingyi Zhang, Jitong Peng, Chengzhi Zhang
arXiv AI
Jul 28

Retrieval-Augmented Generation of Ontologies from Relational Databases

arXiv:2506. 01232v2 Announce Type: replace-cross Abstract: Deriving OWL ontologies from relational database schemas supports semantic interoperability and downstream tasks such as knowledge graph population, ontology-based data access, graph-based learning, and automated reasoning.

By Nadeen Fathallah, Mojtaba Nayyeri, Athish A Yogi, Ratan Bahadur Thapa, Hans-Michael Tautenhahn, Anton Schnurpel, Steffen Staab