arXiv AI

A Variability-Based Framework for Interpretable Naming in Formal and Relational Concept Analysis

arXiv:2606. 08477v1 Announce Type: new Abstract: Knowledge extraction from symbolic data often produces abstractions that are formally defined but not immediately interpretable by users.

arXiv Machine Learning
Sep 2

Convergence issues in Relational Concept Analysis based on AOC-posets

The paper examines convergence problems in Relational Concept Analysis (RCA) when applied to AOC-posets instead of full concept lattices. It explains why RCA’s iterative process may fail to converge in the AOC-poset setting, identifies conditions that can still guarantee convergence, and proposes a convergent variant that preserves the AOC-poset structure by never removing relational attributes. The study also discusses data transformations that can restore convergence.

By Xavier Dolques, Agn\`es Braud, Alain Gutierrez, Marianne Huchard, Florence Le Ber
arXiv AI
Sep 21

On the Limitations of Large Language Models for Conceptual Database Modeling

The article examines how Large Language Models can aid in creating Entity-Relationship diagrams from natural language requirements. It tests three LLMs with three prompting strategies—Zero-Shot, Chain of Thought, and Chain of Thought + Verifier—on scenarios of increasing complexity. Findings show that while LLMs perform adequately on simpler tasks, their reliability drops with more complex requirements, leading to inconsistencies, ambiguities, and constraint representation failures.

By Arthur F. Siqueira, Carlos D. S. Nogueira, Eduarda Farias, Claudio E. C. Campelo, J\'ulia Menezes
arXiv AI
Aug 7

Tytan: Interactive Neurosymbolic Construction of Analytic Semantic Schemas from Relational Data

arXiv:2608. 06331v1 Announce Type: cross Abstract: From natural-language query interfaces to automated report generation, data analysis tools need a description of the data: the real-world entities it contains, which columns function as measures or identifiers, and how tables connect into units of analysis.

By Donna Hooshmand, Shubham Shahi, Cameron Barrie, Abhratanu Dutta, Marko Sterbentz, Harper Pack, Kristian J. Hammond
arXiv AI
Aug 28

Knowledge Cards: Structured Knowledge for AI Systems

The paper introduces Knowledge Cards, a new structured artefact designed to capture validated knowledge about specific concepts that AI systems use to make decisions. Unlike existing model, data, and system cards, Knowledge Cards focus on the layer between inputs and outputs, documenting entities, relationships, reasoning patterns, conditions for validity, and provenance, all grounded in a formal domain ontology and signed off by a domain expert. Prototype cards have been created in the energy and pharmaceutical domains, and the schema is released as a public draft for community engagement.

By Liliana Ferreira
arXiv AI
Jun 19

Toten: Knowledge-Based Ontological Tokenization Of Physical Quantities And Technical Notation In Brazilian Portuguese

arXiv:2606. 19626v1 Announce Type: new Abstract: Byte-Pair Encoding tokenization is statistically efficient for vocabulary compression, but semantically blind to structured technical entities, fragmenting physical quantities, numbers, units, and symbolic expressions into lexically arbitrary subwords.

By Antonio de Sousa Leit\~ao Filho; Allan Kardec Duailibe Barros Filho; Fabr\'icio Saul Lima; Selby Mykael Lima dos Santos; Rejani Bandeira Vieira Sousa