arXiv Machine Learning By Lucas D. Konrad, Nikolas Kuschnig

Testing Most Influential Sets

Read the original on arXiv Machine Learning →

arXiv:2510. 20372v4 Announce Type: replace-cross Abstract: Small influential data subsets can dramatically impact model conclusions, with a few data points overturning key findings.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 10

From Citations to Contributions: LLM-Assisted Credit Scoring of Research Articles

The paper proposes a new method for credit scoring research articles that distinguishes between a paper’s original contribution and the prior work it builds upon. It introduces a hierarchical ‘contribution tree’ framework that conserves importance across a document’s structure and separates original from citation-derived credit. Large language models are employed as noisy comparative estimators to scale the analysis, and the approach is extended to collections of articles via weighted citation graphs to produce corpus-level contributions and normalized influence scores.

By Sana Ebrahimi, Suraj Shetiya, Abolfazl Asudeh
arXiv AI
Sep 10

Optimal Experiments for Partial Causal Effect Identification

The paper tackles selecting a cost‑constrained set of experiments that most effectively tighten bounds on a partially identifiable causal query. It formalizes this as the NP‑hard max‑potency problem, introduces efficient graphical pruning rules to reduce the search space, and demonstrates the approach on synthetic graphs and real NHANES data to estimate the effect of physical activity on diabetes.

By Tobias Maringgele, Jalal Etesami