← Back to all news
arXiv Machine Learning August 27, 2026 By Chandler Smith, Magnus Sesodia, Friedrich Lindenberg, Christian Schroeder de Witt

OpenSanctions Pairs: Large-Scale Entity Matching with LLMs

Read the original on arXiv Machine Learning →

The Flow has not summarised this story yet — read it at arXiv Machine Learning.

  • llms
  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
Jul 28

Beyond Scale and Generation: Understanding Language Model-based Entity Matching

arXiv:2607. 24688v1 Announce Type: cross Abstract: Entity matching identifies records that refer to the same real-world entity.

By Zeyu Zhang, Xue Li, Iacer Calixto, Paul Groth, Sebastian Schelter
llmsragfine-tuningbenchmarks
More like this →
arXiv Machine Learning
Aug 7

Where Privacy Risk Lives in English-Source Multilingual RAG: A Stage-Decomposed Audit Across Five Query Languages

arXiv:2608. 05163v1 Announce Type: cross Abstract: A common assumption holds that switching to a non-English language makes a multilingual RAG system easier to attack for personal information.

By Yanhang Li, Zhichao Fan, Zexin Zhuang
llmsragbenchmarkssafety
More like this →
arXiv AI
Aug 10

Multi-Legal-Bench: Evaluating LLMs on Legal Reasoning Across Jurisdictions, Languages, and Legal Traditions

arXiv:2605. 29738v2 Announce Type: replace-cross Abstract: Legal NLP benchmarks overwhelmingly evaluate a single language or aggregate tasks that differ fundamentally across jurisdictions, making cross-lingual comparison impossible.

By Volodymyr Ovcharov
llmsnlpbenchmarkssafety
More like this →
arXiv AI
Aug 7

OpenAI Privacy Filter: A Cross-Lingual, Cross-Domain PII Evaluation Across 32 Benchmarks

arXiv:2608. 02616v2 Announce Type: replace-cross Abstract: We present what is, to our knowledge, the first systematic evaluation of OpenAI's Privacy Filter (OPF), a 1.

By Rohith Uppala
llmsfine-tuningbenchmarkssafety
More like this →
arXiv AI
5d ago

When Names Cross Scripts: A Source-Grounded Benchmark for Historical Entity Reconciliation in the Mongol World

arXiv:2608.23507v1 Announce Type: cross Abstract: Historical people may appear under different languages, scripts, and transcription traditions, while distinct individuals may share highly similar or...

By Xiang Chen, Zeyu Zhang
nlpbenchmarks
More like this →
arXiv AI
Jul 29

IRIS: Reusable Identity Representations from Frozen LLMs for Entity Alignment

arXiv:2607. 25579v1 Announce Type: cross Abstract: Entity alignment (EA) identifies entities across knowledge graphs (KGs) that refer to the same real-world object.

By Xinran Liu, Shengtao Li, Shouqian Shi, Ge Wang, Xin-Wei Yao
llmsbenchmarkssafety
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea