arXiv AI By Sunghwan Kim, Wooseok Jeong, Serin Kim, Sangam Lee, Dongha Lee

SAGEO Arena: A Realistic Environment for Evaluating Search-Augmented Generative Engine Optimization

Read the original on arXiv AI →

arXiv:2602. 12187v2 Announce Type: replace-cross Abstract: Search-Augmented Generative Engines (SAGE) have emerged as a new paradigm for information access, bridging web-scale retrieval with generative capabilities to deliver synthesized answers.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Jun 2

Characterizing Web Search in The Age of Generative AI

arXiv:2510. 11560v2 Announce Type: replace-cross Abstract: The advent of LLMs has given rise to generative search, a new search paradigm in which LLMs retrieve information from the web related to a query and synthesize it into a single, coherent response.

By Elisabeth Kirsten, Jost Grosse Perdekamp, Qinyuan Wu, Mihir Upadhyay, Krishna P. Gummadi, Muhammad Bilal Zafar
arXiv AI
Sep 24

Query Implied Generative Engine Optimization

The paper "Query Implied Generative Engine Optimization" introduces QI‑GEO, a method that infers user intent directly from documents to enhance visibility in Generative Search Engines. By approximating a document’s intent space, QI‑GEO identifies missing yet relevant content, improving objective scores by up to 15.9% and subjective scores by up to 17.6% on GEO‑Bench datasets. The approach yields nearly twice as many citation gains as losses, demonstrating that document‑derived intent approximations can boost content visibility without explicit query inputs.

By Shilpa Ramakrishna, William B. Andreopoulos
arXiv AI
Sep 1

Agent2UCB: Agentic System for Generative Engine Optimization

Agent2UCB is a new agentic system designed for Generative Engine Optimization (GEO), which refines content to boost its likelihood of being cited or summarized by generative AI search engines. The system autonomously evaluates nine GEO strategies for each content item, selects the most effective one, and speeds up this selection using a bandit-based Agent2UCB policy that blends large language model priors with real-time reward signals. Additionally, it offers a lightweight, text-only SEO readiness check that assesses readability, topical coverage, and EEAT-style credibility, and experiments on GEO-Bench demonstrate consistent visibility gains while maintaining SEO quality.

By Sheldon Yu, Rui Wang, Tong Yu, Sungchul Kim, Doga Dogan, Junda Wu, Julian McAuley
Hugging Face Trending Papers
Sep 8

Q2D-Web: A Large-Scale Benchmark for Retrieval in Agentic RAG Systems

Q2D-Web is a new large‑scale benchmark for agentic Retrieval‑Augmented Generation (RAG) systems, featuring a 190 million‑document web corpus and 70 k machine‑reformulated search queries in ten languages. It supplies three sets of relevance judgments—agent citations, production rankings, and a combined set enriched with LLM‑based labels—to evaluate first‑stage retrievers. Experiments on 13 retrievers show consistent ranking across judgment sets but significant variation across domains, languages, and query types, and demonstrate that a carefully sampled sub‑corpus can approximate full‑corpus evaluation with minimal loss in Recall@1000.