arXiv Machine Learning By Hyunseok Paeng

The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection

Read the original on arXiv Machine Learning →

arXiv:2606. 09204v1 Announce Type: new Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation -- the Injection Paradox -- in which prompt injections embedded in retrieved documents backfire against the attacker, suppressing the target brand below the injection-free baseline.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Aug 25

One Polluted Page Is Enough: Evaluating Web Content Pollution in LLM Recommenders

The paper introduces FORGE, a benchmark that rewrites real product pages into fake ones to test how often search‑augmented large language models (LLMs) recommend these polluted items. Across 12 commercial and open‑weight LLMs, a single polluted page can lead to up to 27% of recommendations being fake, rising to 73.8% when the top‑3 replacements are used. The study finds that reasoning does not help and existing defenses—skepticism prompts, consensus filters, and credibility re‑ranking—are largely ineffective.

By Minghao Luo, Liang Chen
arXiv Machine Learning
Sep 3

Training seeds and model-selection stability in recommender-system evaluation

The paper investigates how the choice of random training seed affects recommender‑system experiments. By fixing the data split and varying seeds across hyperparameter settings, the authors analyze seed effects on user‑level metrics, validation‑based model selection, and recommendation‑list agreement. Their findings show that seed variation can be detectable and its impact depends on configuration separation, validation‑to‑test transfer, and top‑k list similarity, indicating that single‑seed results may overstate evaluation stability.

By Juan Manuel Rodriguez, Oleg Lesota, Antonela Tommasel
arXiv Machine Learning
Sep 25

Who Owns the AI Recommendation? A Multi-Industry Empirical Map of Brand Category Ownership Across Large Language Models

The study examines how large language models (LLMs) like GPT‑5.2, Gemini 3 Flash, and Perplexity sonar‑pro recommend brands across five industries. Using 50 brands and 250 queries repeated five times, the authors measured brand inclusion, recommendation share, competitive vacuum, and co‑mention asymmetry, finding that most queries mention at least one brand and that vacuum prevalence remained stable between February and September 2026. The analysis shows strong cross‑date consistency in recommendation patterns and no emergent clustering of brand mentions, though co‑mention structures deviate from null expectations.

By Dmitrij \.Zatuchin