arXiv AI

Augmenting Fundamental Analysis with Large Language Models: A RAG-Based System for Generating Investor Briefs

arXiv:2607. 09121v1 Announce Type: cross Abstract: In this study, we examine the opportunities brought by Large Language Models (LLMs) to various aspects of fundamental analysis of companies based on their reports as well as data and documents describing macroeconomic situation like GDP and inflation changes as well as documents filled to the U.

arXiv Computation and Language
Sep 4

The Analyst in the Prompt: Role, Retrieval, and Memory Biases in LLM Financial Analysis

The study examines how user context—such as memory, profiles, and role prompts—affects Large Language Models’ (LLMs) financial analysis. Using 3,575 SEC filings and twelve LLMs, the authors distinguish between evidence selection and interpretation, finding that most context spillover arises from differing interpretations under various roles rather than from retrieving different evidence. They evaluate two mitigation strategies—expressing investor mindset as a user profile instead of an assistant role, and separating evidence-based from personalized outputs—both of which reduce but do not eliminate spillover, with effectiveness varying across models.

By Ahmed Asaad, Amr Mohamed, Yang Zhang, Omneya Abdelsalam
arXiv Machine Learning
Sep 11

AI Economist Agent: An Agentic Framework for Evidence-Based Economic and Financial Analysis with RAG, Knowledge Graphs, and Large Language Models

The paper introduces an AI economist agent that integrates large language models, retrieval‑augmented generation, knowledge graphs, and quantitative models to conduct evidence‑based economic and financial scenario analysis. The framework orchestrates LLM agents to plan analyses, retrieve relevant evidence, and structure economic mechanisms, while registered quantitative models produce numerical outcomes and predefined tests validate intermediate results for inclusion in the final report. Applied to European macro‑financial stress scenarios and bank capital analysis, the empirical study demonstrates the agent’s ability to combine flexible evidence retrieval and scenario construction while maintaining traceability to sources and explicit model calculations.

By Masahiro Kato
Hugging Face Trending Papers
Sep 2

The Analyst in the Prompt: Role, Retrieval, and Memory Biases in LLM Financial Analysis

The paper investigates how user context—such as memory, profiles, and role prompts—affects large language models’ financial analysis. By testing 3,575 SEC filings across twelve LLMs, the study distinguishes between evidence selection and interpretation, finding that interpretation under different roles drives most user-context spillover. Two mitigation strategies—using a user profile instead of an assistant role and separating evidence-based from personalized outputs—reduce but do not eliminate this spillover, with effectiveness varying by model.

Hugging Face Trending Papers
Jun 22

IPO Finance Agent: Evaluation of LLM Financial Analysts beyond Finance Agent v2, with Automated Rubric Generation -- the Case of the SpaceX (SPCX) IPO

Finance Agent v2 (by Vals AI) has emerged as the reference benchmark for evaluating both Anthropic Claude and OpenAI ChatGPT frontier language models on financial tasks. However, it narrowly deals with periodic reporting from publicly traded companies (SEC 10-K and 10-Q filings), and its agentic harness relies on naive, unenriched chunk retrieval.

arXiv AI
Sep 10

IGT @ FinMMEval 2026 Task 2: Question-Type Prompting with Targeted Extraction for Multilingual Financial QA

The IGT system tackles PolyFiQA Task 2 of the FinMMEval Lab, a multilingual financial QA challenge involving English SEC filings and news in five languages. It distinguishes two question families: numeric‑structured queries are answered via keyword extraction from filings, while synthesis queries use rule‑based passage selection from news. The approach yields a development ROUGE‑1 of ~0.395, a 60% boost over a generic RAG baseline, and places third among twelve teams on the official test set.

By Yuwen Chiu (Georgia Institute of Technology)
arXiv AI
Jun 12

Fin-RATE: A Real-world Financial Analytics and Tracking Evaluation Benchmark for LLMs on SEC Filings

arXiv:2602. 07294v4 Announce Type: replace-cross Abstract: With the increasing deployment of Large Language Models (LLMs) in the finance domain, LLMs are increasingly expected to parse complex regulatory disclosures.

By Yidong Jiang, Junrong Chen, Eftychia Makri, Jialin Chen, Peiwen Li, Ali Maatouk, Leandros Tassiulas, Eliot Brenner, Bing Xiang, Rex Ying