CIFQA: A Deterministic Tool-Grounded Multi-Agent LLM Framework for Financial Query Answering

arXiv Published Updated AI & Machine Learning
Sign in to save

Affected assets and topics

LLM

Article tone

Neutral How the article is written, as reported by the source.

Expected market reaction

Neutral Confidence 50% How confidence is read Horizon: Short term Impact: Low

Evidence trail

Evidence
Source arXiv
Claim CIFQA: A Deterministic Tool-Grounded Multi-Agent LLM Framework for Financial Query Answering
AI inference Neutral · 50%
Generated 2026-08-28 04:00

AI provenance

Analysed by Free Analysis Rule Based Analysis not AI Methodology v1.0 Generated
Technical identifiers
Provider tag
free-analysis-rule-based-analysis
Analysis version
free-analysis-rule-based-analysis
Article id
123307

Original source

arXiv:2608.26114v1 Announce Type: new Abstract: Calculation-intensive financial question answering requires exact reasoning over structured rates, temporal conditions, numerical formulas, and rule-based constraints. Although Large Language Models (LLMs) perform strongly on natural language tasks, they often produce numerically incorrect yet plausible answers when solving multi-step financial calculations. To address this limitation, we introduce CIFQA (Calculation-Intensive Financial Query Answering), a deterministic tool-grounded multi-agent LLM framework for financial question answering. CIFQA separates language understanding from numerical execution by assigning specialized agents to query interpretation, routing, parameter extraction, computation planning, and response generation, while deterministic Python-based tools perform financial calculations and rule application. We instantiate CIFQA for fixed deposit query answering and evaluate it on a curated benchmark of fixed deposit queries. CIFQA achieves 95.54% accuracy on calculation-intensive queries and 90.87% overall accuracy, substantially outperforming direct LLM baselines even when provided with complete formulas, rate cards, and benchmark instructions. Ablation studies show that deterministic components such as exact rate lookup, tenure computation, rolling-year adjustment, and premature-withdrawal logic are critical contributors to performance. Notably, a 17B open-source backbone operating within CIFQA outperforms substantially larger frontier models evaluated with the same financial information, demonstrating that architectural design is a more important determinant of numerical reliability than model scale. While evaluated on fixed deposit queries, CIFQA provides a generalizable framework for calculation-intensive financial reasoning tasks.

Read the full article on arXiv

Original article published by arXiv on August 28, 2026. Analysis and insights provided by AnalystMarkets AI.

Related coverage