CIFQA: A Deterministic Tool-Grounded Multi-Agent LLM Framework for Financial Query Answering

arXiv نُشر في تم التحديث AI & Machine Learning
سجّل الدخول للحفظ

الأصول والمواضيع المتأثرة

LLM

نبرة المقال

محايد كيف كُتب المقال، وفق ما ذكره المصدر.

التأثير المتوقع على السوق

محايد الثقة 50% كيف تُقرأ نسبة الثقة الأفق الزمني: المدى القصير الأثر: الأدنى

مسار الأدلة

الأدلة
المصدر arXiv
الادّعاء CIFQA: A Deterministic Tool-Grounded Multi-Agent LLM Framework for Financial Query Answering
استنتاج الذكاء الاصطناعي Neutral · 50%
أُنشئ في 2026-08-28 04:00

مصدر التحليل بالذكاء الاصطناعي

حُلِّل بواسطة Free Analysis Rule Based Analysis ليس ذكاءً اصطناعياً المنهجية v1.0 أُنشئ في
المعرّفات التقنية
وسم المزوّد
free-analysis-rule-based-analysis
إصدار التحليل
free-analysis-rule-based-analysis
معرّف المقال
123307

المصدر الأصلي

arXiv:2608.26114v1 Announce Type: new Abstract: Calculation-intensive financial question answering requires exact reasoning over structured rates, temporal conditions, numerical formulas, and rule-based constraints. Although Large Language Models (LLMs) perform strongly on natural language tasks, they often produce numerically incorrect yet plausible answers when solving multi-step financial calculations. To address this limitation, we introduce CIFQA (Calculation-Intensive Financial Query Answering), a deterministic tool-grounded multi-agent LLM framework for financial question answering. CIFQA separates language understanding from numerical execution by assigning specialized agents to query interpretation, routing, parameter extraction, computation planning, and response generation, while deterministic Python-based tools perform financial calculations and rule application. We instantiate CIFQA for fixed deposit query answering and evaluate it on a curated benchmark of fixed deposit queries. CIFQA achieves 95.54% accuracy on calculation-intensive queries and 90.87% overall accuracy, substantially outperforming direct LLM baselines even when provided with complete formulas, rate cards, and benchmark instructions. Ablation studies show that deterministic components such as exact rate lookup, tenure computation, rolling-year adjustment, and premature-withdrawal logic are critical contributors to performance. Notably, a 17B open-source backbone operating within CIFQA outperforms substantially larger frontier models evaluated with the same financial information, demonstrating that architectural design is a more important determinant of numerical reliability than model scale. While evaluated on fixed deposit queries, CIFQA provides a generalizable framework for calculation-intensive financial reasoning tasks.

اقرأ المقال كاملاً على arXiv

المقال الأصلي منشور بواسطة arXiv في أغسطس 28, 2026. التحليل والرؤى المقدمة من AnalystMarkets AI.

تغطية ذات صلة