Finding 2681Emerging EvidenceValidation V0
Surprisingly, even top AI models scored below 60% on the detailed rubric and under 45% on final-answer accuracy, exposing a significant gap between current AI capabilities and professional financial analysis standards.
86%Confidence
1Evidence objects
v1Version
DraftStatus
Evidence trail
Supporting86% linkage confidence
Surprisingly, even top AI models scored below 60% on the detailed rubric and under 45% on final-answer accuracy, exposing a significant gap between current AI capabilities and professional financial analysis standards.
key_findings bullet 2 · key_findings
Inspect source: BigFinanceBench: A Workflow-Grounded Benchmark for Financial-Research Agents →This Finding was extracted from the configured corpus. It is versioned, traceable, and may evolve through editorial review or new corpus evidence.