← Back
Finding 7017Emerging EvidenceValidation V0

SECQUE introduces a benchmark using real SEC filings with 565 expert-written questions across four categories, showing even GPT-4o struggles with complex reasoning and analyst insights, spotlighting evaluation challenges in finance.

86%Confidence
1Evidence objects
v1Version
DraftStatus

Evidence trail

Knowledge status

This Finding was extracted from the configured corpus. It is versioned, traceable, and may evolve through editorial review or new corpus evidence.