New benchmark tests AI reasoning over enterprise financial records

ENTLORE benchmark evaluates how well AI systems recover implicit organizational relationships across enterprise documents—a key challenge for automated financial analysis and reporting.

18 more from arXiv: AI + Accounting

Study questions embedding-similarity gates in AI agent systems

Research validates how AI agent frameworks use text-embedding cosine similarity thresholds as quality gates for deduplication and semantic caching, finding measurement gaps between semantic meaning...

Large Language Models Automate IT Audit Control Evaluation

IntelliAudit uses retrieval-grounded multi-agent LLMs to automate IT audit evidence evaluation by synthesizing heterogeneous organizational data across policies and records—addressing the semantic ...

FinRank benchmark tests AI accuracy on SEC filing questions

New research introduces FinRank, a benchmark for evaluating AI systems on financial question answering over SEC filings, focusing on evidence provenance rather than answer correctness alone.

Scoping Review Finds System Integration Critical Gap in AI Audits

Academic scoping review of 4,259 documents reveals AI audits focused on individual models miss integration risks—drawing lessons from aerospace safety-critical auditing for accounting and complianc...

AI verification tax consumes 15+ hours weekly for half of finance pros

Study reveals accounting firms investing in AI spend 15-30+ hours weekly verifying AI work, contradicting job-loss narratives and highlighting hidden compliance costs of automation.

Stay ahead of AI in accounting

Get the latest news on agentic AI for accounting, audit, and tax delivered to your inbox. Curated by AI, reviewed by professionals.

Subscribe to Newsletter