Castellum.AI’s Financial Crime Benchmark Finds Accuracy & Consistency Gaps in Standalone LLMs
Castellum.AI announced its FinCrime Agent Benchmark (FCA Bench) in September that compares the performance of nine standalone large language models (LLMs) performing financial-crime checks and when those same models operate within the Castellum.AI Harness. The benchmark measures risk identification accuracy (sanctions, politically exposed persons (PEPs) and adverse media screening), applying an institution’s standard operating procedures...




























