Перейти к содержимому

AI Can't Unsee Your Data - We Measured It | FinBoundBench

Compex AI Systems

0:00 / 0:00

AI Can't Unsee Your Data - We Measured It | FinBoundBench

45 просмотров · 2 недели назад
Compex AI Systems
4 подписчика
45 просмотров · 2 недели назад
When an AI model can SEE confidential data it's not allowed to use, it uses it — in 96–100% of decisions we tested. When we removed it under contract: 0%. And doing it right cost zero accuracy. In this video, Amir (CEO) and Luca (COO) walk through FinBoundBench — a preregistered, frozen, independently recomputed study of purpose-selective AI in financial decisions, run across two model families and two financial tasks. What we found: Prohibited-but-visible data changed decisions on 100/100 pairs (primary study) and 96.5% (replication) — against noise floors of 0% and 3.7% Removing the field under a purpose contract: 0% and 0.8% — statistically at the floor Accuracy with proper purpose + data: 0.32 → 0.71. Under governed execution: still 0.71. 100% of the gain kept Every headline number reproduced by an independent implementation — 17/17 checks FinBoundBench is an open benchmark — proposed as a competition track at ICAIF 2026 (ACM International Conference on AI in Finance). Our own runtime is one baseline among several. Come beat it. Paper, protocol & frozen results: https://github.com/CompexTo/finboundb... Compex — governed AI execution: https://getcompex.com Results are specific to the tested model–task pairs. We report influence relative to each system's measured decision floor, never in absolute terms. #AIGovernance #FinBoundBench #TrustworthyAI #AI