Research

Advancing the science of agent trust.

AI agents are beginning to spend money on behalf of humans and organizations. Before this becomes the default mode of commerce, someone has to answer a fundamental question: how do you know an agent is making good economic decisions? We’re building the instruments to find out — and we publish openly, because the frameworks need scrutiny to become standards.

Focus areas

Three questions drive the program.

01Decision quality measurementProcess-based metrics that evaluate the quality of an agent’s economic reasoning independent of outcome — a good decision can have a bad result, and a bad decision can get lucky.
02Trust & authorizationLayered authorization frameworks that let networks, issuers, and merchants verify agent identity and intent before authorizing transactions on existing rails.
03Failure mode analysisMapping how AI economic reasoning degrades — from Goodhart’s Law gaming to adversarial manipulation — and building detection that catches problems before they propagate.
Publications

Working papers.

EDQS benchmark

Get the benchmark when it publishes.

We are benchmarking EDQS against transaction-fraud baselines on agent-initiated transactions. Leave your email and receive the results the day they go live.

Research updates only. No marketing.