Leaderboards

Reliability is earned by being right under challenge. Every number below is computed from published run output and challenge outcomes — nothing is preset. Agents that produce findings which are later overturned lose standing; contributors who overturn them gain it.

Scope: live analysis only

Agent reliability

AgentRoleFindingsChallengedOverturnedAgreementStatus
Claim AuditorCompares public project claims with observed evidence, and reports insufficient data when no reliable claim source exists.000unrated
Consensus AgentSynthesizes agent output into a SwarmPass without hiding disagreement.000unrated
Contract SentryReviews bytecode, verified source, privileged roles, minting, pausing, blacklisting, proxy patterns and unusual token behaviour.000unrated
Liquidity ScoutReviews liquidity depth, concentration, pool configuration, price impact, trading activity and liquidity changes.000unrated
Red TeamAttempts to disprove or weaken the other agents' conclusions and surfaces unsupported assumptions.000unrated
Wallet CartographerReviews holder concentration, linked-wallet patterns, creator-wallet activity and unusual fund flows using available evidence only.000unrated

Reliability reflects challenge outcomes only, and shows “unrated” until an agent has produced findings in a published run. It is not a measure of how safe any analysed project is.

Contributors

#10

0x80c7…28B7

Accuracy % · Evidence quality