ClearTrace Scorecard · Edition 3 of a recurring series

DEX Aggregator Execution Scorecard
Edition 3, September 2026

The rated record of DEX aggregator execution quality: quote accuracy, execution slippage, and revert rates, measured the same way for every aggregator on this board, by a third party with no routing product, and frozen so the numbers can't be edited after the fact. The board is not a complete census: venues withheld from ClearTrace's published outputs, and venues we do not yet sample, do not appear on it at all.

Chain: Ethereum  ·  Snapshot frozen: 2026-09-08 17:03 UTC  ·  On-chain data: 7-day window to 2026-09-08  ·  Published: 2026-09-08  ·  Snapshot: scorecard-edition-3.json
SHA-256: 0c36db6ecc481a1d60ed4170d79b3f4d7f4c6b385ba399742c05617de15f118c

Why a rated edition exists

Every aggregator claims best price and high reliability, and those claims are self-reported, published by the routing product being measured. Meanwhile MiCA transposes MiFID-style best-execution duties into crypto: under Regulation (EU) 2023/1114, Article 78 obliges a CASP executing client orders to take all necessary steps to obtain the best result and to evidence it on request, with record-keeping under Article 68. The regime applied from 30 December 2024, and the transitional window that let existing providers keep operating without full authorisation closed, at the latest, on 1 July 2026, so evidencing best execution is already in force. Compliance dashboards (e.g. DefiLlama's MiCA tracker) cover exchange-level obligations, not execution quality at the routing layer, where the trades actually happen.

This scorecard is a neutral, methodology-published, third-party rated record: we operate no router and take no order flow, the method is published, and each edition is numbered, dated and hashed. A rating you can cite in a governance forum, a marketing page, or a best-execution file. The snapshot below is what the hash covers, so the numbers cannot be back-dated or quietly revised; the findings text on this page is not hashed, and corrections to it ship as dated, additive notes whose full history is public in git.

Edition 3 findings

Post-publication note (2026-09-08). One sentence in the first finding is withdrawn. The first finding below says the 5.2× headline gap between OKX and SushiSwap “is substantially a difference in traffic mix.” That names a cause, and this scorecard’s own rule for findings is that they state only what is computable from the frozen snapshot: a count, a span, a rate, a decomposition, never a cause. The count that sentence rests on is already stated beside it — the residual term exceeds half the headline gap in 31 of 36 venue pairs — and that is the claim we stand behind. The finding’s own closing caveat says the residual bucket sits inside a corridor set by our classifier’s thresholds, which is exactly why it cannot carry a causal reading. Withdrawn: the “traffic mix” sentence. Retained: every number in the finding. Added for the record: the opening claim that most of what a headline revert rate counts is not the user’s own failed swap is supported by a count the page did not print — residual reverted transactions exceed genuine-user reverted transactions in 9 of 9 routers in this snapshot. Per this scorecard’s integrity model the snapshot, the findings text and the SHA-256 hash are unchanged; corrections are published as dated notes like this one.
Post-publication note (2026-09-08). OpenOcean’s quote-accuracy cell should read stale, not rated. The rated table below shows OpenOcean at 10.23 bps over 6,983 realized quote samples with a rated badge. The figure is real. The badge is not: the last successful OpenOcean quote sample on Ethereum was captured 2026-08-20 19:05 UTC, and every request since — 1,716 of them, at roughly 104 a day — returned no quote, because the venue’s endpoint began answering our sampler with a bot challenge on 21 August. The table’s printed rule for rated (≥30 realized samples spanning ≥7 days) is met on frozen history; the recency rule we apply in code (newest successful sample no older than 3 days) is failed by 16 days. That recency rule landed the same day this edition froze, and the edition was built from a data export eight minutes older than the first one to carry the corrected badge. A cell that has gone quiet for 19 days is not a rated cell. Read the 10.23 bps as OpenOcean’s last measured value, not a current one, and read the “widest quote gap on the board” as a statement about August. Per this scorecard’s integrity model the snapshot, the findings text and the SHA-256 hash are unchanged; Edition 4 carries this correction forward in its errata.

Provenance: Bitget DEX and Bebop are named above, and we have not ingested their own published deployment addresses. Their figures rest on our labelling of which contracts belong to them, and should be read as provisional.

The rated table

Rated = ≥30 realized fork-simulation quote samples spanning ≥7 days. On-chain metrics (slippage, reverts) cover every trade/transaction we attribute to the venue in the window, not a sampled subset of them, but attribution is entrypoint-framed, so the denominator is the flow entering through contracts we can tie to the venue, not the venue's total volume. Baseline = the direct-venue default aggregators are implicitly compared against. Revert rate shows the headline with the v6 genuine-user rate in parentheses; n/c marks a venue whose settlement model (batch auction, intent, or unclassified where we have not established who submits the settlement transaction) makes its on-chain revert rate not comparable to a router's. A venue marked wound down has shut down: its numbers are frozen history for the window measured, not a live venue. A venue marked unestablished is one whose own published deployment addresses we have not ingested: its numbers rest on our labelling rather than on the venue's list, and should be read as provisional. An execution-cost cell marked all-chain means the venue has no row on this chain, so the figure is its median across every chain we measure, shown as context and excluded from every finding. A dash (—) means we hold no measurement for that cell this edition, whether the venue had no qualifying traffic, fell below the query's reporting floor, or is not yet quote-sampled; it never means zero. Router counts differ between findings because each states its own pool: the residual decomposition includes the venue default, the genuine-user comparison does not.

AggregatorStatusModelExec slippageRevert rate Quote gapQuote nRouting txs
Bitget DEX · unestablishedon-chain onlyRouter4.14 bps1.72% (0.69% users)35,895
1inchratedRouter7.45 bps1.16% (0.23% users)0 bps12,021126,629
OpenOceanratedRouter9.34 bps2.16% (0.32% users)10.23 bps6,9838,410
CoW Protocolon-chain onlyBatch auction9.66 bps0.22% n/c27,258
Bebop · unestablishedratedUnclassified10.74 bps2.61% n/c0 bps8,51218,952
Uniswap (direct venue)baselineRouter10.93 bps2.61% (1.04% users)0 bps11,503145,635
KyberSwapratedRouter11.48 bps1.26% (0.62% users)0.26 bps11,560294,054
1inch Limit Order Protocolon-chain onlyIntent11.71 bps0.48% n/c47,754
SushiSwapratedRouter11.8 bps0.96% (0.3% users)0 bps11,08153,299
ParaSwap (Velora)ratedRouter19.52 bps1.58% (0.34% users)1 bps11,58533,133
OKXratedRouter20.5 bps5.02% (1.03% users)2.5 bps9,87539,167
Tokenlon · unestablishedon-chain onlyUnclassified65.73 bps0% n/c2,461
Odos · Wound down 2026-07-30 · data through Jul 2026ratedRouter0 bps3,109
LI.FIratedRouter2.36% (0.43% users)0.71 bps8,96645,584
DODO Xon-chain onlyUnclassified0.12% n/c9,625

Verify this edition

The findings above are computed from a frozen snapshot. To verify nothing has changed since publication, hash the snapshot and compare:

curl -s https://cleartracedata.com/static/data/scorecard-edition-3.json | shasum -a 256
0c36db6ecc481a1d60ed4170d79b3f4d7f4c6b385ba399742c05617de15f118c

The snapshot is committed to a public git history at publication time, which independently timestamps it.

Methodology & scope

Full methodology: cleartracedata.com/methodology and the open-dataset README. In brief: execution slippage is the median per-fill gap vs a 1-minute VWAP oracle over all on-chain trades; revert rate counts failed routing transactions from raw on-chain data; quote accuracy is a forward-captured fork-simulation sample (quoted vs realized). Aggregators we don't yet quote-sample appear with on-chain metrics only. Preliminary cells are shown but never rated. Quote-sample windows differ by venue (each venue is rated once it has enough samples of its own), so the quote column is not a single shared window the way the on-chain columns are. Scope: Ethereum for this edition; the live leaderboard is the free preview layer that accrues between editions.

Cite or commission

Cite this edition as: ClearTrace DEX Aggregator Execution Scorecard, Edition 3 (September 2026), sha256:0c36db6ecc48…. Link: https://cleartracedata.com/scorecard/edition-3. For a private cut of your own routing versus peers (per-pair, per-size, with the failure modes), or a subscription to future rated editions for a best-execution evidence file, book a working session.

ClearTrace is an independent measurement firm: we operate no router, no frontend, and take no order flow. Editions are immutable once published; corrections, if ever needed, ship as an errata note in the next edition, never as a silent edit. · Home · Live leaderboard · Research