AI API relay leaderboard

pangolin aggregates public community detections and ranks relays by Bayesian-weighted score (more tests + steadier scores → higher rank). Based on 28 public reports across 7 relays (6 with ≥2 tests qualify for the main board). Refreshed every 10 minutes · Want to test a relay? Back home

🏆 Main board Top 7

Sorted by Bayesian-weighted score — more tests make the score more trustworthy

7

gw.hbapis.com

1 detections · OpenAI · Last 2026-09-17 · Single sample — reference only · Full history →

Common issues: long context authenticityMessage structure specification

83 Median
How is the board scored? Click to expand the rules

Each public /r/{job_id} report feeds this board, grouped by base_url domain. Ranking uses a Bayesian-weighted score (prior=50, prior_weight=5): a one-off perfect 100 only ranks ~58.3 — you need ≥10 steady near-perfect runs to climb the main board. That stops a new relay from topping the chart after a lucky single run. The UI “Median” is the median of past scores (for intuition); sort order uses the Bayesian weight.

Single-sample relays stay off the main board (noise control) but still appear at the bottom of the full list.

2026-09-08 · Sharing, languages, and scoring tweaks

This update improves report sharing and language switching, and lightly rebalances OpenAI / Gemini scoring emphasis.