test.apiqik-kratos.apiqik.com
Common issues: Extreme reasoning ability (number theory counting)、 Think supportive
pangolin aggregates public community detections and ranks relays by Bayesian-weighted score (more tests + steadier scores → higher rank). Based on 28 public reports across 7 relays (6 with ≥2 tests qualify for the main board).
Sorted by Bayesian-weighted score — more tests make the score more trustworthy
Common issues: Extreme reasoning ability (number theory counting)、 Think supportive
Common issues: Extreme reasoning ability (number theory counting)、 Minimalist Input Token Audit、 Token billing
Common issues: Extreme reasoning ability (number theory counting)、 Claude Tokenizer verification
Common issues: Extreme reasoning ability (number theory counting)、 Think supportive
Each public /r/{job_id} report feeds this board, grouped by base_url domain. Ranking uses a Bayesian-weighted score (prior=50, prior_weight=5): a one-off perfect 100 only ranks ~58.3 — you need ≥10 steady near-perfect runs to climb the main board. That stops a new relay from topping the chart after a lucky single run. The UI “Median” is the median of past scores (for intuition); sort order uses the Bayesian weight.
Single-sample relays stay off the main board (noise control) but still appear at the bottom of the full list.