When this trader buys at $0.70, they imply 70% probability. Perfect calibration = the event happens 70% of the time.
| Bucket | Bets | Expected | Actual | Error |
|---|---|---|---|---|
| 0.00-0.10 | 20 | 5% | 45% | 39.9% |
| 0.10-0.20 | 4 | 14% | 25% | 11.2% |
| 0.20-0.30 | 10 | 24% | 10% | 14.1% |
| 0.30-0.40 | 16 | 35% | 6% | 29.2% |
| 0.40-0.50 | 38 | 46% | 0% | 45.6% |
| 0.50-0.60 | 38 | 54% | 0% | 54.3% |
| 0.60-0.70 | 36 | 65% | 6% | 59.1% |
| 0.70-0.80 | 22 | 75% | 9% | 66.0% |
| 0.80-0.90 | 18 | 85% | 17% | 68.1% |
| 0.90-1.00 | 20 | 96% | 40% | 56.0% |
Skill measures calibration quality (0-100). Variance measures return volatility (0-100, higher = more volatile).
On-chain verification: wallet age 26 days, 2645 txs, provenance grade C. Bot score: 0/100, wash trading score: 20/100.
Polymarket on-chain coverage: $23,360 in / $0 out across 21 withdrawal tx since 2026-03-24.
3100 total trades across 208 markets.
222 bets on resolved markets available for calibration scoring.
Calibration error: 50.3% — needs improvement.
Skill: 4/100 (calibration quality). Variance: 42/100 (higher = more volatile returns).
Brier Skill Score: -234.2% vs naive baseline (>0% = better than always predicting base rate).
Brier decomposition: REL=0.2733 RES=0.0233 UNC=0.1068.
Log loss: 1.0373 (skill: -180.3% vs naive). Lower log loss = better calibration on rare events.
Below average. The data shows poor calibration, thin evidence, or both. When this trader expresses high confidence, events don't happen at the rate they imply.
Confidence: D/40 ± 3 (high confidence, 222 resolved bets). This score is highly reliable — enough resolved bets to be confident.
Methodology: Brier Score Decomposition (Murphy 1973), Log Loss, On-Chain USDC Verification. Same approach used by IARPA to identify superforecasters.