Compare AI models
Two models, one window, one asset class, one horizon. Accuracy is shown with a 95% confidence interval, because a rate without one cannot be compared to another rate.
Rule-Based Analysis
- Sample size8365
- Correct1021
- Mean stated confidence 32.1%
- Confidence gap 19.9 pts
- Symbols covered 1005 / 1016
- Coverage of the field 98.9%
- Horizons used1
| Confidence | n | Accuracy |
|---|---|---|
| 0–50% | 8013 | 11.4% |
| 50–70% | 190 | 38.9% |
| 70–85% | 42 | 14.3% |
| 85–100% | 120 | 22.5% |
Mistral Small Latest
- Sample size1051
- Correct377
- Mean stated confidence 74.5%
- Confidence gap 38.7 pts
- Symbols covered 136 / 1016
- Coverage of the field 13.4%
- Horizons used2
| Confidence | n | Accuracy |
|---|---|---|
| 0–50% | 19 | 52.6% |
| 50–70% | 377 | 44.8% |
| 70–85% | 233 | 37.3% |
| 85–100% | 422 | 26.3% |
Over the last 30 days the two confidence intervals do not overlap, so this sample does separate the models.
Accuracy counts a prediction correct when the direction it stated matches the direction the price moved past a 1% threshold over the stated horizon. Coverage is the share of symbols scored in this window that the model expressed an opinion on — a high rate over four symbols is not the same claim as the same rate over ninety. Full method and disclosures