Official benchmark
Model Accuracy — Official Benchmark
The official accuracy comes from one fixed benchmark: the same 18 stocks, predicted once per NYSE session at 16:15 New York time and scored at the next session's official close. Every call stays on the record — right, wrong or pending.
16:15 ET, every NYSE sessionNo deletions, no rewrites
What is — and is not — counted
Only Benchmark predictions are scored here. Predictions you request any time on the website, the REST API, the Python SDK or MCP are separate and never enter this figure, no matter how many there are.
Accuracy
…
awaiting first scored session
Brier score
…
lower is better · 0.25 = coin flip
Evaluated
…
Correct: 0
Pending
…
Total predictions: 0
"Always up" baseline
…
same sessions, guessing up every time
Evaluation period
…
Latest run
…
void
…
0 void (corporate action, not scored)
Per-stock accuracy
| Ticker | Total predictions | Evaluated | Pending | Correct | Accuracy | Brier score |
|---|---|---|---|---|---|---|
| … | ||||||
Calibration
When the model says X% chance of an up close, how often does it actually close up?
| P(up) range | n | Avg predicted | Actual up rate |
|---|
Recent benchmark predictions(every call, including misses and pending)
| Predicted on | Target session | Ticker | Call | P(up) | Reference close | Target close | Result | Generated (ET) |
|---|
How the benchmark works
- Universe: the AI Stocks and Semiconductor Stocks lists from our US markets page — 18 unique tickers (NVDA and AMD are in both and count once).
- Timing: a scheduled job runs at 16:15 America/New_York on every NYSE session, after the close, using that day's complete data. Predictions are accepted only inside 16:15–16:45; a missed window is never filled in later.
- Target: the next real NYSE session from the exchange calendar — Friday predicts Monday, and a holiday Monday moves the target to Tuesday.
- Scoring: after the target session's official close, the close is compared with the reference close. Up = higher; an unchanged close counts as not up. Brier score = (P(up) − outcome)².
- Audit: a prediction cannot be edited or deleted once written — the database refuses it. A stock split between the two closes makes the pair incomparable; that row is kept and marked void.
- The benchmark began clean on launch day. No earlier predictions were copied into it.
Past accuracy does not guarantee future results. This page is for informational purposes only and is not investment advice.