Part of event: OpenAI’s Astra Model: Text Arena Debut? · openais-astra-model-text-arena-debut
Market structure and 24h change as reported by Polymarket's public API.
Top of book
Best bid
0.600
Best ask
0.610
Spread
1.7% of mid
1.00 pts
As reported by Polymarket's public API at 05:20 UTC; top of book, not full depth; per-side depth not shown. The spread is shown against the midpoint of the two quotes above, which is arithmetic on the venue's figures, not a figure the venue reports.
Volume
24h volume
$7,867
Total volume
$29,988
Volume as reported by Polymarket's public API.
Resolution rules
Rules as stated by Polymarket, reproduced verbatim. Convexly has not independently reviewed them and does not endorse them; the rating's resolution-reliability pillar treats rule text as present but not independently reviewed.
On August 1, 2026, OpenAI publicly disclosed “Astra,” describing an internal version of Astra as its “next major model.” You can read more about that here: https://openai.com/index/ten-advances-in-mathematics/. This market will resolve to "Yes" if the next OpenAI “Astra” model added to the Arena.AI Leaderboard (https://arena.ai/leaderboard/text/overall-no-style-control) has at least the specified score at 12:00 PM ET on the calendar date following the date on which it first appears on the leaderboard. Otherwise, this market will resolve to "No". A qualifying model must be named "Astra" (including variants or successor branding such as Astra 1, Astra 6, Astra X, GPT-6 Astra, or similar naming or rebranded versions), or be confirmed to be the same model referenced in the announcement by OpenAI or by a consensus of credible reporting. Results from the "Score" column under the "Text Arena | Overall" Leaderboard tab at https://arena.ai/leaderboard/text/overall-no-style-control with style control off will be used to resolve this market. This market will resolve solely based on the specified score in the Score column of the leaderboard, regardless of any underlying granular or unrounded data presented elsewhere. A model marked “AutoEval” will not be considered added to the leaderboard. Only scores displayed without the “AutoEval” label will be considered. If multiple models are added to the leaderboard on the same calendar date (ET), the highest-scoring model will be used for resolution. Models added to the leaderboard on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Arena.AI Leaderboard. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the leaderboard is irrelevant for this market. The resolution source for this market is the Chatbot Arena LLM Leaderboard found at https://arena.ai/leaderboard/text/overall-no-style-control. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the leaderboard, this market will resolve based on the first subsequent instance at which such a score becomes available on the leaderboard. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the leaderboard or if no qualifying model release occurs by December 31, 2026, 11:59 PM ET, this market will resolve to "No".Show the full rules textHide the full rules text
On August 1, 2026, OpenAI publicly disclosed “Astra,” describing an internal version of Astra as its “next major model.” You can read more about that here: https://openai.com/index/ten-advances-in-mathematics/. This market will resolve to "Yes" if the next OpenAI “Astra” model added to the Arena.AI Leaderboard (https://arena.ai/leaderboard/text/overall-no-style-control) has at least the specified score at 12:00 PM ET on the calendar date following the date on which it first appears on the leaderboard. Otherwise, this market will resolve to "No". A qualifying model must be named "Astra" (including variants or successor branding such as Astra 1, Astra 6, Astra X, GPT-6 Astra, or similar naming or rebranded versions), or be confirmed to be the same model referenced in the announcement by OpenAI or by a consensus of credible reporting. Results from the "Score" column under the "Text Arena | Overall" Leaderboard tab at https://arena.ai/leaderboard/text/overall-no-style-control with style control off will be used to resolve this market. This market will resolve solely based on the specified score in the Score column of the leaderboard, regardless of any underlying granular or unrounded data presented elsewhere. A model marked “AutoEval” will not be considered added to the leaderboard. Only scores displayed without the “AutoEval” label will be considered. If multiple models are added to the leaderboard on the same calendar date (ET), the highest-scoring model will be used for resolution. Models added to the leaderboard on the calendar date following the initial qualifying model’s first appearance will not be considered. A qualifying model must be newly added to the Arena.AI Leaderboard. Whether the model was previously released, publicly accessible, in beta, or otherwise available before appearing on the leaderboard is irrelevant for this market. The resolution source for this market is the Chatbot Arena LLM Leaderboard found at https://arena.ai/leaderboard/text/overall-no-style-control. If this resolution source is unavailable at 12:00 PM ET on the calendar date following the date on which the qualifying model first appears on the leaderboard, this market will resolve based on the first subsequent instance at which such a score becomes available on the leaderboard. If it remains unavailable through the end of the seventh day after the qualifying model first appears on the leaderboard or if no qualifying model release occurs by December 31, 2026, 11:59 PM ET, this market will resolve to "No".
Resolution mechanics
Resolution parameters as stated by Polymarket.
- Resolved by (UMA adapter)
- 0x65070BE91477460D8A7AeEb94ef92fe056C2f2A7
- Proposer bond
- $500 proposer bond
- Proposer reward
- $5 proposer reward
- Stated resolution source
- https://arena.ai/leaderboard/text/overall-no-style-control
Market Trust diagnostic · v0.2 candidate
Will OpenAI’s Astra model debut on the Arena Leaderboard at a score of at least 1490?
This market is not in the v0.1 daily card set yet. The v0.2 diagnostic scorecard below shows the per-pillar contribution from the latest materializer run, with visible caveats and evidence-status separation.
Why this rating
Why this rating
Composite 53 · Do not cite
Measurement certainty
Low
1 measured · 3 modeled · 2 not assessed
Certainty reflects how much of this market we measured, not how good the market is.
- CoherenceModeled estimatefixed rules, not a per-market measurementweight 25.0
No full-provenance CME signal is linked to this condition in the current ledger window; absence of a signal is not proof of coherence.
- · no linked cme signal in window
- · absence of signal not proof of coherence
- · CME signal window: 7 days
- · recent full provenance signals: 15
50+17.9 - ExecutabilityModeled estimatefixed rules, not a per-market measurementweight 20.0
Executability scored as worse-of(spread, depth) from a single orderbook snapshot; the two were merged from the former liquidity + depth pillars to stop double-counting one correlated source.
- · executability takes the worse of spread and depth
- · spread: top-of-book spread 1.0c
- · spread: spread is 3% of mid, tight relative to price
- · spread: depth approximated from a probability-point band and split evenly per side; a per-side breakdown is not yet available
15+4.3 - Resolution reliabilityModeled estimatefixed rules, not a per-market measurementweight 20.0
Resolution source and rule text are present but not independently reviewed.
- · resolution metadata present manual review pending
70+20.0 - Participant flowNot assessednot yet measured for this marketweight 20.0
Participant flow inputs exist but do not meet identity-coverage floors; descriptive only.
- · below publishable sample floor
- · high participant concentration
- · high unknown flow share
- · insufficient scored accounts
not assessed - Integrity-risk screenNot assessednot yet measured for this marketweight 10.0
One or more screen indicators flagged at a low level. A reason to look, not a finding about anyone.
- · top-3 wallets control 100% of notional flow
- · flow events: 1
- · distinct wallets: 1
- · wash round trips: 0
75 - Audit completenessMeasuredweight 5.0
3 linked artifacts; 0 missing hashes; 0 stale/failing sources.
- · artifact hashes present
- · source watermarks healthy
100+7.1
How it adds up: composite = Σ(score × weight) / Σ(weight) across measured + heuristic pillars only. Pending and data-rights-blocked pillars are excluded from both numerator and denominator. Total contributing weight = 70.0 of 100.0; the remaining 30.0 is unmeasured.
Caveat: 2 pillars not assessed yet; the score will recompute as the substrate fills in.
Structural coherence read
The Coherent Markets Engine checks this market's pricing for structural inconsistencies against related markets. Whether a current signal exists, and its construction, is included with Researcher.
See ResearcherMarket opened 2026-08-13 · minimum order size 5 shares · price tick size 0.010 · as stated by Polymarket.