Reporting tells you what happened. This page tells you what we are waiting to find out, and keeps a permanent record of what we got right and wrong once the answer arrives.
Nothing here is a prediction we are proud of in advance. These are dated, checkable questions raised by a story we already published. When one resolves, it moves to the track record below rather than quietly disappearing — including the times our implied expectation was wrong.
Track record
Every item that has resolved, kept permanently. We do not remove a wrong call.
RESOLVED NO — Did Claude Sonnet 5 rise 50 percent on 1 September 2026?
Sonnet 5 was priced to rise from USD 2.00/10.00 to USD 3.00/15.00 per million tokens on 1 September 2026. On 10 August 2026, Anthropic made the introductory rate permanent instead. The watchlist resolved NO, and the direction of our implied expectation was wrong — we had flagged the scheduled rise as the figure to model against. See how this resolved and the Claude Sonnet 5 profile.
RESOLVED YES, PARTIALLY — Does Alibaba release the Qwen3.8-Max open weights?
Alibaba committed to open weights within a week of the 3 August 2026 announcement. The weights arrived two days late, and what shipped was the text-only base checkpoint under a custom licence — not the full Max product served through the API. The watchlist resolved YES, with the commitment only partially met. See how this resolved and the Qwen3.8-Max profile.
Open — pricing that is scheduled to change
Does DeepSeek raise its API pricing?
DeepSeek’s own pricing page warns that an overall increase is expected, described as significant. At USD 0.14 per million input tokens it is currently the cheapest capable model in the directory, and that is the basis on which a lot of people are choosing it. See the DeepSeek-V4-Flash profile.
Open — commitments that have not landed yet
Does Gemini 3.1 Pro lose its preview label?
Google’s reasoning flagship launched on 19 February 2026 and still ships under a preview model identifier. Google has confirmed Gemini 3.5 Pro is in testing and Gemini 4 is only in pre training. This is the clearest single test of whether the August leadership split changes Google’s shipping cadence. See our analysis and the Gemini 3.1 Pro profile.
Open — figures we could not verify
Do OpenAI’s two pricing pages agree?
OpenAI’s GPT-5.6 launch post and its developer pricing page quote different rates for Terra and Luna. We use the docs pricing page because that is the one OpenAI maintains for billing, and we say so on both profiles rather than picking one silently. See the Terra and Luna profiles.
Does Anthropic publish the Claude Max 20x price?
The Max tier is quoted only as from USD 100 per month across a 5x and a 20x option. The price of the 20x tier is not published, so we record it as unverified. See the Claude profile.
Do Canva and CapCut publish a global price list?
Neither vendor made current consumer pricing readable to us. CapCut’s own help article states that cost varies by region, device, and promotion. We would rather print unverified than a number we could not read on the vendor’s own page. See the Canva and CapCut profiles.
Open — claims waiting on independent evidence
Does anyone reproduce Qwen3.8-Max’s benchmark scores?
Alibaba reports 86.6 on Terminal-Bench 2.1, the strongest agentic coding figure any lab has published. Every number comes from its own release materials. That is normal for a new model and it is not a criticism, but a self reported score from any lab is a claim until someone else reproduces it. See the Qwen3.8-Max profile.
How this page works
An item is added when a published 18BYTE story raises a question with a specific, checkable answer arriving later. When the answer arrives, the item moves from the open sections above into the track record, permanently — we do not delete a resolution, including the ones where we called it wrong.
If something here has resolved and we have not caught it, tell us at hello@18byte.com. See also our editorial standards and corrections policy.