New model launches, official specifications and independent data · Updated September 4, 2026
The AI Model Release Tracker follows important GPT, Claude, Gemini and Grok launches without turning every announcement into a winner. It puts first-party release facts beside independent benchmark availability so you can decide what is ready to test, what it costs and what is still unknown.
The short answer
- Use the release date and access status to check whether a new model should actually be visible in your account.
- Use the provider specifications to compare context, output limits, caching and headline API rates.
- Use the independent-data panel only when a matching benchmark row exists. “Pending” means HiseHub will not invent a score from launch material.
- Open the source before changing a subscription or production API because rollouts and documentation can change quickly.
HiseHub Release Watch
Track new AI models without confusing launch claims with benchmark evidence
See when a model launched, what its provider officially lists for price, context and access, and whether a comparable independent benchmark row has reached HiseHub yet.
4 releases shown
GPT-6 Astra
API model: gpt-6-astra
Adds asynchronous tool calls, mid-turn steering and effort changes for long-running OpenAI Responses API workflows.
- Input / output
- $10 / $50base; >272K input changes rates
- Cached input
- $1per 1M tokens
- Context
- 1.05M1,050,000 input tokens
- Max output
- 128K128,000 output tokens
Trusted Access enterprises first; API and paid ChatGPT plans are rolling out in stages.
- Intelligence
- 52.8
- Coding
- 76.9
- Agentic
- 51.5
- Output speed
- 56.5 tok/s
Highest-scoring available row: GPT-6 Astra (max). Scores are a snapshot, not a universal winner.
Gemini 3.8 Flash
API model: gemini-3.8-flash
Targets long-horizon software, agent and enterprise workflows at a substantially lower token-price tier.
- Input / output
- $0.75 / $3.75introductory through Dec 31, 2026
- Cached input
- Supportedcheck the current pricing page
- Context
- 1.05M1,048,576 input tokens
- Max output
- 65.5K65,536 output tokens
Stable in the Gemini API and announced for Google AI Pro and Ultra; product-surface and regional access can still differ.
- Intelligence
- 41.2
- Coding
- 76.3
- Agentic
- 41.1
- Output speed
- 280.8 tok/s
Highest-scoring available row: Gemini 3.8 Flash (high). Scores are a snapshot, not a universal winner.
Claude Fable 5.1
API model: claude-fable-5-1
Adds adaptive reasoning controls for sustained agentic work, with migration changes for tool choice and thinking blocks.
- Input / output
- $10 / $50per 1M tokens
- Cached input
- $0.25per 1M tokens
- Context
- 1M1,000,000 input tokens
- Max output
- 128K128,000 output tokens
Documented for the Claude API and multiple cloud platforms; Mythos 5.1 remains a separate restricted route.
- Intelligence
- 53.4
- Coding
- 81.6
- Agentic
- 58.0
- Output speed
- 69.4 tok/s
Highest-scoring available row: Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback). Scores are a snapshot, not a universal winner.
Grok 4.6
API model: grok-4-6
A newer Grok API generation focused on reasoning and agent work with a 500K-token context window.
- Input / output
- $2 / $6base; >200K input changes rates
- Cached input
- $0.50per 1M tokens
- Context
- 500K500,000 input tokens
- Max output
- Not separately listedCheck the live model documentation
Available through the xAI API; consumer-product access and API billing should be verified separately.
- Intelligence
- 44.4
- Coding
- 76.8
- Agentic
- 53.4
- Output speed
- 58.5 tok/s
Highest-scoring available row: Grok 4.6 (high). Scores are a snapshot, not a universal winner.
Two evidence layers, kept separate
Provider facts—release date, token price, context, model ID and access—come from the linked first-party documentation and are checked editorially. Benchmark signals come from HiseHub’s cached Artificial Analysis feed and update when that existing daily sync receives a matching model row.
The benchmark panel may update automatically after the independent source adds or revises a model. The editorial release facts do not silently scrape provider pages; HiseHub rechecks those when providers change documentation. Last benchmark sync: September 9, 2026 06:36 UTC.
AI Updates Official-source events; automatic candidates are filtered, deduplicated and labelled. Last source run: September 10, 2026 01:35 UTC. Langflow v1.12.1 is listed as a stable release. Source notes: fix frontend : render no icon for plain select trigger fix: detect watsonx models behind LiteLLM proxy HiseHub links the official source; rollout and compatibility remain source-specific. Affects: Teams evaluating or operating Langflow in a hosted or self-hosted workflow. Action: Read the official notes first, check compatibility and backups, and verify the relevant plan, region or deployment before changing production. n8n n8n@2.38.4 is listed as a stable release. Source notes: core: Keep a serving external secrets provider active when its replacement fails core: Prevent Anthropic agent threads from breaking permanently and keep run errors visible HiseHub links the official source; rollout and compatibility remain source-specific. Affects: Teams evaluating or operating n8n in a hosted or self-hosted workflow. Action: Read the official notes first, check compatibility and backups, and verify the relevant plan, region or deployment before changing production. n8n 2.37.11 is listed as a stable release. HiseHub only records the official event here; availability, compatibility and rollout details still belong to the linked source. Affects: People evaluating or operating n8n, especially users whose workflow depends on the released project or product. Action: Read the official notes first, check compatibility and backups, and verify the relevant plan, region or deployment before changing production. Source notes: WeatherNext 3, our most advanced global weather AI model, is now in Search, Gemini, Maps, Google Maps Platform, and Cloud. HiseHub links the official source; rollout and compatibility remain source-specific. Affects: Weather and forecasting developers or services evaluating the model; this is not a Gemini subscription entitlement. Action: Check the official model documentation, API or service availability, and regional terms before planning an integration. Langflow v1.12.0 is listed as a stable release. HiseHub only records the official event here; availability, compatibility and rollout details still belong to the linked source. Affects: People evaluating or operating Langflow, especially users whose workflow depends on the released project or product. Action: Read the official notes first, check compatibility and backups, and verify the relevant plan, region or deployment before changing production. These are short, source-linked notices rather than long news articles. Release status, compatibility and access can change; verify the official source before upgrading or changing a plan.
Curated changes worth checking
1.12.1
n8n@2.38.4
n8n 2.37.11
Introducing WeatherNext 3, our most advanced and accurate global weather AI model
Langflow v1.12.0
What changed in the September 2026 model wave?
GPT-6 Astra, Gemini 3.8 Flash and Claude Fable 5.1 arrived within three days. That makes a normal “which model is best?” headline especially unreliable: the products target different price bands, appear through different access routes and do not receive comparable third-party evaluation data at the same time.
GPT-6 Astra and Claude Fable 5.1 share the same published base token rates, but their caching and long-context billing differ. Gemini 3.8 Flash is listed in a much lower API price tier. A useful comparison therefore needs to separate model capability from availability, integration risk and total workload cost.
For a detailed early analysis, read GPT-6 Astra vs Claude Fable 5.1: Price, Context, Access & Early Verdict. For an interactive ranking after independent data arrives, use the AI Model Comparison.
How HiseHub verifies the tracker
Provider facts come from first-party documentation
Release dates, API model IDs, context windows, maximum output, token prices and stated access routes are checked against documentation from OpenAI, Anthropic, Google and xAI. Each release card links directly to its provider page. HiseHub does not treat a social post, reseller page or copied pricing table as the source of record.
Benchmark data comes from a separate independent feed
Capability scores and measured output speed are matched against the cached Artificial Analysis data already used by HiseHub’s comparison tool. That feed refreshes separately. When several reasoning or effort variants exist, the release card displays the highest-scoring available row and names it explicitly rather than hiding the variant.
This separation prevents three common errors: copying a provider’s self-reported launch benchmark into an “independent” ranking, assigning an old model family’s score to a new version, and calling an untested model the winner because it was announced most recently.
How to use a new-model announcement before paying
- Verify live access. A release page can describe a phased rollout before the model reaches every plan, account, region, workspace or API project.
- Check the correct billing route. A consumer subscription and an API project are normally separate products. One does not automatically include the other.
- Compare the complete rate card. Input, output, cached input, cache writes, tools, long-context multipliers and regional processing can affect the final bill.
- Wait for comparable evidence when necessary. Provider system cards are valuable, but they are not a neutral head-to-head evaluation.
- Test representative work. Measure successful completion, retries, latency and total cost on tasks that resemble your real use case.
Use the AI API Cost Calculator to turn token rates into a workload estimate. If your question is which consumer tier exposes a model, check the AI Model Access Map instead.
Why a release tracker is different from an AI news feed
A broad AI news feed becomes stale quickly and usually adds little value beyond rewriting announcements. This tracker is narrower: it follows launches that change a purchase, access or integration decision. Every entry needs a first-party source, structured specifications and a clear independent-data status.
That also means HiseHub will not add every preview, rumour, fine-tune or renamed product. A model belongs here when users can meaningfully ask whether they can access it, what it costs or whether independent comparison data is available.
Frequently asked questions
Does the AI Model Release Tracker update automatically?
The independent benchmark panel reads HiseHub’s existing daily model-data cache and can update when a matching model row appears. Official product facts are reviewed editorially because provider pages can change structure or wording and should not be copied blindly.
Why is a new model marked “benchmark pending”?
It means the independent feed does not yet contain a comparable scored row. HiseHub keeps that gap visible instead of estimating a result from an older model or a provider claim.
Do higher benchmark scores always mean a better subscription?
No. A consumer plan can have different usage limits, model access, tools and regional availability. Benchmarks can help form a shortlist, but the right plan depends on the task, access route, reliability and total price.
Are the prices consumer subscription prices?
No. The release cards show published text API rates per million tokens. For consumer and App Store pricing, use HiseHub’s AI Subscription Price Comparison.
Editorial facts last checked September 4, 2026. Model names, access, prices, limits and third-party scores can change. HiseHub is independent and is not affiliated with OpenAI, Anthropic, Google, xAI, X or Artificial Analysis.