• Skip to primary navigation
  • Skip to main content
  • Skip to footer
Sq Magazine LogoSQ Magazine

Smarter Insights for a Fast-Moving Digital World

  • Latest News
  • Statistics
  • About
  • Contact
Subscribe

SQ Magazine » The Tech Index » AI Model Tracker

AI Model Tracker

A curated timeline of the models that matter: who shipped what, the specs and license you can actually build on, and what it costs. Every entry is sourced from the developer's own release.

33 models tracked · verified through 12 Aug 2026 · last verified 11 Sep 2026 · next review by 11 Sep 2026

Barry Elad
Maintained By
Barry Elad
Barry Elad
Founder & Senior Journalist • 728 Articles
Barry Elad is a seasoned journalist and analyst specializing in finance, technology, AI, and founder of SQ Magazine. He explores the world o...
LATEST POSTS:
OpenAI Taps Samsung for Breakthrough Next-Gen Chips
How Many Bitcoins Are There in 2026? Supply, Mined and Remaining Statistics
OpenAI Agents Hijacked German Wiki, Researchers Say

Models tracked
33
Open weights
8
Developers
11
Re-verify cycle
30day
Largest context window GPT-6 Astra · 1.1M tokens
8 open weights 9 community 16 proprietary

Latest change DeepSeek released V4.1-Flash on 10 Sep 2026 and retired V4-Flash and V4-Flash-Vision-Exp from its API, so both rows move to Retired with weights-only access. V4-Pro moves to Deprecated as DeepSeek phases it out. Source: api-docs.deepseek.com/news/news260910 (10 Sep 2026) All changes

All records

33 models, newest release first. Filter or search below.

What is in scope Model releases the developer itself published specs or pricing for.

33 models from 11 developers, of which 8 ship open weights. 14 of 33 disclose a training-data cutoff; the rest publish none, and this table leaves those cells empty rather than estimating them.

What changed 11 Sep 2026

DeepSeek released V4.1-Flash on September 10, 2026, and the tracker now carries it as a draft. DeepSeek describes it as a 552B-parameter mixture-of-experts model that processes images and text, with MIT-licensed weights on Hugging Face and a context window of 1,048,576 tokens in its configuration file. The same release notes state that V4-Flash and V4-Flash-Vision-Exp are retired, and the DeepSeek API now routes both model names to V4.1-Flash. Both rows move to Retired with weights-only access. DeepSeek also says it is phasing out V4-Pro, so that row moves to Deprecated. All three keep public MIT weights.

A sourced timeline of model releases, not a benchmark board. Specs and pricing are as published by the developer; “undisclosed” stays undisclosed, and price is a relative tier with the exact figure on the provider’s page via each row’s source. Informational only.
Model Developer Released Modality License Price Access Our coverage Record detail
OP GPT-6 Astra OpenAI Sep 2026 Text + Vision Proprietary High tier$$$ Limited OpenAI Releases GPT-6 Astra After Largest Training Run Yet →
Context window
1.1M tokens
Parameters
Undisclosed
Knowledge cutoff
Apr 2026
Lifecycle
Current
Stated use
Hardest end-to-end work across computer use, browsing and software engineering
License terms
Commercial API; OpenAI terms. Trusted Access Program for enterprises at launch
Primary source
OpenAI release ↗
Last confirmed
4 Sep 2026
Recent change
Second independent verification 4 Sep 2026 against OpenAI docs: 1,050,000 context window, 128,000 max output, Apr 30 2026 knowledge cutoff, price $10 input / $50 output per MTok, text+image in and te…
GO Gemini 3.8 Flash Google Sep 2026 Multimodal Proprietary Low tier$ API Google announcement ↗
Context window
1M tokens
Parameters
Undisclosed
Lifecycle
Current
Stated use
Long-horizon software engineering, autonomous agents, and complex enterprise workflows
Benchmarks (developer-reported)
HLE-Verified 54.9%
License terms
Commercial API; Google terms
Primary source
Google announcement ↗
Last confirmed
3 Sep 2026
Recent change
GA 2 Sep 2026. ai.google.dev lists 1,048,576 in and 65,536 out, no cutoff. Bench corrected 3 Sep: DeepSWE 1.1 73.7% is on no vendor page; the announcement states HLE-Verified 54.9%.
Security & safety
Model card published for Gemini 3.8 Flash.
AN Claude Fable 5.1 Anthropic Sep 2026 Text + Vision Proprietary High tier$$$ API Anthropic release ↗
Context window
1M tokens
Parameters
Undisclosed
Knowledge cutoff
Jun 2026
Lifecycle
Current
Stated use
Demanding reasoning and long-horizon agentic work
Benchmarks (developer-reported)
CursorBench 3.2.0 73.4% · Terminal-Bench 4.0 55.8% · OSWorld 2.0 77.9% (partial)
License terms
Commercial API; Claude Mythos 5.1 shares its specs and pricing, by invitation only
Primary source
Anthropic release ↗
Last confirmed
3 Sep 2026
Recent change
Announcement states Fable 5.1 is generally available and priced as Fable 5 was, $10/1M input and $50/1M output, with cache reads cut 75% to $0.25/1M
Security & safety
System card published for Fable 5.1 and Mythos 5.1; Anthropic states the benchmark scores were produced with production safeguards enabled.
ME Muse-Glimmer-30B Meta Aug 2026 Text + Vision Open weights Self-host Open weights Meta model card ↗
Context window
128K tokens
Parameters
~29.6B
Knowledge cutoff
Jan 2026
Lifecycle
Current
Stated use
Multimodal agentic model sized for local deployment
Benchmarks (developer-reported)
SWE-Bench Pro 51.2 · AIME 2026 94.7 · MCP Atlas 75.5 · Charxiv Reasoning 78.8
License terms
Apache 2.0 covering the repository and the weights, per the model card.
Primary source
Meta model card ↗
Last confirmed
4 Sep 2026
Recent change
Second verification 4 Sep 2026: the HF org meta-models is the verified "Meta Inc." account and the card credits Meta Superintelligence Lab, so the Meta attribution holds. Context corrected 131 to 128…
Security & safety
Open weights: you own the data boundary and the safety tuning. The repository ships no system card or safety evaluation beside the weights.
DE DeepSeek-V4-Flash-Vision-Exp DeepSeek Aug 2026 Multimodal Open weights Self-host Open weights DeepSeek model card ↗
Context window
1M tokens
Parameters
305B
Lifecycle
Retired
Stated use
Experimental open-weights multimodal agent work across image and text inputs
Benchmarks (developer-reported)
Terminal Bench 2.1 83.9 · DeepSWE 59.3 · Cybergym 75.3 · Chartography 64.3
License terms
MIT License covering the repository, per the model card. The card states no additional commercial restriction.
Primary source
DeepSeek model card ↗
Last confirmed
11 Sep 2026
Recent change
DeepSeek release notes of 10 Sep 2026 state V4-Flash-Vision-Exp is retired and the API now routes deepseek-v4-flash-vision-exp to V4.1-Flash. Weights stay public under MIT.
Security & safety
Weights public and ungated on Hugging Face across 48 safetensors shards, so self-hosting keeps data in your own boundary. The repository ships no system card or safety evaluation, and the card marks the model experimental.
AL Qwen3.8-Flash-Next Alibaba Aug 2026 Multimodal Community Self-host Open weights Qwen model card ↗
Context window
256K tokens
Parameters
125B (6B active)
Lifecycle
Current
Stated use
Long-context agentic coding and computer use at 6B active parameters
Benchmarks (developer-reported)
GPQA Diamond 91.7 · SWE-bench Pro 62.5 · LiveCodeBench v6 91.9 · HLE 35.9
License terms
Qwen Community License 1.0. Commercial use allowed; name display above 100M MAU or USD 20M monthly revenue; separate license for MaaS or AI assistants.
Primary source
Qwen model card ↗
Last confirmed
4 Sep 2026
Recent change
Second verification 4 Sep 2026 from the cited repo: text_config.max_position_embeddings 262,144 confirms the 256K context; model_type is qwen4_exp and HF licenses it as other, matching the experiment…
Security & safety
Weights public and ungated on Hugging Face across 131 safetensors shards; self-host data boundary. Unlike the Qwen3.8-Max License on the 2.4T row, this one needs a separate Qwen license for Model-as-a-Service or AI Work Assistant use with no revenue floor.
AL Qwen3.8-27B Alibaba Aug 2026 Multimodal Open weights Self-host Open weights Qwen model card ↗
Context window
256K tokens
Parameters
27B (dense)
Lifecycle
Current
Stated use
Agentic coding and computer use in a compact dense model
Benchmarks (developer-reported)
GPQA Diamond 89.2 · SWE-bench Pro 61.7 · Terminal Bench 2.1 73.0 · OSWorld-Verified 84.3
License terms
Apache-2.0
Primary source
Qwen model card ↗
Last confirmed
4 Sep 2026
Recent change
Second verification 4 Sep 2026 from the cited repo: text_config.max_position_embeddings 262,144 confirms the 256K context, HF metadata gives license apache-2.0. Published.
Security & safety
Apache-2.0 weights, public and ungated on Hugging Face across 18 safetensors shards; self-host data boundary. The repo LICENSE is unmodified Apache-2.0 and adds no scale caps, MAU thresholds or revenue thresholds, unlike the Qwen3.8-Max License on the 2.4T row.
GO Gemini 3.7 Flash Google Aug 2026 Multimodal Proprietary Low tier$ API Google announcement ↗
Context window
1M tokens
Parameters
Undisclosed
Lifecycle
Superseded
Stated use
Complex coding, agentic workflows, and multi-step execution
Benchmarks (developer-reported)
FrontierCode 1.1 Main 43.6% · DeepSWE v1.1 65.3% · AutomationBench 30.4%
License terms
Commercial API; Google terms
Primary source
Google announcement ↗
Last confirmed
3 Sep 2026
Recent change
Announced 13 Aug 2026. 1,048,576 in and 65,536 out, no cutoff. $0.75/$3.75 per M to 31 Dec 2026, then $1.50/$7.50. Superseded by Gemini 3.8 Flash 2 Sep 2026; re-confirmed 3 Sep, pricing unchanged.
XA Grok 4.6 xAI Aug 2026 Text + Vision Proprietary Mid tier$$unconfirmed API xAI announcement ↗
Context window
500K tokens
Parameters
Undisclosed
Knowledge cutoff
Feb 2026
Lifecycle
Current
Stated use
Coding, agentic tasks, and knowledge work
Benchmarks (developer-reported)
CursorBench v3.2 69.9% · DeepSWE v1.1 65.9% · FrontierCode v1.1 61.3%
License terms
Commercial API
Primary source
xAI announcement ↗
Last confirmed
12 Aug 2026
Recent change
x.ai announcement gives benchmark scores and states neither a context window nor a price, so both figures on this row come from outside the citation.
AL Qwen3.8-2.4T-A95B Alibaba Aug 2026 Text Community Self-host Open weights Qwen model card ↗
Context window
256K tokens
Parameters
2.4T (95B active)
Lifecycle
Current
Stated use
Long-context reasoning and agentic coding
Benchmarks (developer-reported)
GPQA Diamond 92.6 · SWE-bench Pro 67.7 · Terminal Bench 2.1 86.6 · PaperBench 93.0
License terms
Qwen3.8-Max License. Commercial use allowed; name display required above 100M MAU or $20M monthly revenue; separate license for MaaS above $50M/12 months.
Primary source
Qwen model card ↗
Last confirmed
3 Sep 2026
Recent change
Cited repo config.json gives max_position_embeddings 262,144, confirming the 256K context
Security & safety
Full weights public and ungated on Hugging Face across 213 safetensors shards; self-host data boundary. The Qwen3.8-Max License requires a separate license from Qwen for Model-as-a-Service use above $50M revenue in 12 consecutive months.
DE DeepSeek-V4-Pro DeepSeek Aug 2026 Text Open weights Low tier$ API + weights DeepSeek model card ↗
Context window
1M tokens
Parameters
1.65T (MoE)
Lifecycle
Deprecated
Stated use
Frontier-scale open-weights MoE with 1M context at low hosted-API prices
Benchmarks (developer-reported)
Terminal Bench 2.1 87.9% · DeepSWE 62.7% · Cybergym 83.3%
License terms
MIT License covering both the repository and the model weights, per the model card.
Primary source
DeepSeek model card ↗
Last confirmed
11 Sep 2026
Recent change
DeepSeek release notes of 10 Sep 2026 say it is phasing out V4-Pro, with deepseek-v4-pro API requests routed to V4.1-Flash from 14 Sep 2026. Weights stay public under MIT.
Security & safety
MIT weights allow self-hosting for full data control; review the DeepSeek provider data policy before using the hosted API. The repository ships no system card or safety evaluation beside the weights.
DE DeepSeek-V4-Flash DeepSeek Jul 2026 Text Open weights Self-host Open weights DeepSeek model card ↗
Context window
1M tokens
Parameters
304B (MoE)
Lifecycle
Retired
Stated use
Low-cost 1M-context open-weights MoE for agentic and long-context work
Benchmarks (developer-reported)
Terminal Bench 2.1 82.7% · DeepSWE 54.4% · Cybergym 76.7%
License terms
MIT License covering both the repository and the model weights, per the model card.
Primary source
DeepSeek model card ↗
Last confirmed
11 Sep 2026
Recent change
DeepSeek release notes of 10 Sep 2026 state V4-Flash is retired and the API now routes deepseek-v4-flash to V4.1-Flash. Weights stay public under MIT on Hugging Face.
Security & safety
MIT weights allow self-hosting for full data control; review the DeepSeek provider data policy before using the hosted API. The repository ships no system card or safety evaluation beside the weights.
XA Grok 4.5 xAI Jul 2026 Text + Vision Proprietary Mid tier$$unconfirmed API xAI announcement ↗
Context window
500K tokens
Parameters
Undisclosed
Lifecycle
Superseded
Stated use
Coding, agentic tasks, and knowledge work
Benchmarks (developer-reported)
SWE Marathon pass@1 29.0% · Terminal Bench 2.1 83.3% · SWE-Bench Pro 64.7%
License terms
Commercial API
Primary source
xAI announcement ↗
Last confirmed
12 Aug 2026
Recent change
x.ai announcement gives benchmark scores and states neither a context window nor a price, so both figures on this row come from outside the citation.
AN Claude Opus 5 Anthropic Jul 2026 Text + Vision Proprietary High tier$$$ API Anthropic release ↗
Context window
1M tokens
Parameters
Undisclosed
Knowledge cutoff
May 2026
Lifecycle
Current
Stated use
Complex agentic coding and enterprise work
Benchmarks (developer-reported)
SWE-bench Verified 96.0% · OSWorld 2.0 70.6% · DeepSWE v1.1 68.8%
License terms
Commercial API
Primary source
Anthropic release ↗
Last confirmed
3 Sep 2026
Recent change
Added in release week; $5/$25 pricing, 1M context, May 2026 knowledge cutoff per the models overview. SWE-bench Verified figure appears in system card section 8.2 prose, not the summary table
Security & safety
System card published. Cyber classifiers allow source-code vulnerability discovery but block binary scanning, penetration testing, and exploit generation.
MO Kimi K3 Moonshot Jul 2026 Multimodal Community Mid tier$$ API + weights Moonshot’s Kimi K3 Beats Top U.S. AI Models in Blind Tests →
Context window
1M tokens
Parameters
2.8T (104B active)
Lifecycle
Current
Stated use
Long-context reasoning, agentic coding
Benchmarks (developer-reported)
Terminal-Bench 2.1 88.3% · DeepSWE 67.5% · SWE Marathon 42.0%
License terms
Kimi K3 License: MIT-style grant with a Model-as-a-Service condition, plus a "Kimi K3" attribution requirement above 100M monthly active users or $20M monthly…
Primary source
Moonshot model card ↗
Last confirmed
3 Sep 2026
Recent change
Weights published 13 Jun 2026; specs, licence tier and modality verified against the Moonshot model card.
Security & safety
Full weights published on Hugging Face under the Kimi K3 License. First-party Kimi API also live.
GO Gemini 3.5 Flash-Lite Google Jul 2026 Multimodal Proprietary Low tier$ API Google model card ↗
Context window
1M tokens
Parameters
Undisclosed
Knowledge cutoff
Mar 2026
Lifecycle
Current
Stated use
Cheapest, low-latency, high throughput
Benchmarks (developer-reported)
SWE-Bench Pro 54.2% · OSWorld-Verified 74.0% · Terminal-Bench 2.1 54.0%
License terms
Commercial API
Primary source
Google model card ↗
Last confirmed
3 Sep 2026
Recent change
Current confirmed: Google lists Gemini 3.5 Flash-Lite as Stable on the models page, and its deprecations row announces no shutdown date
Security & safety
Model card; safety evals.
GO Gemini 3.6 Flash Google Jul 2026 Multimodal Proprietary Low tier$ API Google model card ↗
Context window
1M tokens
Parameters
Undisclosed
Knowledge cutoff
Mar 2026
Lifecycle
Superseded
Stated use
Fast workhorse, up to 17% fewer tokens
Benchmarks (developer-reported)
SWE-Bench Pro 58.7% · OSWorld-Verified 83.0% · Terminal-Bench 2.1 78.0%
License terms
Commercial API; Google terms
Primary source
Google model card ↗
Last confirmed
3 Sep 2026
Recent change
Launch post prices 3.6 Flash at $1.50/1M input and $7.50/1M output
Security & safety
Model card and safety evals; enterprise data controls.
OP GPT-5.6 Sol OpenAI Jul 2026 Text + Vision Proprietary High tier$$$ API OpenAI Launches GPT 5.6 Sol With Powerful New AI Features →
Context window
1.1M tokens
Parameters
Undisclosed
Knowledge cutoff
Feb 2026
Lifecycle
Current
Stated use
Flagship of the GPT-5.6 series (Sol/Terra/Luna)
Benchmarks (developer-reported)
SWE-Bench Pro 64.6% · GPQA Diamond 94.6% · Terminal-Bench 2.1 88.8%
License terms
Commercial API; OpenAI terms
Primary source
OpenAI release ↗
Last confirmed
3 Sep 2026
Recent change
Model page states a 1,050,000 context window, 128,000 max output tokens and a 16 Feb 2026 knowledge cutoff
Security & safety
System card and preparedness framework; enterprise data controls.
AN Claude Sonnet 5 Anthropic Jun 2026 Text + Vision Proprietary Mid tier$$ API Anthropic release ↗
Context window
1M tokens
Parameters
Undisclosed
Knowledge cutoff
Jan 2026
Lifecycle
Current
Stated use
Most agentic Sonnet; near-Opus at lower cost
Benchmarks (developer-reported)
SWE-bench Verified 85.2% · OSWorld-Verified 81.2% · Terminal-Bench 2.1 80.4%
License terms
Commercial API; Anthropic usage policy
Primary source
Anthropic release ↗
Last confirmed
3 Sep 2026
Recent change
Upgrade to Sonnet 4.6; standard pricing 3/15, introductory 2/10 through Aug 31 2026. 1M context, 128K output. Released Jun 30 2026.
Security & safety
System card published; cyber safeguards enabled by default; enterprise data controls.
AN Claude Fable 5 Anthropic Jun 2026 Text + Vision Proprietary High tier$$$ API Claude Fable 5 Ends Free Access For Pro Subscribers →
Context window
1M tokens
Parameters
Undisclosed
Knowledge cutoff
Jan 2026
Lifecycle
Current
Stated use
Frontier intelligence for long-running agents
License terms
Commercial API; Mythos-class 30-day data retention, not used for training
Primary source
Anthropic release ↗
Last confirmed
3 Sep 2026
Security & safety
System card published; enterprise data controls.
AN Claude Opus 4.8 Anthropic May 2026 Text + Vision Proprietary High tier$$$ API Anthropic Unveils Claude Science to Transform Research →
Context window
1M tokens
Parameters
Undisclosed
Knowledge cutoff
Jan 2026
Lifecycle
Superseded
Stated use
Frontier reasoning, long-horizon agents
Benchmarks (developer-reported)
SWE-bench Verified 88.6% · GPQA Diamond 93.6% · OSWorld-Verified 83.4%
License terms
Commercial API; Anthropic usage policy
Primary source
Anthropic model card ↗
Last confirmed
3 Sep 2026
Recent change
Superseded by Claude Opus 5 (Jul 2026); remains available and priced at $5/$25 on the current models page
Security & safety
System card and Responsible Scaling Policy published; customer API data not used for training.
GO Gemini 3 Pro Google Nov 2025 Multimodal Proprietary — API Google, Gemini API deprecations ↗
Context window
1M tokens
Parameters
Undisclosed
Lifecycle
Retired
Stated use
Frontier multimodal reasoning with Deep Think
Benchmarks (developer-reported)
SWE-bench Verified 76.2% · GPQA Diamond 91.9% · Terminal-Bench 2.0 54.2%
License terms
Commercial API
Primary source
Google, Gemini API deprecations ↗
Last confirmed
3 Sep 2026
Recent change
Retired confirmed against Google directly: the models page lists Gemini 3 Pro only under Previous models as (Shut down), and the deprecations table gives gemini-3-pro-preview a shutdown date of 9 Mar…
Security & safety
Model card and safety evals; enterprise data controls.
AN Claude Haiku 4.5 Anthropic Oct 2025 Text + Vision Proprietary Low tier$ API Anthropic model card ↗
Context window
200K tokens
Parameters
Undisclosed
Knowledge cutoff
Feb 2025
Lifecycle
Current
Stated use
Low-latency, cost-efficient
Benchmarks (developer-reported)
SWE-bench Verified 73.3% · GPQA Diamond 73.0% · AIME 2025 80.7% (no tools)
License terms
Commercial API
Primary source
Anthropic model card ↗
Last confirmed
3 Sep 2026
Security & safety
System card; enterprise data controls.
XA Grok 2 xAI Aug 2025 Text Community Self-host Open weights xAI (Hugging Face) ↗
Context window
128K tokens
Parameters
Undisclosed
Lifecycle
Superseded
Stated use
Open frontier-class weights
Benchmarks (developer-reported)
HumanEval 88.4% · GPQA 56.0% · MMLU-Pro 75.5%
License terms
Grok 2 Community License Agreement
Primary source
xAI (Hugging Face) ↗
Last confirmed
3 Sep 2026
Recent change
Confirmed from the cited repo: config.json gives max_position_embeddings 131072 (128K) and the README licenses the weights under the Grok 2 Community License Agreement
Security & safety
Open weights under the Grok 2 Community License (restricted); deployer owns safety tuning.
OP GPT-5 OpenAI Aug 2025 Text + Vision Proprietary Mid tier$$ API OpenAI Launches ChatGPT Work With GPT-5.6 Agents →
Context window
400K tokens
Parameters
Undisclosed
Knowledge cutoff
Sep 2024
Lifecycle
Deprecated
Stated use
Broad reasoning and tools
Benchmarks (developer-reported)
SWE-bench Verified 74.9% · AIME 2025 94.6% (no tools) · MMMU 84.2%
License terms
Commercial API
Primary source
OpenAI model card ↗
Last confirmed
3 Sep 2026
Recent change
Superseded by GPT-5.6 (Jul 2026); OpenAI lists GPT-5 as prior-generation.
Security & safety
System card published.
AL Qwen3 Alibaba Apr 2025 Text Open weights Self-host Open weights Alibaba model card ↗
Context window
128K tokens
Parameters
0.6B-235B
Lifecycle
Superseded
Stated use
Broad open family, multilingual
Benchmarks (developer-reported)
LiveCodeBench 70.7% · AIME 2025 81.5% · BFCL v3 70.8%
License terms
Apache-2.0
Primary source
Alibaba model card ↗
Last confirmed
3 Sep 2026
Recent change
Alibaba current Model Studio API lists qwen3.7-max/plus and qwen3.6-flash; the open Apr-2025 variant qwen3-235b-a22b is not on the current first-party pricing page, so price kept free and access set…
Security & safety
Apache-2.0 weights; self-host data boundary.
ME Llama 4 Meta Apr 2025 Text + Vision Community Self-host Open weights Meta model card ↗
Context window
1M tokens
Parameters
up to 400B (MoE)
Lifecycle
Current
Stated use
Open MoE, long context
Benchmarks (developer-reported)
LiveCodeBench 43.4% · GPQA Diamond 69.8% · MMLU-Pro 80.5%
License terms
Llama Community License (scale cap)
Primary source
Meta model card ↗
Last confirmed
3 Sep 2026
Recent change
Meta operates no first-party inference API and points developers to third-party hosts, so access is weights-only. Represents Llama 4 Maverick (1M context, 17B active / 400B total, 128 experts, releas…
Security & safety
Open weights: you own the data boundary and the safety tuning. Scale-cap clause for very large deployments.
GO Gemma 3 Google Mar 2025 Text + Vision Community Self-host Open weights Google model card ↗
Context window
128K tokens
Parameters
1B-27B
Lifecycle
Superseded
Stated use
Efficient open family
Benchmarks (developer-reported)
LiveCodeBench 29.7% · GPQA Diamond 42.4% · MMLU-Pro 67.5%
License terms
Gemma terms of use (restricted)
Primary source
Google model card ↗
Last confirmed
3 Sep 2026
Recent change
Open family (1B/4B/12B/27B); 4B/12B/27B support vision (text plus image). Built on Gemini 2.0 research. ShieldGemma 2 image safety checker built on Gemma 3.
Security & safety
Open weights under Gemma terms; self-host safety is your responsibility.
DE DeepSeek-R1 DeepSeek Jan 2025 Text Open weights Low tier$ API + weights DeepSeek model card ↗
Context window
128K tokens
Parameters
671B (MoE)
Lifecycle
Superseded
Stated use
Open reasoning model
Benchmarks (developer-reported)
SWE-bench Verified 49.2% · GPQA Diamond 71.5% · MMLU-Pro 84.0%
License terms
MIT
Primary source
DeepSeek model card ↗
Last confirmed
3 Sep 2026
Recent change
Updated as R1-0528 (May 2025).
Security & safety
MIT weights; self-host data boundary. Reasoning traces can be verbose; review before logging.
MI Phi-4 Microsoft Dec 2024 Text Open weights Self-host Open weights Microsoft model card ↗
Context window
16K tokens
Parameters
14B
Knowledge cutoff
Jun 2024
Lifecycle
Current
Stated use
Small-model reasoning
Benchmarks (developer-reported)
HumanEval 82.6% · GPQA 56.1% · MMLU 84.8%
License terms
MIT
Primary source
Microsoft model card ↗
Last confirmed
3 Sep 2026
Recent change
Phi-4 family expanded in 2025 (Phi-4-mini, Phi-4-multimodal, Phi-4-reasoning).
Security & safety
MIT weights, small enough to run locally; a good fit where data cannot leave the device.
DE DeepSeek-V3 DeepSeek Dec 2024 Text Community Low tier$ API + weights DeepSeek model card ↗
Context window
128K tokens
Parameters
671B (MoE)
Lifecycle
Superseded
Stated use
Efficient open MoE, low-cost API
Benchmarks (developer-reported)
SWE-bench Verified 42.0% · GPQA Diamond 59.1% · MMLU-Pro 75.9%
License terms
DeepSeek Model License (code is MIT; weights carry use restrictions)
Primary source
DeepSeek model card ↗
Last confirmed
3 Sep 2026
Recent change
License corrected: V3 weights ship under the DeepSeek Model License, not MIT — MIT covers only the code repository (R3 audit, Jul 2026)
Security & safety
MIT weights plus low-cost hosted API. Self-host for data control; review provider data policy if using the API.
CO Command R+ Cohere Aug 2024 Text Community Mid tier$$ API + weights Cohere model card ↗
Context window
128K tokens
Parameters
104B
Lifecycle
Superseded
Stated use
RAG and tool use
License terms
CC-BY-NC (non-commercial)
Primary source
Cohere model card ↗
Last confirmed
3 Sep 2026
Recent change
Command R+ 08-2024 (current version), a refresh of the original Apr 2024 release; 128K context, CC-BY-NC weights plus paid Cohere API. Superseded by Command A (Mar 2025) and Command A Reasoning.
Security & safety
Weights non-commercial (CC-BY-NC); enterprise license required to ship commercially.
MI Mistral Large 2 Mistral Jul 2024 Text Community Self-host Open weights Mistral model card ↗
Context window
128K tokens
Parameters
123B
Lifecycle
Retired
Stated use
Strong European frontier
Benchmarks (developer-reported)
MMLU 84.0% (pretrained)
License terms
Mistral Research License (non-commercial); commercial license to self-deploy
Primary source
Mistral model card ↗
Last confirmed
3 Sep 2026
Recent change
Mistral retired mistral-large-2407 on 30 Mar 2025 and names Mistral Large 3 as the replacement; the card states access has ended, so the row moves from Superseded to Retired
Security & safety
EU-based provider; commercial license with enterprise data terms.

No records match the current filters.

What the data shows

Every figure below comes from the verified table above, redrawn as the trends and comparisons a table cannot show.

Model releases by monthNotable releases per month, dated from the developer's own announcement.02468Dec 2024: 2 releases2 releasesDec 2024Jan 2025: 1 release1 releaseMar 2025: 1 release1 releaseMar 2025Apr 2025: 2 releases2 releasesAug 2025: 2 releases2 releasesAug 2025Oct 2025: 1 release1 releaseNov 2025: 1 release1 releaseNov 2025May 2026: 1 release1 releaseJun 2026: 2 releases2 releasesJun 2026Jul 2026: 7 releases7 releasesAug 2026: 8 releases8 releasesAug 2026Sep 2026: 3 releases3 releasesSep 2026
Context windows comparedTokens of context, in thousands, per the developer's own specification.05001K1.5KGPT-6 AstraGPT-6 Astra: 1,050K1,050KGPT-5.6 SolGPT-5.6 Sol: 1,050K1,050KGemini 3.8 FlashGemini 3.8 Flash: 1,000K1,000KClaude Fable 5.1Claude Fable 5.1: 1,000K1,000KDeepSeek-V4-Flash…DeepSeek-V4-Flash…: 1,000K1,000KGemini 3.7 FlashGemini 3.7 Flash: 1,000K1,000KDeepSeek-V4-ProDeepSeek-V4-Pro: 1,000K1,000KDeepSeek-V4-FlashDeepSeek-V4-Flash: 1,000K1,000KClaude Opus 5Claude Opus 5: 1,000K1,000KKimi K3Kimi K3: 1,000K1,000K
Releases by developerTracked releases per developer.0246GoogleGoogle: 66AnthropicAnthropic: 66DeepSeekDeepSeek: 55AlibabaAlibaba: 44OpenAIOpenAI: 33xAIxAI: 33MetaMeta: 22MoonshotMoonshot: 11MicrosoftMicrosoft: 11CohereCohere: 11

Verification ledger

5 most recent of 16 logged updates
  • DeepSeek released V4.1-Flash on 10 Sep 2026 and retired V4-Flash and V4-Flash-Vision-Exp from its API, so both rows move to Retired with weights-only access. V4-Pro moves to Deprecated as DeepSeek phases it out. Source: api-docs.deepseek.com/news/news260910 10 Sep 2026
  • Added GPT-6 Astra as a draft. OpenAI docs give 1,050,000 context, 128,000 max output, Apr 30 2026 cutoff, $10 in / $1 cached / $50 out per MTok. Trusted Access Program enterprises only at launch, not GA, so it stays a draft. No benchmarks on any reachable primary page. 3 Sep 2026
  • Gemini 3.8 Flash bench corrected: the DeepSWE 1.1 73.7% figure appears on no Google page, replaced with HLE-Verified 54.9% from the 2 Sep announcement. Gemini 3.7 Flash re-confirmed superseded. Gemini 3.8 Flash Cyber staged as a draft, limited access via the Fairwind Program, not GA. 2 Sep 2026
  • Gemini 3.8 Flash added (GA 2 Sep, gemini-3.8-flash): 1M input context, text/image/video/audio/PDF in, 0.75 and 3.75 dollars per million through 2026, DeepSWE 1.1 73.7 percent. Gemini 3.7 Flash marked superseded. Read from the Google model and pricing pages. 2 Sep 2026
  • Added Claude Fable 5.1, published: Anthropic announced it 1 Sep 2026 at 10/50 dollars per million tokens, 1M context, Jun 2026 cutoff. Mythos 5.1 shares the row. Also staged Muse-Glimmer-30B as a draft, Meta open weights on Hugging Face since 9 Aug 2026. 1 Sep 2026

How this tracker is maintained

Every model passes the same checks before it appears, and the row keeps pace with each new release.

  1. 01

    Sourced

    Specs, license and pricing come from the developer’s own release: model card, license file, or pricing page. No benchmark screenshots, no third-hand numbers.

  2. 02

    Dated

    Each row carries its release date and when we last confirmed it. Each version bump gets its own new row.

  3. 03

    Re-checked

    Reviewed every 30 days because the field moves fast. License changes, price cuts and deprecations are logged in the ledger.

Why not a leaderboard?
Live benchmarks already exist and shift daily. This is the durable record: what shipped, when, under which license, at what cost. Row details carry the scores each developer published at release; for current head-to-head rankings, use a live board such as LMArena.
How often is it updated?
Every 30 days, and immediately when a major model ships. The ledger logs each addition and reclassification.

This is informational content, not a procurement or security recommendation. Specs and pricing are the developer’s own disclosures and change with new versions; benchmark scores are vendor-reported and are not independently reproduced here. Confirm current terms on the developer’s page before acting.

Quoting a figure with a link to this page needs no permission. Cite it as you would any source. Reuse of the compiled dataset itself is licensed under CC BY 4.0: credit SQ Magazine and link back.

Sources

  • Alibaba model card
  • Anthropic model card
  • Anthropic release
  • Cohere model card
  • DeepSeek model card
  • Google announcement
  • Google model card
  • Google, Gemini API deprecations
  • Meta model card
  • Microsoft model card
  • Mistral model card
  • Moonshot model card
  • OpenAI model card
  • OpenAI release
  • Qwen model card
  • xAI (Hugging Face)
  • xAI announcement

Footer

SQ Magazine Logo

Smarter Insights for a Fast-Moving Digital World

Connect With Us

Follow Us on Google News

Editorial & Trust

  • About
  • Publishing Principles
  • Fact-Check Policy
  • Corrections Policy
  • Ethics Policy
  • Disclaimer

Worth Checking

  • Social Media Attention Span Stats
  • Gen Z Social Media Statistics
  • TikTok vs. Instagram Statistics
  • LLM Hallucination Statistics
  • Spotify User Statistics
  • Apple Customer Loyalty Statistics
  • Data Breach Tracker
  • Patch Tuesday Dashboard
  • AI Model Tracker
  • AI Funding Tracker
Contact Us
13570 Grove Dr #189,
Maple Grove, MN 55311,
United States
10 a.m. to 6 p.m. | Every day

Copyright © 2022–2026 SQ Magazine. All Rights Reserved. Powered by the Neural Stack.

  • Privacy Policy
  • Terms
  • Accessibility Statement
Company
  • About Us
  • Our Team
  • Our Mission
  • Core Values
Discover
  • Brand Assets
    Brand Assets
  • Stats Methodology
    Stats Research Process
  • Glossary
    Glossary
Categories
  • Internet
  • Technology
  • Artificial Intelligence
  • Gaming
  • Cybersecurity
Internet
How Many People Work at WhatsApp
How Many People Work at WhatsApp 2026: Employee Count and History
Spotify Listening Statistics
Spotify Listening Statistics 2026: Average Listening Time
How Many Subscribers Does MrBeast Have
How Many Subscribers Does MrBeast Have in 2026? Channel Growth Statistics
WhatsApp Business Statistics
WhatsApp Business Statistics 2026: Real Market Insights
Udemy Statistics
Udemy Statistics 2026: Revenue and Learner Data
Coursera Statistics
Coursera Statistics 2026: Learners, Revenue and Growth Data
Technology
How Many iPhones Has Apple Sold
How Many iPhones Has Apple Sold in 2026? Units Sold by Year
How Many Employees Does Amazon Have
How Many Employees Does Amazon Have 2026: Workforce Growth
Netflix vs. Hulu Statistics
Netflix vs Hulu Statistics 2026: Viewer Growth Data
TripAdvisor Statistics
TripAdvisor Statistics 2026: Revenue, Reviews, Viator and TheFork Data
Search Engine Statistics
Search Engine Statistics 2026: Market Share, Volume & AI Shift
NVIDIA Employee Count Statistics
NVIDIA Employee Count Statistics 2026: Headcount, R&D, and Revenue
Artificial Intelligence
AI Music Statistics
AI Music Statistics 2026: Generation, Adoption and Industry Impact
AI Coding Statistics
AI Coding Statistics 2026: Adoption, Productivity and Market Data
How Much Content on Social Media Is AI Generated Statistics
How Much Content on Social Media Is AI Generated Statistics 2026: Hidden Truths
ChatGPT vs DeepSeek Statistics
ChatGPT vs DeepSeek Statistics 2026: Users, Benchmarks & Pricing
ChatGPT vs Claude vs Gemini vs Perplexity Statistics
ChatGPT vs Claude vs Gemini vs Perplexity Statistics 2026: Users, Revenue & Market Share
How Many People Work At Midjourney
How Many People Work At Midjourney 2026: Lean Team, Big Revenue
Gaming
Gaming Statistics
Gaming Statistics 2026: Market Size, Players, Revenue, and Platforms
Roblox vs Minecraft Statistics
Roblox vs Minecraft Statistics 2026: Players, Revenue, Creators
Online Gambling Regulations Statistics
Online Gambling Regulations Statistics 2026: Global Compliance and Enforcement Data
Fantasy Sports Statistics
Fantasy Sports Statistics 2026: Users, Revenue & Trends
Apex Legends Statistics
Apex Legends Statistics 2026: Players, Revenue, and Esports
Fortnite Statistics
Fortnite Statistics 2026: Players, Revenue, Esports, and Engagement
Cybersecurity
Signal Statistics
Signal Statistics 2026: Users, Finances and Encryption Adoption
Password Statistics
Password Statistics 2026: Credential Theft, MFA, and the Passkey Tipping Point
Identity Theft Statistics
Identity Theft Statistics 2026: Key Fraud Data and Trends
CVE Statistics
CVE Statistics 2026: Severity Distribution and Top Affected Vendors
Dark Web AI Tool Marketplace Statistics
Dark Web AI Tool Marketplace Statistics 2026: Explosive Market Growth
API Security Breach Statistics
API Security Breach Statistics 2026: Hidden Threats
Categories
  • Cybersecurity
  • Artificial Intelligence
  • Internet
  • Technology
  • Gaming
Cybersecurity
Idscan Data Breach Confirmation
IDScan Confirms Massive Data Breach of Drivers License Records
Checkpoint Vpn Certification Flaw
Urgent Check Point VPN Flaws Expose Systems to RCE
Microsoft Bitlocker Rce Patch
Critical Windows BitLocker Flaw Sparks Urgent Patch
Cpanel Mailtrack Flaw Patched
cPanel Fixes Powerful Root Access Bug in EmailTrack
Quickfox Supply Chain Security
QuickFox Partners with Murphy Security to Strengthen Client Software Supply Chain Security
Chrome Zero Day Exploited Patched
Google Fixes 12 Chrome Flaws: Urgent Update Released
Artificial Intelligence
Openai Ends 1 Us Government Deal
OpenAI Ends $1 Government Deal, Offers 50% Discount
Openai Samsung Ai Chip Alliance
OpenAI Taps Samsung for Breakthrough Next-Gen Chips
Openai Agents Hijack German Wiki Site
OpenAI Agents Hijacked German Wiki, Researchers Say
Gpt 6 Astra Launched For Daybreak Users
OpenAI Releases GPT-6 Astra After Largest Training Run Yet
Claude Down Opus 5 And Fable 5
Claude Services Disrupted as Multiple Models Report Elevated Errors
Nvidia Huggingface Acquisition
Nvidia Confirms Acquisition of Hugging Face for $12.9 Billion
Internet
Apple Wallet Ids Launch In Oklahoma
Apple Wallet IDs Launch in Oklahoma in Major Expansion
Meta to Pay 18 Billion in Landmark Teen Safety Deal
Meta to Pay $18 Billion in Landmark Teen Safety Deal
Whatsapp Brings Passkeys 2fa
WhatsApp Hits 1 Billion Passkey Users, Adds 2FA Passwords
Apple Eu App Store Fee Reduction
Apple Sets New EU App Store Fees, Effective October 1
Github Outage Aug 2026
GitHub Down: Outage Hits Thousands of Users Worldwide
Russia S Fsb Charges Telegram Founder Durov With Terrorism
Russia’s FSB Charges Telegram Founder Durov With Terrorism
Technology
Snapchat Social Event Planning Feature
Snap Brings Social Event Planning Feature With Private Invites
Apple Iphone 18 And 18 Pro Launched
iPhone 18 Pro Debuts With Breakthrough Camera Upgrades
Iphone Foldable Launch Rumours Mark Gurmann
Apple Foldable iPhone To Top $2,000 In Leaked Roadmap
Eu Dsa Chatgpt Reddit Roblox Compliance
EU Expands Powerful DSA Oversight to ChatGPT and Reddit
Apple Confirms September 9 iPhone Event Under CEO Ternus
Apple Confirms September 9 iPhone Event Under CEO Ternus
Apple Mac Studio M5 Chip
New Mac Studio M5 Ultra Brings Massive On-Device AI Power
Gaming
Xbox Live Down Again
Xbox Live Down Again: Sign-In Error 0x80004005 Hits Players
Gta Vi Official Cover Art
GTA 6 Pre-Orders Start June 25, New Cover Art Unveiled
Epic Games Teases Unreal Engine 6 For Rocket League
Epic Games Teases Unreal Engine 6 for Rocket League
Stardew Valley Launched For Nintendo Switch 2 Edition
Stardew Valley Switch 2 Edition Arrives with Online Co-op
Hogwarts Legacy Game Crosses 40m Downloads
Hogwarts Legacy Crosses 40M Sales, Beating Industry Giants
Pubg Black Budget Closed Alpha Launched
PUBG: Black Budget Launches Closed Alpha Test With a Bold PvPvE Twist