AI model database

Major large language models: specs, pricing and platform support

From OpenAI, Anthropic, Google and xAI to DeepSeek, Qwen, Kimi and GLM — current flagship models, release timelines, context windows, API pricing and whether each can be used inside Edor.ai. Every figure links to official documentation so schools can verify it.

Prices are the vendors' published API rates (USD per 1M tokens) and do not represent what a school spends on Edor.ai — our plans are billed as a per-school annual fee with no token charges.

Prices and specifications in this article are current as of 2026-09

Native support

OpenAI

GPT-6 Astra
Released
2026-09-03
Context window
1,050K tokens
API pricing
$10 / $50

Native provider. OpenAI is the platform default and also supplies the embedding, speech-to-text and text-to-speech models behind the knowledge base and oral practice. Administrators can switch models in the admin panel.

Read the full page
Native support

Anthropic

Claude Fable 5.1
Released
2026-09-01
Context window
1,000K tokens
API pricing
$10 / $50

Native provider. Claude is reliable for Chinese writing feedback and long-document work, but it has no embedding or speech models, so the knowledge base and oral practice still need an OpenAI or Azure key.

Read the full page
Via Poe

Google

Gemini 3.8 Flash
Released
2026-09-02
Context window
1,049K tokens
API pricing
$0.75 / $3.75

Not a native provider. Schools can reach the Gemini family through the Poe provider, and Google's open-weight Gemma family can be self-hosted with Ollama. Schools already on Google Workspace for Education can run both side by side.

Read the full page
Via Poe

xAI

Grok 4.6
Released
2026-08-12
Context window
500K tokens
API pricing
$2 / $6

Not a native provider; reachable through Poe. Grok is tightly integrated with X platform content, so schools should assess its content-safety posture before using it with students.

Read the full page
Self-hosted via OllamaOpen weights · Apache 2.0

Meta

Muse Glimmer 30B
Released
2026-08-09
Context window
131K tokens
API pricing
$0.35 / $1.5

Not a native provider, but Meta's open-weight models are among the most practical choices for on-premises use: Muse Glimmer 30B ships under Apache 2.0 and runs on a single consumer GPU via Ollama or LM Studio, so student data never leaves the campus network. Note that in April 2026 Meta moved its frontier line from Llama to the closed-weight Muse Spark, making the Llama family legacy.

Read the full page
Self-hosted via OllamaOpen weights · Apache 2.0

Mistral AI

Mistral Large 3
Released
2025-12
Context window
256K tokens
API pricing
$0.5 / $1.5

Not a native provider. Mistral releases its whole line under Apache 2.0 and is the most-cited choice under European data-governance frameworks; the small Ministral 3B / 8B models suit older school hardware or offline use.

Read the full page
Not supported

Cursor (Anysphere)

Composer 2.5
Released
2026-05-18
Context window
Not published
API pricing
$0.5 / $2.5

Not applicable to a school platform. Composer is a coding-specific model built into the Cursor editor, and the documentation states it is available only inside Cursor's own products, with no inference API that can be called as a model endpoint. Cursor does offer a Cloud Agents API and SDK, but those drive coding agents inside Cursor's environment rather than exposing a chat model a learning platform can call, and it is not designed for classrooms or students. Teachers can use Cursor as a tool when teaching programming, but the platform's model choice belongs among the other providers listed here.

Read the full page
Self-hosted via OllamaOpen weights · MIT

DeepSeek

DeepSeek-V4.1-Flash
Released
2026-09-10
Context window
1,000K tokens
API pricing
$0.15 / $0.6

Not a native provider. The practical route is to run a distilled or quantised build on a school server through Ollama so no data leaves the campus network. Calling DeepSeek's own API is a cross-border transfer and needs a privacy assessment first.

Read the full page
Self-hosted via OllamaOpen weights · 自訂授權 / Apache 2.0(視型號)

Alibaba Qwen

Qwen3.8-Max
Released
2026-09-03
Context window
1,000K tokens
API pricing
$2 / $6

Not a native provider. Qwen is the open family most respected for Chinese-language work, and Edor.ai's Ollama suggestions already include `qwen2.5:14b`. The newer Qwen3.8-27B ships under Apache 2.0 and is documented to run in 24GB of VRAM, making it a practical on-premises choice for Chinese material. The flagship Qwen3.8-Max is a closed cloud API and a cross-border transfer, so student data needs assessment first.

Read the full page
Not supportedOpen weights · Kimi K3 License

Moonshot AI (Kimi)

Kimi K3
Released
2026-07-16
Context window
1,000K tokens
API pricing
$3 / $15

Not wired up today. Kimi K3 publishes weights, but at 2.8T parameters (about 104B active) it is far beyond what a school server can host, so self-hosting is unrealistic. Its official API sits in mainland China and is a cross-border transfer. Schools wanting to try its long-document strength should have teachers use a personal account on de-identified material only.

Read the full page
Self-hosted via OllamaOpen weights · MIT(GLM-5.2 與 5.3-Flash)

Zhipu AI / Z.ai (GLM)

GLM-5.3
Released
2026-08-14
Context window
1,000K tokens
API pricing
$1.4 / $4.4

Not a native provider. GLM is one of the few teams publishing large models under MIT (GLM-5.2 and GLM-5.3-Flash), which appeals to sponsoring bodies wanting full autonomy, though the parameter counts still need server-class hardware and weight releases do not land at the same time for every version — verify against the official Hugging Face page before committing.

Read the full page
Not supportedOpen weights · 自訂授權(見 Hugging Face model card)

MiniMax

MiniMax M3
Released
2026-06
Context window
1,000K tokens
API pricing
$0.3 / $1.2

Not wired up today. MiniMax M3 takes a sparse-attention, low-cost long-context approach — about 428B total parameters with only ~23B active per token — making it one of the cheapest long-context models in its class. Its API is in mainland China and its licence terms need line-by-line review, so an IT coordinator should assess privacy and licensing first.

Read the full page
Not supported

ByteDance Doubao

Doubao Seed 2.1 Pro
Released
2026-06-23
Context window
256K tokens
API pricing
Not published

Not wired up, and not recommended for student-facing use in Hong Kong schools. Doubao is served only through Volcano Engine in mainland China, its weights are closed, its pricing is in RMB (Seed 2.1 Pro at ¥6 input / ¥30 output per 1M tokens), and its terms and data location sit under mainland regulation. Where Chinese-language strength is the goal, self-hosting Qwen or GLM open weights is the more controllable route.

Read the full page
Self-hosted via OllamaOpen weights · Apache 2.0

Tencent Hunyuan

Hunyuan Hy4 Preview
Released
2026-08-28
Context window
1,000K tokens
API pricing
$0.83 / $?

Not a native provider, but notable: Hunyuan Hy4 Preview publishes 770B-parameter weights (49B active) under a fully permissive Apache 2.0 licence, which is clearer than the bespoke licences on GLM-5.3 or Kimi K3. Self-hosting still needs server-class hardware, and calling Tencent Cloud's API directly is a cross-border transfer.

Read the full page
Not supportedOpen weights · ERNIE 4.5 系列開源;5.x 未公開

Baidu ERNIE

ERNIE 5.1
Released
2026-08
Context window
128K tokens
API pricing
$0.55 / $?

Not wired up today. ERNIE's main strength is search-grounded answers from Baidu, which is precisely where Hong Kong schools should be careful: its sources, content policy and data handling all sit in the mainland framework. Its 128K context window is also well below other current flagships.

Read the full page

Latest release timeline

Recent public releases across vendors, newest first. Generations turn over quickly, so ask any vendor whether model upgrades are already included in the annual fee.

  1. 2026-09-10DeepSeek-V4.1-FlashDeepSeek · FlagshipThe current model: a 552B-parameter MoE that activates only 8B (prefill) to 16B (decode) at inference, with native image support. The $0.15 / $0.60 rate is off-peak; peak hours double it.
  2. 2026-09-03GPT-6 AstraOpenAI · FlagshipCurrent flagship. `reasoning.effort` from low to max; any request above 272K input tokens is billed at 2x input rates for the whole request.
  3. 2026-09-03Qwen3.8-Max-0902Alibaba Qwen · FlagshipThe newest post-trained snapshot of the flagship: a 2.4T-parameter MoE activating about 95B per token, with text, image and video input. This snapshot is cloud-API only and its weights are not published.
  4. 2026-09-02Gemini 3.8 FlashGoogle · FlagshipThe current workhorse with `thinking_level` set to low, medium or high. The $0.75 / $3.75 rate is introductory; standard pricing of $1.50 / $7.50 starts 1 January 2027.
  5. 2026-09-02Gemini 3.8 Flash CyberGoogle · FlagshipA cybersecurity variant offered only to trusted defenders through the Fairwind Program.
  6. 2026-09-01Claude Fable 5.1Anthropic · FlagshipThe current top tier with adaptive thinking always on. Cache reads dropped to $0.25 per 1M tokens, the lowest in the family.
  7. 2026-08-28Hunyuan Hy4 PreviewTencent Hunyuan · Flagship770B total and 49B active parameters with a 1M-token context window under Apache 2.0; Tencent's own blind evaluation places it slightly ahead of GLM-5.3 and Kimi K3.
  8. 2026-08-26Qwen3.8-FlashAlibaba Qwen · LightweightA small MoE with 125B total and 6B active parameters, extending context from 256K to 1M at a very low price.
  9. 2026-08-26GLM-5.3-FlashZhipu AI / Z.ai (GLM) · LightweightA natively multimodal MoE with 320B total and 18B active parameters under MIT, at a promotional $0.075 / $0.25 per 1M tokens (list $0.15 / $0.50).
  10. 2026-08-21Doubao-Seed-EvolvingByteDance Doubao · FlagshipA continuously iterating model updated at least weekly; a single fixed model ID always resolves to the newest build.
  11. 2026-08-16DeepSeek-V4-Pro / V4-FlashDeepSeek · FlagshipThe V4 generation that introduced peak / off-peak pricing. From 14 September 2026 V4-Pro requests route to V4.1-Flash.
  12. 2026-08-14Qwen3.8-27BAlibaba Qwen · Open weightsA 27.8B dense vision-language model under Apache 2.0, documented to run in 24GB of VRAM — the most realistic of this batch for school self-hosting.
  13. 2026-08-14GLM-5.3Zhipu AI / Z.ai (GLM) · FlagshipThe current flagship aimed at coding and agentic work, priced the same as GLM-5.2. Check the official Hugging Face page for the current status of the flagship weights.
  14. 2026-08-13Gemini 3.7 FlashGoogle · WorkhorseThe previous version at the same price, the second of three Flash iterations inside six weeks.

How a school should read this table

  • Do not shop on token price alone. Teacher tools generate limited tokens; the real cost difference is whether you must buy a separate licence per teacher.
  • Context length decides how much school material the AI can read at once, but long context does not mean accurate recall — see the LLM Classroom lesson on "lost in the middle".
  • The value of open weights is not that they are free but that they can run on your own server, so student data never leaves the campus network.
  • Support status changes as we add providers. If your school requires a specific model, raise it during quotation and we will assess it.
Subscribe to the AI in Education newsletter

One email a month: practical AI teaching articles for Hong Kong schools, platform updates and grant news. Unsubscribe any time.

We only use this address for the newsletter and never share it.