Structured cards
Model database
Filter by provider, architecture family, or full-text search across descriptions.
Frontier models
Verified flagships and this month’s launches—Claude Opus 5.5, GPT-6 Sol and Luna, GPT-6 Astra, Claude Fable 5.1, Gemini 3.8, and more. Scroll sideways for the full shelf.
By company
Latest verified models from each frontier company, summarized on one page.
All models
Filter and paginate the full catalog. Tabs control lifecycle scope.
Alibaba Qwen
Qwen3.8-Max
CurrentLatestQwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents. Effective September 5, 2026 (UTC+8), the qwen3.8-max endpoint automatically serves the September 2 snapshot qwen3.8-max-0902 (alias qwen3.8-max-2026-09-02). QwenCloud says 0902 improves coding depth, multi-tool agent delivery, and visual understanding while keeping the 1M context, thinking mode, tools, and unchanged billing.
Alibaba Qwen
Qwen3.8-27B
CurrentLatestQwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.
Alibaba Qwen
Qwen3.8-Flash-Next
CurrentLatestQwen3.8-Flash-Next is Qwen's experimental open-weight preview of the architecture intended to underpin Qwen4. The official Hugging Face card documents a 125B MoE with 6B activated parameters plus 51B n-gram embeddings and a 4B MTP head, native 262,144-token context extensible to 1M, and text, image, and video input.
Alibaba Qwen
Qwen3.8-2.4T-A95B
CurrentLatestQwen3.8-2.4T-A95B is Qwen's downloadable 2.4-trillion-parameter sparse mixture-of-experts model with 95B activated parameters. The official model card documents text-only input and output, mandatory thinking, 262,144 native context extensible to roughly 1.01M, configurable reasoning effort, and a dedicated Qwen3.8-Max license.
Alibaba Qwen
Qwen-Image-3.0
CurrentLatestQwen-Image-3.0 is Alibaba's third-generation image model for image generation, editing, complex layouts, and multilingual text rendering. The official Qwen release highlights inputs up to 4.5K, small text rendering down to 10 pixels, and native support for twelve languages.
Alibaba Qwen
Qwen3.6-27B
LegacyQwen3.6-27B is Alibaba Qwen's Apache 2.0 open-weight multimodal model for coding, repository-level reasoning, tool-driven workflows, and long-context tasks. The official model card documents a 27B language model with a vision encoder, 262,144 tokens of native context, optional extension to 1,010,000 tokens, and support in Transformers, vLLM, SGLang, and KTransformers.
Alibaba Qwen
Qwen3.8-Max-Preview
LegacyQwen3.8-Max-Preview is Alibaba's July 2026 preview of its newest Qwen Max model for coding, agentic, and general reasoning workflows. Alibaba's Qwen Code materials identify qwen3.8-max as qwen3.8-max-preview; because this is a preview, endpoint behavior, limits, pricing, and availability should be treated as changeable.
Recommended
Top current models by information quality score — good defaults when you are not sure where to start.
Anthropic
Claude Opus 5.5
CurrentLatestClaude Opus 5.5 is Anthropic's September 22, 2026 Opus-class model for long-running agentic coding and knowledge work, and the first model in the Claude 5.5 family. The Claude API ID is claude-opus-5-5, with a 1M-token context window, 128K synchronous max output, adaptive thinking that cannot be disabled, and default effort medium. Anthropic prices it at $4 input and $20 output per million tokens and says typical workloads cost about 40% less to run than Opus 5.
OpenAI
GPT-Image-2.5 Sunburst
CurrentLatestGPT-Image-2.5 Sunburst is OpenAI's September 8, 2026 most capable image generation and editing model. The API ID is gpt-image-2.5-sunburst, with dated snapshot gpt-image-2.5-sunburst-2026-09-08. It takes text and image input and outputs images. OpenAI positions it for workflows where editing precision matters most, with quality settings low, medium, high, xhigh, max, and auto. Call it on the Images API or as the model of the Responses API image generation tool.
Alibaba Qwen
Qwen3.8-27B
CurrentLatestQwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.
Meta
Muse Glimmer 30B
CurrentLatestMuse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.
OpenAI
GPT-6 Sol
CurrentLatestGPT-6 Sol is OpenAI's September 22, 2026 model for complex coding and agentic workflows, positioned as a faster and cheaper GPT-6 lane than GPT-6 Astra. The API model ID is gpt-6-sol, with a 1,050,000-token context window, 128,000 max output tokens, text and image input, and reasoning.effort from none through max (default medium). Standard text pricing is $2 input and $10 output per million tokens.
Missing a frontier release? Add a model (editors)