Structured cards
Model database
Filter by provider, architecture family, or full-text search across descriptions.
Frontier models
Verified flagships and this month’s launches—Claude Opus 5.5, GPT-6 Sol and Luna, GPT-6 Astra, Claude Fable 5.1, Gemini 3.8, and more. Scroll sideways for the full shelf.
By company
Latest verified models from each frontier company, summarized on one page.
All models
Filter and paginate the full catalog. Tabs control lifecycle scope.
Gemini 3.5 Flash
LegacyGemini 3.5 Flash is Google's stable Gemini 3-series Flash model for agentic and coding tasks where teams need strong performance with lower latency and cost than Pro.
Gemini 1.5 Pro
LegacyGoogle DeepMind Gemini 1.5 Pro targets long-context multimodal workloads—large effective context for retrieval-heavy document pipelines, plus image, audio, and video inputs on supported surfaces. It is often paired with Vertex AI or the Gemini API for enterprise workloads on GCP.
Gemini 2.5 Flash-Lite
LegacyGoogle's fastest and most budget-friendly multimodal model in the Gemini 2.5 family, according to the Gemini API model documentation.
Gemini 1.5 Flash
LegacyGemini 1.5 Flash targets low-latency, cost-efficient multimodal chat and retrieval workloads on the Gemini API and Vertex AI. It keeps much of the long-context family behavior with faster responses for interactive apps.
Gemini 1.0 Pro
LegacyGemini 1.0 Pro represents Google’s first broadly marketed Gemini-era general model for text and basic multimodal tasks on Vertex and consumer surfaces. New projects should prefer 1.5+ generations unless constrained by legacy integrations—verify availability.
Gemini 2.0 Flash
LegacyGemini 2.0 Flash is Google’s efficiency-oriented multimodal model generation aimed at fast agentic and interactive experiences. Capabilities and naming evolve—validate against the current Gemini API reference for tool use and context limits.
Recommended
Top current models by information quality score — good defaults when you are not sure where to start.
Anthropic
Claude Opus 5.5
CurrentLatestClaude Opus 5.5 is Anthropic's September 22, 2026 Opus-class model for long-running agentic coding and knowledge work, and the first model in the Claude 5.5 family. The Claude API ID is claude-opus-5-5, with a 1M-token context window, 128K synchronous max output, adaptive thinking that cannot be disabled, and default effort medium. Anthropic prices it at $4 input and $20 output per million tokens and says typical workloads cost about 40% less to run than Opus 5.
OpenAI
GPT-Image-2.5 Sunburst
CurrentLatestGPT-Image-2.5 Sunburst is OpenAI's September 8, 2026 most capable image generation and editing model. The API ID is gpt-image-2.5-sunburst, with dated snapshot gpt-image-2.5-sunburst-2026-09-08. It takes text and image input and outputs images. OpenAI positions it for workflows where editing precision matters most, with quality settings low, medium, high, xhigh, max, and auto. Call it on the Images API or as the model of the Responses API image generation tool.
Alibaba Qwen
Qwen3.8-27B
CurrentLatestQwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.
Meta
Muse Glimmer 30B
CurrentLatestMuse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.
OpenAI
GPT-6 Sol
CurrentLatestGPT-6 Sol is OpenAI's September 22, 2026 model for complex coding and agentic workflows, positioned as a faster and cheaper GPT-6 lane than GPT-6 Astra. The API model ID is gpt-6-sol, with a 1,050,000-token context window, 128,000 max output tokens, text and image input, and reasoning.effort from none through max (default medium). Standard text pricing is $2 input and $10 output per million tokens.
Missing a frontier release? Add a model (editors)