GenAIWiki

Structured cards

Model database

Filter by provider, architecture family, or full-text search across descriptions.

Frontier models

Verified flagships and recent launches across major providers—scroll sideways for the full shelf.

SpaceXAI

Grok 4.6

FrontierLatest

Grok 4.6 is SpaceXAI's frontier model for coding, long-running agents, knowledge work, and interactive visual projects. Official documentation lists text and image input, text output, a 500,000-token context window, configurable low through xhigh reasoning, function calling, web and X search, code execution, and the API model ID grok-4.6.

FeaturedUpdated 9 days ago
xaispacexai

Google

Gemini 3.7 Flash

FrontierLatest

Gemini 3.7 Flash is Google's workhorse model for coding, agents, web development, knowledge work, and multimodal workflows. Google documents availability through the Gemini API, Google AI Studio, Android Studio, Gemini Enterprise Agent Platform, and Gemini Spark, with a 1M-token input context and up to 64K output.

FeaturedUpdated 9 days ago
googlegemini

Meta

Muse Spark 1.2

FrontierLatest

Muse Spark 1.2 is Meta's hosted coding and agent model for code generation, complex debugging, codebase understanding, long-horizon workflows, and tool use. It is available through Muse Code and Meta Model API with a 1M-token context window and separate standard and contributor data-use tiers.

FeaturedUpdated 9 days ago
metamuse

Alibaba Qwen

Qwen3.8-Max

FrontierLatest

Qwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents. QwenCloud documents text, image, and video input, text output, a 1M-token context, up to 131K output, built-in tools, and the API model ID qwen3.8-max.

FeaturedUpdated 9 days ago
alibabaqwen

Z.ai

GLM-5.3

FrontierLatest

GLM-5.3 is Z.ai's post-trained coding and agent model built on the same base model as GLM-5.2. Z.ai positions it for complex coding, long-horizon tasks, and cybersecurity evaluation, with mandatory thinking, low, high, and max reasoning effort, and availability through GLM Coding Plan and ZCode at launch.

FeaturedUpdated 9 days ago
zaiglm

OpenAI

GPT-5.6 Sol

FrontierLatest

GPT-5.6 Sol is OpenAI's strongest GPT-5.6-class general-purpose model for coding, research, and defensive cybersecurity workflows. In the August 10, 2026 Daybreak expansion, Sol is the recommended starting point for most vetted defenders under Daybreak Blue, where system-level cyber guardrails are adjusted for authorized security work without switching to a purpose-trained cyber model.

FeaturedUpdated 9 days ago
openaigpt-5-6

Moonshot AI

Kimi K3

FrontierLatest

Kimi K3 is Moonshot AI's July 2026 multimodal model for long-horizon coding, reasoning, and knowledge work. Official Kimi materials document native vision, a one-million-token context window, thinking-only API behavior, and access through Kimi, Kimi Work, Kimi Code, and the Kimi API.

FeaturedUpdated 9 days ago
frontierchina

Meta

Muse Glimmer 30B

FrontierLatest

Muse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.

FeaturedUpdated 9 days ago
metamuse

Sarvam AI

Sarvam 105B

FrontierLatest

Sarvam 105B is Sarvam AI's flagship 105B+ parameter Mixture-of-Experts reasoning model for Indian-language and English chat, complex reasoning, coding, long-context document analysis, and agentic tool-use workflows. Sarvam documents it as a 128K-context OpenAI-compatible chat model with Multi-head Latent Attention, 12T tokens of pre-training data, Apache 2.0 open weights, and production use powering Indus. Its strongest fit is Indian-language enterprise assistants, multilingual reasoning, and agent workflows where native script, romanized, and code-mixed inputs matter.

Updated 9 days ago
sarvamindian languages

Anthropic

Claude Fable 5

FrontierLatest

Anthropic's highest-capability widely released Claude model, documented for deep reasoning, codebase-scale work, long-context enterprise workloads, and multimodal inputs.

FeaturedUpdated 9 days ago
frontierclaude

Alibaba Qwen

Qwen3.8-27B

FrontierLatest

Qwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.

FeaturedUpdated 9 days ago
alibabaqwen

OpenAI

GPT-5.6-Cyber

FrontierLatest

GPT-5.6-Cyber is OpenAI's cybersecurity-specific model announced August 10, 2026 for Daybreak Red. Built on GPT-5.6 Sol, it is trained to improve specialized cyber tasks for trusted, authorized defenders—such as vulnerability research and exploit-chain development in approved environments—and to reduce refusals on certain higher-risk dual-use security prompts that general Sol still blocks.

FeaturedUpdated 9 days ago
openaigpt-5-6

All models

Filter and paginate the full catalog. Tabs control lifecycle scope.

Alibaba Qwen

Qwen3.6-27B

Legacy

Qwen3.6-27B is Alibaba Qwen's Apache 2.0 open-weight multimodal model for coding, repository-level reasoning, tool-driven workflows, and long-context tasks. The official model card documents a 27B language model with a vision encoder, 262,144 tokens of native context, optional extension to 1,010,000 tokens, and support in Transformers, vLLM, SGLang, and KTransformers.

FeaturedUpdated 9 days ago
qwenalibaba

Z.ai

GLM-5.2

Legacy

GLM-5.2 is Z.ai's open-weight model for long-horizon reasoning, coding, and agent workflows. Official Z.ai and Hugging Face materials document a one-million-token context window, flexible reasoning effort, a mixture-of-experts architecture, and MIT-licensed weights.

FeaturedUpdated 9 days ago
frontierchina

Alibaba Qwen

Qwen3.8-Max-Preview

Legacy

Qwen3.8-Max-Preview is Alibaba's July 2026 preview of its newest Qwen Max model for coding, agentic, and general reasoning workflows. Alibaba's Qwen Code materials identify qwen3.8-max as qwen3.8-max-preview; because this is a preview, endpoint behavior, limits, pricing, and availability should be treated as changeable.

FeaturedUpdated 9 days ago
frontierchina

xAI

Grok 4.5

Legacy

Grok 4.5 is xAI's frontier model for coding, agentic tasks, and knowledge work, available through the xAI API and in Cursor as a first-party model option.

FeaturedUpdated 9 days ago
frontiercoding

Anthropic

Claude Opus 4.7

Legacy

Anthropic's most capable generally available Claude model for complex reasoning and agentic coding, documented in the Claude model overview.

FeaturedUpdated 9 days ago
frontierreasoning

Anthropic

Claude Opus 4.8

Legacy

Anthropic's current Opus-tier Claude model, documented for complex reasoning, coding, and multimodal enterprise workloads below the newer Fable tier.

FeaturedUpdated 9 days ago
frontierclaude

OpenAI

GPT-5.5

Legacy

OpenAI's current flagship model for complex reasoning, coding, and professional work, documented in the OpenAI API model guide as the default starting point for high-complexity workloads.

FeaturedUpdated 9 days ago
frontierreasoning

Google

Gemini 3.5 Flash

Legacy

Gemini 3.5 Flash is Google's stable Gemini 3-series Flash model for agentic and coding tasks where teams need strong performance with lower latency and cost than Pro.

Updated 9 days ago
googleflash

Google

Gemini 1.5 Pro

Legacy

Google DeepMind Gemini 1.5 Pro targets long-context multimodal workloads—large effective context for retrieval-heavy document pipelines, plus image, audio, and video inputs on supported surfaces. It is often paired with Vertex AI or the Gemini API for enterprise workloads on GCP.

FeaturedUpdated 9 days ago
googlelong-context

DeepSeek

DeepSeek-V3

Legacy

DeepSeek-V3 is a large-scale language model family noted for strong coding and math performance under open or research-friendly terms (verify the exact license for your deployment). Teams adopt it for cost-sensitive research, self-hosted inference, or comparison against frontier APIs.

FeaturedUpdated 9 days ago
researchcoding

Anthropic

Claude 3.5 Sonnet

Legacy

Anthropic’s balanced Sonnet-tier model tuned for long-context reasoning, careful instruction following, and strong performance on coding and analysis workloads. It is a common enterprise choice on the Anthropic API and on AWS Bedrock when teams need large context for RAG and document review.

FeaturedUpdated 9 days ago
codingagents

Mistral AI

Mistral Large 2

Legacy

Mistral’s frontier-class multilingual model emphasizing JSON adherence, agent-friendly behavior, and competitive reasoning within the Mistral API ecosystem. European teams often evaluate it for GDPR-adjacent deployment patterns alongside US-hosted alternatives.

FeaturedUpdated 9 days ago
euenterprise

OpenAI

GPT-5.4

Legacy

OpenAI's GPT-5.4 model, documented in the official OpenAI API model guide as part of the current GPT-5 family below the GPT-5.5 flagship lane.

FeaturedUpdated 9 days ago
frontieropenai

OpenAI

GPT-4o

Legacy

GPT-4o is an OpenAI multimodal model that accepts text and image inputs and produces text. It supports streaming, function calling, Structured Outputs, fine-tuning, and predicted outputs for vision-heavy assistants and structured extraction workflows.

FeaturedUpdated 9 days ago
frontiermultimodal

xAI

Grok-2

Legacy

Grok-2 is xAI’s flagship chat model positioned for real-time knowledge integrations and high-throughput conversational products on xAI’s API. Availability and pricing evolve—treat capabilities as vendor-specific.

FeaturedUpdated 9 days ago
frontierapi

Alibaba

Qwen 2.5 72B Instruct

Legacy

Qwen 2.5 72B Instruct is a large multilingual open-weights model from Alibaba’s Qwen family with strong coding and general chat performance. Common in APAC deployments and on Hugging Face inference endpoints—check license terms for commercial use.

FeaturedUpdated 9 days ago
open-weightsmultilingual

Anthropic

Claude Sonnet 4.6

Legacy

Anthropic's Sonnet-tier model documented as the best combination of speed and intelligence in the Claude model overview.

FeaturedUpdated 9 days ago
frontiersonnet

OpenAI

o1

Legacy

OpenAI’s o1 series emphasizes extended internal reasoning before answering—useful for competition-style math, complex debugging, and multi-step planning where latency is acceptable. It behaves differently from standard chat models: tune prompts for chain-of-thought style tasks and measure time-to-first-token.

FeaturedUpdated 9 days ago
reasoningstem

Google

Gemini 2.5 Flash-Lite

Legacy

Google's fastest and most budget-friendly multimodal model in the Gemini 2.5 family, according to the Gemini API model documentation.

Updated 9 days ago
googleflash-lite

DeepSeek

DeepSeek-V3.2

Legacy

DeepSeek's documented successor to the V3.2 experimental line, positioned in official DeepSeek API news as live on app, web, and API.

FeaturedUpdated 9 days ago
frontierdeepseek

OpenAI

GPT-5.4 mini

Legacy

OpenAI's smaller GPT-5.4 mini model, documented in the official OpenAI API model guide for lower-latency or lower-cost GPT-5 family workloads.

Updated 9 days ago
openaimini

OpenAI

GPT-5.4 nano

Legacy

OpenAI's smallest GPT-5.4 nano model, documented in the official OpenAI API model guide for very low-latency or economical GPT-5 family routing.

Updated 9 days ago
openainano

Google

Gemini 1.5 Flash

Legacy

Gemini 1.5 Flash targets low-latency, cost-efficient multimodal chat and retrieval workloads on the Gemini API and Vertex AI. It keeps much of the long-context family behavior with faster responses for interactive apps.

Updated 9 days ago
googlelatency

Anthropic

Claude 3.5 Haiku

Legacy

Claude 3.5 Haiku is Anthropic’s fast, cost-efficient tier for high-volume classification, routing, and simple chat. It targets latency-sensitive paths and agent pre-processing before escalating to Sonnet-class models.

Updated 9 days ago
latencycost

Recommended

Top current models by information quality score — good defaults when you are not sure where to start.

Missing a frontier release? Add a model (editors)