GenAIWiki

Structured cards

Model database

Filter by provider, architecture family, or full-text search across descriptions.

Frontier models

Verified flagships and recent launches across major providers—scroll sideways for the full shelf.

SpaceXAI

Grok 4.6

FrontierLatest

Grok 4.6 is SpaceXAI's frontier model for coding, long-running agents, knowledge work, and interactive visual projects. Official documentation lists text and image input, text output, a 500,000-token context window, configurable low through xhigh reasoning, function calling, web and X search, code execution, and the API model ID grok-4.6.

FeaturedUpdated 9 days ago
xaispacexai

Google

Gemini 3.7 Flash

FrontierLatest

Gemini 3.7 Flash is Google's workhorse model for coding, agents, web development, knowledge work, and multimodal workflows. Google documents availability through the Gemini API, Google AI Studio, Android Studio, Gemini Enterprise Agent Platform, and Gemini Spark, with a 1M-token input context and up to 64K output.

FeaturedUpdated 9 days ago
googlegemini

Meta

Muse Spark 1.2

FrontierLatest

Muse Spark 1.2 is Meta's hosted coding and agent model for code generation, complex debugging, codebase understanding, long-horizon workflows, and tool use. It is available through Muse Code and Meta Model API with a 1M-token context window and separate standard and contributor data-use tiers.

FeaturedUpdated 9 days ago
metamuse

Alibaba Qwen

Qwen3.8-Max

FrontierLatest

Qwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents. QwenCloud documents text, image, and video input, text output, a 1M-token context, up to 131K output, built-in tools, and the API model ID qwen3.8-max.

FeaturedUpdated 9 days ago
alibabaqwen

Z.ai

GLM-5.3

FrontierLatest

GLM-5.3 is Z.ai's post-trained coding and agent model built on the same base model as GLM-5.2. Z.ai positions it for complex coding, long-horizon tasks, and cybersecurity evaluation, with mandatory thinking, low, high, and max reasoning effort, and availability through GLM Coding Plan and ZCode at launch.

FeaturedUpdated 9 days ago
zaiglm

OpenAI

GPT-5.6 Sol

FrontierLatest

GPT-5.6 Sol is OpenAI's strongest GPT-5.6-class general-purpose model for coding, research, and defensive cybersecurity workflows. In the August 10, 2026 Daybreak expansion, Sol is the recommended starting point for most vetted defenders under Daybreak Blue, where system-level cyber guardrails are adjusted for authorized security work without switching to a purpose-trained cyber model.

FeaturedUpdated 9 days ago
openaigpt-5-6

Moonshot AI

Kimi K3

FrontierLatest

Kimi K3 is Moonshot AI's July 2026 multimodal model for long-horizon coding, reasoning, and knowledge work. Official Kimi materials document native vision, a one-million-token context window, thinking-only API behavior, and access through Kimi, Kimi Work, Kimi Code, and the Kimi API.

FeaturedUpdated 9 days ago
frontierchina

Meta

Muse Glimmer 30B

FrontierLatest

Muse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.

FeaturedUpdated 9 days ago
metamuse

Sarvam AI

Sarvam 105B

FrontierLatest

Sarvam 105B is Sarvam AI's flagship 105B+ parameter Mixture-of-Experts reasoning model for Indian-language and English chat, complex reasoning, coding, long-context document analysis, and agentic tool-use workflows. Sarvam documents it as a 128K-context OpenAI-compatible chat model with Multi-head Latent Attention, 12T tokens of pre-training data, Apache 2.0 open weights, and production use powering Indus. Its strongest fit is Indian-language enterprise assistants, multilingual reasoning, and agent workflows where native script, romanized, and code-mixed inputs matter.

Updated 9 days ago
sarvamindian languages

Anthropic

Claude Fable 5

FrontierLatest

Anthropic's highest-capability widely released Claude model, documented for deep reasoning, codebase-scale work, long-context enterprise workloads, and multimodal inputs.

FeaturedUpdated 9 days ago
frontierclaude

Alibaba Qwen

Qwen3.8-27B

FrontierLatest

Qwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.

FeaturedUpdated 9 days ago
alibabaqwen

OpenAI

GPT-5.6-Cyber

FrontierLatest

GPT-5.6-Cyber is OpenAI's cybersecurity-specific model announced August 10, 2026 for Daybreak Red. Built on GPT-5.6 Sol, it is trained to improve specialized cyber tasks for trusted, authorized defenders—such as vulnerability research and exploit-chain development in approved environments—and to reduce refusals on certain higher-risk dual-use security prompts that general Sol still blocks.

FeaturedUpdated 9 days ago
openaigpt-5-6

All models

Filter and paginate the full catalog. Tabs control lifecycle scope.

Alibaba Qwen

Qwen3.8-27B

CurrentLatest

Qwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.

FeaturedUpdated 9 days ago
alibabaqwen

Alibaba Qwen

Qwen3.8-Max

CurrentLatest

Qwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents. QwenCloud documents text, image, and video input, text output, a 1M-token context, up to 131K output, built-in tools, and the API model ID qwen3.8-max.

FeaturedUpdated 9 days ago
alibabaqwen

Meta

Muse Spark 1.2

CurrentLatest

Muse Spark 1.2 is Meta's hosted coding and agent model for code generation, complex debugging, codebase understanding, long-horizon workflows, and tool use. It is available through Muse Code and Meta Model API with a 1M-token context window and separate standard and contributor data-use tiers.

FeaturedUpdated 9 days ago
metamuse

Google

Gemini 3.7 Flash

CurrentLatest

Gemini 3.7 Flash is Google's workhorse model for coding, agents, web development, knowledge work, and multimodal workflows. Google documents availability through the Gemini API, Google AI Studio, Android Studio, Gemini Enterprise Agent Platform, and Gemini Spark, with a 1M-token input context and up to 64K output.

FeaturedUpdated 9 days ago
googlegemini

SpaceXAI

Grok 4.6

CurrentLatest

Grok 4.6 is SpaceXAI's frontier model for coding, long-running agents, knowledge work, and interactive visual projects. Official documentation lists text and image input, text output, a 500,000-token context window, configurable low through xhigh reasoning, function calling, web and X search, code execution, and the API model ID grok-4.6.

FeaturedUpdated 9 days ago
xaispacexai

Meta

Muse Glimmer 30B

CurrentLatest

Muse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.

FeaturedUpdated 9 days ago
metamuse

Moonshot AI

Kimi K3

CurrentLatest

Kimi K3 is Moonshot AI's July 2026 multimodal model for long-horizon coding, reasoning, and knowledge work. Official Kimi materials document native vision, a one-million-token context window, thinking-only API behavior, and access through Kimi, Kimi Work, Kimi Code, and the Kimi API.

FeaturedUpdated 9 days ago
frontierchina

Sarvam AI

Sarvam 105B

CurrentLatest

Sarvam 105B is Sarvam AI's flagship 105B+ parameter Mixture-of-Experts reasoning model for Indian-language and English chat, complex reasoning, coding, long-context document analysis, and agentic tool-use workflows. Sarvam documents it as a 128K-context OpenAI-compatible chat model with Multi-head Latent Attention, 12T tokens of pre-training data, Apache 2.0 open weights, and production use powering Indus. Its strongest fit is Indian-language enterprise assistants, multilingual reasoning, and agent workflows where native script, romanized, and code-mixed inputs matter.

Updated 9 days ago
sarvamindian languages

OpenAI

GPT-5.6-Cyber

CurrentLatest

GPT-5.6-Cyber is OpenAI's cybersecurity-specific model announced August 10, 2026 for Daybreak Red. Built on GPT-5.6 Sol, it is trained to improve specialized cyber tasks for trusted, authorized defenders—such as vulnerability research and exploit-chain development in approved environments—and to reduce refusals on certain higher-risk dual-use security prompts that general Sol still blocks.

FeaturedUpdated 9 days ago
openaigpt-5-6

Sarvam AI

Sarvam 30B

CurrentLatest

Sarvam 30B is a 30B parameter Mixture-of-Experts chat and reasoning model from Sarvam AI, optimized for Indian languages, real-time conversation, high-throughput voice-agent pipelines, coding, and practical deployment. Sarvam documents 2.4B active parameters per token, 16T tokens of pre-training data, a 64K context window, Grouped Query Attention, Apache 2.0 open weights, and OpenAI-compatible chat completions.

Updated 9 days ago
sarvamindian languages

MiniMax

MiniMax M3

CurrentLatest

MiniMax M3 is a June 2026 open-weight multimodal model for coding, agentic workflows, computer use, and long-context work. MiniMax documents a one-million-token context window, native image and video understanding, and deployment through hosted or downloadable model paths.

FeaturedUpdated 9 days ago
frontierchina

Alibaba Qwen

Qwen3.8-2.4T-A95B

CurrentLatest

Qwen3.8-2.4T-A95B is Qwen's downloadable 2.4-trillion-parameter sparse mixture-of-experts model with 95B activated parameters. The official model card documents text-only input and output, mandatory thinking, 262,144 native context extensible to roughly 1.01M, configurable reasoning effort, and a dedicated Qwen3.8-Max license.

FeaturedUpdated 9 days ago
alibabaqwen

Microsoft Research

VibeVoice-ASR-BitNet

CurrentLatest

VibeVoice-ASR-BitNet is Microsoft Research's compressed automatic speech recognition model for real-time CPU and edge transcription without a GPU. Its official model card documents a 1.58 GB footprint, seven-language support, and real-time CPU performance under the tested configuration.

FeaturedUpdated 9 days ago
microsoftspeech-to-text

OpenAI

GPT-5.6 Sol

CurrentLatest

GPT-5.6 Sol is OpenAI's strongest GPT-5.6-class general-purpose model for coding, research, and defensive cybersecurity workflows. In the August 10, 2026 Daybreak expansion, Sol is the recommended starting point for most vetted defenders under Daybreak Blue, where system-level cyber guardrails are adjusted for authorized security work without switching to a purpose-trained cyber model.

FeaturedUpdated 9 days ago
openaigpt-5-6

Alibaba Qwen

Qwen-Image-3.0

CurrentLatest

Qwen-Image-3.0 is Alibaba's third-generation image model for image generation, editing, complex layouts, and multilingual text rendering. The official Qwen release highlights inputs up to 4.5K, small text rendering down to 10 pixels, and native support for twelve languages.

FeaturedUpdated 9 days ago
chinaqwen

Z.ai

GLM-5.3

CurrentLatest

GLM-5.3 is Z.ai's post-trained coding and agent model built on the same base model as GLM-5.2. Z.ai positions it for complex coding, long-horizon tasks, and cybersecurity evaluation, with mandatory thinking, low, high, and max reasoning effort, and availability through GLM Coding Plan and ZCode at launch.

FeaturedUpdated 9 days ago
zaiglm

Microsoft Research

Mage-Flow

CurrentLatest

Mage-Flow is Microsoft Research's MIT-licensed 4B image generation and editing family. The official model card provides base, reinforcement-learning-aligned, turbo, and editing variants for teams evaluating downloadable image models and custom serving.

FeaturedUpdated 9 days ago
microsoftopen-weights

Anthropic

Claude Sonnet 5

CurrentLatest

Claude Sonnet 5 is Anthropic's balanced current Sonnet model for agentic coding, tool use, and production knowledge work, offering a stronger cost-performance lane than prior Sonnet releases.

FeaturedUpdated 9 days ago
frontiersonnet

Anthropic

Claude Opus 5

CurrentLatest

Claude Opus 5 is Anthropic's advanced Claude model for complex agentic coding and enterprise work, with a 1M-token context window, adaptive thinking, and strong long-horizon task behavior.

FeaturedUpdated 9 days ago
frontiercoding

Anthropic

Claude Fable 5

CurrentLatest

Anthropic's highest-capability widely released Claude model, documented for deep reasoning, codebase-scale work, long-context enterprise workloads, and multimodal inputs.

FeaturedUpdated 9 days ago
frontierclaude

OpenAI

GPT-5.6 Terra

CurrentLatest

OpenAI's GPT-5.6 Terra is the balanced GPT-5.6 tier for teams that want strong reasoning and coding quality at a lower cost than the Sol flagship lane.

Updated 9 days ago
reasoningcoding

Mistral AI

Mistral Medium 3.5

CurrentLatest

Mistral Medium 3.5 is Mistral's frontier-class multimodal model optimized for agentic and coding use cases, released as open weights under a modified MIT license.

FeaturedUpdated 9 days ago
frontiermistral

AWS

Amazon Nova

CurrentLatest

Amazon Nova is AWS’s multimodal foundation model family for text, image, and video workloads delivered through Amazon Bedrock with enterprise IAM, VPC, and governance patterns. Model IDs and modalities vary by region—verify Bedrock model access lists.

FeaturedUpdated 9 days ago
awsbedrock

Stability AI

Stable Diffusion XL

CurrentLatest

Stable Diffusion XL (SDXL) 1.0 is Stability AI's latent diffusion text-to-image model for native 1024x1024 generation. The base model can run standalone or feed an optional refiner for the final denoising steps, and the published weights support self-hosted Diffusers workflows.

Updated 9 days ago
imageopen-weights

Recommended

Top current models by information quality score — good defaults when you are not sure where to start.

Missing a frontier release? Add a model (editors)