GenAIWiki

Structured cards

Model database

Filter by provider, architecture family, or full-text search across descriptions.

Frontier models

Verified flagships and recent launches across major providers—scroll sideways for the full shelf.

SpaceXAI

Grok 4.6

FrontierLatest

Grok 4.6 is SpaceXAI's frontier model for coding, long-running agents, knowledge work, and interactive visual projects. Official documentation lists text and image input, text output, a 500,000-token context window, configurable low through xhigh reasoning, function calling, web and X search, code execution, and the API model ID grok-4.6.

FeaturedUpdated 9 days ago
xaispacexai

Google

Gemini 3.7 Flash

FrontierLatest

Gemini 3.7 Flash is Google's workhorse model for coding, agents, web development, knowledge work, and multimodal workflows. Google documents availability through the Gemini API, Google AI Studio, Android Studio, Gemini Enterprise Agent Platform, and Gemini Spark, with a 1M-token input context and up to 64K output.

FeaturedUpdated 9 days ago
googlegemini

Meta

Muse Spark 1.2

FrontierLatest

Muse Spark 1.2 is Meta's hosted coding and agent model for code generation, complex debugging, codebase understanding, long-horizon workflows, and tool use. It is available through Muse Code and Meta Model API with a 1M-token context window and separate standard and contributor data-use tiers.

FeaturedUpdated 9 days ago
metamuse

Alibaba Qwen

Qwen3.8-Max

FrontierLatest

Qwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents. QwenCloud documents text, image, and video input, text output, a 1M-token context, up to 131K output, built-in tools, and the API model ID qwen3.8-max.

FeaturedUpdated 9 days ago
alibabaqwen

Z.ai

GLM-5.3

FrontierLatest

GLM-5.3 is Z.ai's post-trained coding and agent model built on the same base model as GLM-5.2. Z.ai positions it for complex coding, long-horizon tasks, and cybersecurity evaluation, with mandatory thinking, low, high, and max reasoning effort, and availability through GLM Coding Plan and ZCode at launch.

FeaturedUpdated 9 days ago
zaiglm

OpenAI

GPT-5.6 Sol

FrontierLatest

GPT-5.6 Sol is OpenAI's strongest GPT-5.6-class general-purpose model for coding, research, and defensive cybersecurity workflows. In the August 10, 2026 Daybreak expansion, Sol is the recommended starting point for most vetted defenders under Daybreak Blue, where system-level cyber guardrails are adjusted for authorized security work without switching to a purpose-trained cyber model.

FeaturedUpdated 9 days ago
openaigpt-5-6

Moonshot AI

Kimi K3

FrontierLatest

Kimi K3 is Moonshot AI's July 2026 multimodal model for long-horizon coding, reasoning, and knowledge work. Official Kimi materials document native vision, a one-million-token context window, thinking-only API behavior, and access through Kimi, Kimi Work, Kimi Code, and the Kimi API.

FeaturedUpdated 9 days ago
frontierchina

Meta

Muse Glimmer 30B

FrontierLatest

Muse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.

FeaturedUpdated 9 days ago
metamuse

Sarvam AI

Sarvam 105B

FrontierLatest

Sarvam 105B is Sarvam AI's flagship 105B+ parameter Mixture-of-Experts reasoning model for Indian-language and English chat, complex reasoning, coding, long-context document analysis, and agentic tool-use workflows. Sarvam documents it as a 128K-context OpenAI-compatible chat model with Multi-head Latent Attention, 12T tokens of pre-training data, Apache 2.0 open weights, and production use powering Indus. Its strongest fit is Indian-language enterprise assistants, multilingual reasoning, and agent workflows where native script, romanized, and code-mixed inputs matter.

Updated 9 days ago
sarvamindian languages

Anthropic

Claude Fable 5

FrontierLatest

Anthropic's highest-capability widely released Claude model, documented for deep reasoning, codebase-scale work, long-context enterprise workloads, and multimodal inputs.

FeaturedUpdated 9 days ago
frontierclaude

Alibaba Qwen

Qwen3.8-27B

FrontierLatest

Qwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.

FeaturedUpdated 9 days ago
alibabaqwen

OpenAI

GPT-5.6-Cyber

FrontierLatest

GPT-5.6-Cyber is OpenAI's cybersecurity-specific model announced August 10, 2026 for Daybreak Red. Built on GPT-5.6 Sol, it is trained to improve specialized cyber tasks for trusted, authorized defenders—such as vulnerability research and exploit-chain development in approved environments—and to reduce refusals on certain higher-risk dual-use security prompts that general Sol still blocks.

FeaturedUpdated 9 days ago
openaigpt-5-6

All models

Filter and paginate the full catalog. Tabs control lifecycle scope.

DeepSeek

DeepSeek-V4-Pro

CurrentLatest

DeepSeek-V4-Pro is DeepSeek's V4 model for high-capability reasoning and agentic coding, available through DeepSeek's OpenAI-compatible and Anthropic-compatible API surfaces.

FeaturedUpdated 9 days ago
deepseekreasoning

OpenAI

GPT-5.6 Luna

CurrentLatest

OpenAI's GPT-5.6 Luna is the cost-sensitive GPT-5.6 tier for high-volume workloads that still need current GPT-5.6 behavior, vision input, structured outputs, and tool support.

Updated 9 days ago
cost-efficienthigh-volume

Meta

Llama 3.1 405B Instruct

CurrentLatest

Meta’s largest open-weights instruct checkpoint in the Llama 3.1 family, aimed at strong reasoning and coding quality with a permissive license for research and customization. It is typically served on dedicated GPU clusters or via partners (cloud inference, on-prem) rather than a single vendor API.

FeaturedUpdated 9 days ago
open-weightsself-host

DeepSeek

DeepSeek-R1

CurrentLatest

DeepSeek-R1 is a reasoning-focused model family emphasizing chain-of-thought style behavior for math, code, and structured problem solving. Deployment options include API and open-weight variants—verify licensing and hosting constraints for your region.

FeaturedUpdated 9 days ago
reasoningresearch

Anysphere

Cursor Composer 2.5

CurrentLatest

Cursor Composer 2.5 is Anysphere's price-efficient first-party coding model for long-running agentic tasks inside Cursor, with improved sustained work and instruction following over Composer 2.

Updated 9 days ago
codingcursor

Cohere

Command R+

CurrentLatest

Cohere’s enterprise-oriented Command R+ emphasizes retrieval-grounded answers and tool orchestration patterns for business data. It targets teams building RAG-heavy assistants where citation-style behavior and connector patterns matter more than raw chat novelty.

FeaturedUpdated 9 days ago
enterpriserag

Google

Gemini 3.1 Flash-Lite

CurrentLatest

Gemini 3.1 Flash-Lite is Google's stable Gemini 3-series workhorse model for cost-efficient, high-volume multimodal workloads.

Updated 9 days ago
googleflash-lite

Anthropic

Claude Haiku 4.5

CurrentLatest

Claude Haiku 4.5 is Anthropic's fastest current Claude model with near-frontier intelligence for high-volume and latency-sensitive workloads.

Updated 9 days ago
haikulatency

DeepSeek

DeepSeek-V4-Flash

CurrentLatest

DeepSeek-V4-Flash is DeepSeek's faster and more economical V4 model, supporting thinking and non-thinking modes through the current DeepSeek API.

Updated 9 days ago
deepseekflash

Microsoft AI

MAI-Code-1-Flash

CurrentLatest

Microsoft AI's agentic coding model in the MAI family, announced for fast code editing, debugging, and tool-driven developer workflows.

FeaturedUpdated 9 days ago
codingmicrosoft

Mistral AI

Mistral Large 3

CurrentLatest

Mistral's open-weight general-purpose multimodal model listed in official Mistral model documentation.

Updated 9 days ago
frontiermistral

Microsoft AI

MAI-Image-2.5

CurrentLatest

Microsoft AI's MAI image model for generation, editing, and visual content workflows, announced as part of the June 2026 MAI model release.

Updated 9 days ago
imagemicrosoft

Microsoft AI

MAI-Voice-2

CurrentLatest

Microsoft AI's voice generation model in the MAI family, announced for natural text-to-speech and voice experiences.

Updated 9 days ago
speechvoice

Microsoft AI

MAI-Transcribe-1.5

CurrentLatest

Microsoft AI's speech-to-text model in the MAI family, announced for fast, accurate transcription across product surfaces.

Updated 9 days ago
speechtranscription

Microsoft AI

MAI-Image-2.5-Flash

CurrentLatest

Microsoft AI's faster MAI image variant, announced for lower-latency image generation and editing workflows.

Updated 9 days ago
imagemicrosoft

Mistral AI

Mistral Small 3

CurrentLatest

Mistral Small 3 is Mistral’s efficiency tier for fast, affordable chat and tool use at high QPS—positioned between tiny open models and Mistral Large. Exact naming and versioning appear in Mistral’s API catalog; pin versions in production.

Updated 9 days ago
latencyapi

Meta

Llama 3.1 70B Instruct

CurrentLatest

Llama 3.1 70B Instruct is a mid-size open-weights instruct model balancing quality and deployability on a single large GPU or small multi-GPU nodes. Common for private assistants, on-prem pilots, and fine-tunes where 405B is impractical.

Updated 9 days ago
open-weightsself-host

OpenAI

Whisper large-v3

CurrentLatest

Whisper large-v3 is OpenAI’s ASR model for transcription and translation across many languages, with strong robustness to accents and noise. It is commonly self-hosted or used via API partners; latency depends heavily on hardware and chunking strategy.

Updated 9 days ago
audioopen-weights

Snowflake

Snowflake Arctic

CurrentLatest

Snowflake Arctic is an enterprise-oriented open model emphasizing efficient training recipes and SQL-adjacent enterprise tasks inside the Snowflake ecosystem. It targets teams that want LLM features colocated with governed data in Snowflake Cortex.

Updated 9 days ago

NVIDIA

NVIDIA Nemotron-4 340B

CurrentLatest

NVIDIA Nemotron-4 340B is a large open-weights model suite aimed at enterprise and research users who train and serve on NVIDIA stacks (NeMo, NGC). It targets GPU-native teams that need customizable checkpoints with NVIDIA-optimized tooling.

Updated 9 days ago
open-weightsenterprise

Mistral AI

Mistral 7B Instruct v0.3

CurrentLatest

Mistral 7B Instruct is a compact dense model that popularized efficient open-weight chat quality at small scale. It remains a baseline for fine-tunes and on-prem pilots where 13B+ models are too heavy.

Updated 9 days ago
open-weightsslm

OpenAI

o3-mini

CurrentLatest

Compact reasoning-focused model in OpenAI’s o-series line aimed at strong STEM and coding performance with lower cost than full o3. Intended for developers who want reasoning without always paying flagship prices—confirm exact API availability and snapshot names in OpenAI docs.

Updated 9 days ago
reasoningstem

OpenAI

text-embedding-3-large

CurrentLatest

text-embedding-3-large produces high-dimensional text embeddings for semantic search, clustering, and classification. Teams pair it with pgvector or SaaS vector DBs for RAG; output dimensions can be reduced with tradeoffs described in OpenAI documentation.

Updated 9 days ago
embeddingsretrieval

Microsoft

Phi-4

CurrentLatest

Phi-4 is Microsoft Research’s small language model line focused on strong reasoning per parameter for on-device and low-cost cloud scenarios. Deployment often happens via Azure AI or Hugging Face hubs—confirm license for your channel.

Updated 9 days ago
slmmicrosoft

Recommended

Top current models by information quality score — good defaults when you are not sure where to start.

Missing a frontier release? Add a model (editors)