GenAIWiki

Structured cards

Model database

Filter by provider, architecture family, or full-text search across descriptions.

Frontier models

Verified flagships and recent launches across major providers—scroll sideways for the full shelf.

SpaceXAI

Grok 4.6

FrontierLatest

Grok 4.6 is SpaceXAI's frontier model for coding, long-running agents, knowledge work, and interactive visual projects. Official documentation lists text and image input, text output, a 500,000-token context window, configurable low through xhigh reasoning, function calling, web and X search, code execution, and the API model ID grok-4.6.

FeaturedUpdated 9 days ago
xaispacexai

Google

Gemini 3.7 Flash

FrontierLatest

Gemini 3.7 Flash is Google's workhorse model for coding, agents, web development, knowledge work, and multimodal workflows. Google documents availability through the Gemini API, Google AI Studio, Android Studio, Gemini Enterprise Agent Platform, and Gemini Spark, with a 1M-token input context and up to 64K output.

FeaturedUpdated 9 days ago
googlegemini

Meta

Muse Spark 1.2

FrontierLatest

Muse Spark 1.2 is Meta's hosted coding and agent model for code generation, complex debugging, codebase understanding, long-horizon workflows, and tool use. It is available through Muse Code and Meta Model API with a 1M-token context window and separate standard and contributor data-use tiers.

FeaturedUpdated 9 days ago
metamuse

Alibaba Qwen

Qwen3.8-Max

FrontierLatest

Qwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents. QwenCloud documents text, image, and video input, text output, a 1M-token context, up to 131K output, built-in tools, and the API model ID qwen3.8-max.

FeaturedUpdated 9 days ago
alibabaqwen

Z.ai

GLM-5.3

FrontierLatest

GLM-5.3 is Z.ai's post-trained coding and agent model built on the same base model as GLM-5.2. Z.ai positions it for complex coding, long-horizon tasks, and cybersecurity evaluation, with mandatory thinking, low, high, and max reasoning effort, and availability through GLM Coding Plan and ZCode at launch.

FeaturedUpdated 9 days ago
zaiglm

OpenAI

GPT-5.6 Sol

FrontierLatest

GPT-5.6 Sol is OpenAI's strongest GPT-5.6-class general-purpose model for coding, research, and defensive cybersecurity workflows. In the August 10, 2026 Daybreak expansion, Sol is the recommended starting point for most vetted defenders under Daybreak Blue, where system-level cyber guardrails are adjusted for authorized security work without switching to a purpose-trained cyber model.

FeaturedUpdated 9 days ago
openaigpt-5-6

Moonshot AI

Kimi K3

FrontierLatest

Kimi K3 is Moonshot AI's July 2026 multimodal model for long-horizon coding, reasoning, and knowledge work. Official Kimi materials document native vision, a one-million-token context window, thinking-only API behavior, and access through Kimi, Kimi Work, Kimi Code, and the Kimi API.

FeaturedUpdated 9 days ago
frontierchina

Meta

Muse Glimmer 30B

FrontierLatest

Muse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.

FeaturedUpdated 9 days ago
metamuse

Sarvam AI

Sarvam 105B

FrontierLatest

Sarvam 105B is Sarvam AI's flagship 105B+ parameter Mixture-of-Experts reasoning model for Indian-language and English chat, complex reasoning, coding, long-context document analysis, and agentic tool-use workflows. Sarvam documents it as a 128K-context OpenAI-compatible chat model with Multi-head Latent Attention, 12T tokens of pre-training data, Apache 2.0 open weights, and production use powering Indus. Its strongest fit is Indian-language enterprise assistants, multilingual reasoning, and agent workflows where native script, romanized, and code-mixed inputs matter.

Updated 9 days ago
sarvamindian languages

Anthropic

Claude Fable 5

FrontierLatest

Anthropic's highest-capability widely released Claude model, documented for deep reasoning, codebase-scale work, long-context enterprise workloads, and multimodal inputs.

FeaturedUpdated 9 days ago
frontierclaude

Alibaba Qwen

Qwen3.8-27B

FrontierLatest

Qwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.

FeaturedUpdated 9 days ago
alibabaqwen

OpenAI

GPT-5.6-Cyber

FrontierLatest

GPT-5.6-Cyber is OpenAI's cybersecurity-specific model announced August 10, 2026 for Daybreak Red. Built on GPT-5.6 Sol, it is trained to improve specialized cyber tasks for trusted, authorized defenders—such as vulnerability research and exploit-chain development in approved environments—and to reduce refusals on certain higher-risk dual-use security prompts that general Sol still blocks.

FeaturedUpdated 9 days ago
openaigpt-5-6

All models

Filter and paginate the full catalog. Tabs control lifecycle scope.

Google

Gemini 1.5 Pro

Legacy

Google DeepMind Gemini 1.5 Pro targets long-context multimodal workloads—large effective context for retrieval-heavy document pipelines, plus image, audio, and video inputs on supported surfaces. It is often paired with Vertex AI or the Gemini API for enterprise workloads on GCP.

FeaturedUpdated 9 days ago
googlelong-context

DeepSeek

DeepSeek-V3

Legacy

DeepSeek-V3 is a large-scale language model family noted for strong coding and math performance under open or research-friendly terms (verify the exact license for your deployment). Teams adopt it for cost-sensitive research, self-hosted inference, or comparison against frontier APIs.

FeaturedUpdated 9 days ago
researchcoding

Anthropic

Claude 3.5 Sonnet

Legacy

Anthropic’s balanced Sonnet-tier model tuned for long-context reasoning, careful instruction following, and strong performance on coding and analysis workloads. It is a common enterprise choice on the Anthropic API and on AWS Bedrock when teams need large context for RAG and document review.

FeaturedUpdated 9 days ago
codingagents

Mistral AI

Mistral Large 2

Legacy

Mistral’s frontier-class multilingual model emphasizing JSON adherence, agent-friendly behavior, and competitive reasoning within the Mistral API ecosystem. European teams often evaluate it for GDPR-adjacent deployment patterns alongside US-hosted alternatives.

FeaturedUpdated 9 days ago
euenterprise

OpenAI

GPT-5.4

Legacy

OpenAI's GPT-5.4 model, documented in the official OpenAI API model guide as part of the current GPT-5 family below the GPT-5.5 flagship lane.

FeaturedUpdated 9 days ago
frontieropenai

OpenAI

GPT-4o

Legacy

GPT-4o is an OpenAI multimodal model that accepts text and image inputs and produces text. It supports streaming, function calling, Structured Outputs, fine-tuning, and predicted outputs for vision-heavy assistants and structured extraction workflows.

FeaturedUpdated 9 days ago
frontiermultimodal

xAI

Grok-2

Legacy

Grok-2 is xAI’s flagship chat model positioned for real-time knowledge integrations and high-throughput conversational products on xAI’s API. Availability and pricing evolve—treat capabilities as vendor-specific.

FeaturedUpdated 9 days ago
frontierapi

Alibaba

Qwen 2.5 72B Instruct

Legacy

Qwen 2.5 72B Instruct is a large multilingual open-weights model from Alibaba’s Qwen family with strong coding and general chat performance. Common in APAC deployments and on Hugging Face inference endpoints—check license terms for commercial use.

FeaturedUpdated 9 days ago
open-weightsmultilingual

Anthropic

Claude Sonnet 4.6

Legacy

Anthropic's Sonnet-tier model documented as the best combination of speed and intelligence in the Claude model overview.

FeaturedUpdated 9 days ago
frontiersonnet

OpenAI

o1

Legacy

OpenAI’s o1 series emphasizes extended internal reasoning before answering—useful for competition-style math, complex debugging, and multi-step planning where latency is acceptable. It behaves differently from standard chat models: tune prompts for chain-of-thought style tasks and measure time-to-first-token.

FeaturedUpdated 9 days ago
reasoningstem

Google

Gemini 2.5 Flash-Lite

Legacy

Google's fastest and most budget-friendly multimodal model in the Gemini 2.5 family, according to the Gemini API model documentation.

Updated 9 days ago
googleflash-lite

DeepSeek

DeepSeek-V3.2

Legacy

DeepSeek's documented successor to the V3.2 experimental line, positioned in official DeepSeek API news as live on app, web, and API.

FeaturedUpdated 9 days ago
frontierdeepseek

OpenAI

GPT-5.4 mini

Legacy

OpenAI's smaller GPT-5.4 mini model, documented in the official OpenAI API model guide for lower-latency or lower-cost GPT-5 family workloads.

Updated 9 days ago
openaimini

OpenAI

GPT-5.4 nano

Legacy

OpenAI's smallest GPT-5.4 nano model, documented in the official OpenAI API model guide for very low-latency or economical GPT-5 family routing.

Updated 9 days ago
openainano

Google

Gemini 1.5 Flash

Legacy

Gemini 1.5 Flash targets low-latency, cost-efficient multimodal chat and retrieval workloads on the Gemini API and Vertex AI. It keeps much of the long-context family behavior with faster responses for interactive apps.

Updated 9 days ago
googlelatency

Anthropic

Claude 3.5 Haiku

Legacy

Claude 3.5 Haiku is Anthropic’s fast, cost-efficient tier for high-volume classification, routing, and simple chat. It targets latency-sensitive paths and agent pre-processing before escalating to Sonnet-class models.

Updated 9 days ago
latencycost

Microsoft

Phi-3 Medium

Legacy

Phi-3 Medium is a compact instruct model aimed at strong quality per parameter for on-device and cost-sensitive cloud inference. It competes with other SLMs on coding and reasoning benchmarks—validate on your domain prompts.

Updated 9 days ago
edgeopen-weights

Meta

Llama 3.2 1B Instruct

Legacy

Llama 3.2 1B Instruct is among the smallest Llama instruct checkpoints for extreme latency and footprint constraints. Use for routing, tagging, and toy assistants—not for complex reasoning without retrieval augmentation.

Updated 9 days ago
edgetiny

Mistral AI

Mixtral 8x7B Instruct

Legacy

Mixtral 8x7B Instruct is a sparse mixture-of-experts open model noted for strong quality per active parameter and efficient inference vs dense models of similar capability. Widely hosted on inference clouds and self-hosted stacks.

Updated 9 days ago
open-weightsmoe

OpenAI

GPT-3.5 Turbo

Legacy

GPT-3.5 Turbo is a long-standing cost-efficient chat model family on the OpenAI API for simple assistants, classification, and legacy integrations. Many teams still use it for non-critical paths or as a fallback when newer models are rate-limited.

Updated 9 days ago
legacycost

Anthropic

Claude 3 Opus

Legacy

Claude 3 Opus was Anthropic’s highest-capability Claude 3-era model for difficult reasoning, nuanced writing, and complex analysis before later Sonnet generations. Teams still reference it for historical benchmarks and legacy deployments—verify current availability in API and Bedrock model lists.

Updated 9 days ago
frontierwriting

Anthropic

Claude 3 Sonnet

Legacy

Claude 3 Sonnet balanced cost and capability in the Claude 3 generation—useful for general assistants and document workflows where Opus was unnecessary. New deployments should compare against Claude 3.5 Sonnet for pricing and quality.

Updated 9 days ago
general-purposelegacy

Meta

Llama 3.2 3B Instruct

Legacy

Llama 3.2 3B Instruct is a compact instruct model in Meta’s 3.2 generation aimed at mobile and edge scenarios with multilingual support on supported checkpoints. Verify hardware targets and license terms for your distribution channel.

Updated 9 days ago
edgeslm

xAI

Grok-3

Legacy

Grok-3 represents xAI’s newer generation aimed at stronger reasoning and tool use versus Grok-2. Capabilities and rollout are version-specific—validate against xAI documentation for your account tier.

Updated 9 days ago
frontierapi

Recommended

Top current models by information quality score — good defaults when you are not sure where to start.

Missing a frontier release? Add a model (editors)