Structured cards
Model database
Filter by provider, architecture family, or full-text search across descriptions.
Frontier models
Verified flagships and recent launches across major providers—scroll sideways for the full shelf.
All models
Filter and paginate the full catalog. Tabs control lifecycle scope.
Alibaba Qwen
Qwen3.8-27B
CurrentLatestQwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.
Alibaba Qwen
Qwen3.8-Max
CurrentLatestQwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents. QwenCloud documents text, image, and video input, text output, a 1M-token context, up to 131K output, built-in tools, and the API model ID qwen3.8-max.
Meta
Muse Spark 1.2
CurrentLatestMuse Spark 1.2 is Meta's hosted coding and agent model for code generation, complex debugging, codebase understanding, long-horizon workflows, and tool use. It is available through Muse Code and Meta Model API with a 1M-token context window and separate standard and contributor data-use tiers.
Gemini 3.7 Flash
CurrentLatestGemini 3.7 Flash is Google's workhorse model for coding, agents, web development, knowledge work, and multimodal workflows. Google documents availability through the Gemini API, Google AI Studio, Android Studio, Gemini Enterprise Agent Platform, and Gemini Spark, with a 1M-token input context and up to 64K output.
SpaceXAI
Grok 4.6
CurrentLatestGrok 4.6 is SpaceXAI's frontier model for coding, long-running agents, knowledge work, and interactive visual projects. Official documentation lists text and image input, text output, a 500,000-token context window, configurable low through xhigh reasoning, function calling, web and X search, code execution, and the API model ID grok-4.6.
Meta
Muse Glimmer 30B
CurrentLatestMuse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.
Moonshot AI
Kimi K3
CurrentLatestKimi K3 is Moonshot AI's July 2026 multimodal model for long-horizon coding, reasoning, and knowledge work. Official Kimi materials document native vision, a one-million-token context window, thinking-only API behavior, and access through Kimi, Kimi Work, Kimi Code, and the Kimi API.
Sarvam AI
Sarvam 105B
CurrentLatestSarvam 105B is Sarvam AI's flagship 105B+ parameter Mixture-of-Experts reasoning model for Indian-language and English chat, complex reasoning, coding, long-context document analysis, and agentic tool-use workflows. Sarvam documents it as a 128K-context OpenAI-compatible chat model with Multi-head Latent Attention, 12T tokens of pre-training data, Apache 2.0 open weights, and production use powering Indus. Its strongest fit is Indian-language enterprise assistants, multilingual reasoning, and agent workflows where native script, romanized, and code-mixed inputs matter.
OpenAI
GPT-5.6-Cyber
CurrentLatestGPT-5.6-Cyber is OpenAI's cybersecurity-specific model announced August 10, 2026 for Daybreak Red. Built on GPT-5.6 Sol, it is trained to improve specialized cyber tasks for trusted, authorized defenders—such as vulnerability research and exploit-chain development in approved environments—and to reduce refusals on certain higher-risk dual-use security prompts that general Sol still blocks.
Sarvam AI
Sarvam 30B
CurrentLatestSarvam 30B is a 30B parameter Mixture-of-Experts chat and reasoning model from Sarvam AI, optimized for Indian languages, real-time conversation, high-throughput voice-agent pipelines, coding, and practical deployment. Sarvam documents 2.4B active parameters per token, 16T tokens of pre-training data, a 64K context window, Grouped Query Attention, Apache 2.0 open weights, and OpenAI-compatible chat completions.
MiniMax
MiniMax M3
CurrentLatestMiniMax M3 is a June 2026 open-weight multimodal model for coding, agentic workflows, computer use, and long-context work. MiniMax documents a one-million-token context window, native image and video understanding, and deployment through hosted or downloadable model paths.
Alibaba Qwen
Qwen3.8-2.4T-A95B
CurrentLatestQwen3.8-2.4T-A95B is Qwen's downloadable 2.4-trillion-parameter sparse mixture-of-experts model with 95B activated parameters. The official model card documents text-only input and output, mandatory thinking, 262,144 native context extensible to roughly 1.01M, configurable reasoning effort, and a dedicated Qwen3.8-Max license.
Microsoft Research
VibeVoice-ASR-BitNet
CurrentLatestVibeVoice-ASR-BitNet is Microsoft Research's compressed automatic speech recognition model for real-time CPU and edge transcription without a GPU. Its official model card documents a 1.58 GB footprint, seven-language support, and real-time CPU performance under the tested configuration.
OpenAI
GPT-5.6 Sol
CurrentLatestGPT-5.6 Sol is OpenAI's strongest GPT-5.6-class general-purpose model for coding, research, and defensive cybersecurity workflows. In the August 10, 2026 Daybreak expansion, Sol is the recommended starting point for most vetted defenders under Daybreak Blue, where system-level cyber guardrails are adjusted for authorized security work without switching to a purpose-trained cyber model.
Alibaba Qwen
Qwen-Image-3.0
CurrentLatestQwen-Image-3.0 is Alibaba's third-generation image model for image generation, editing, complex layouts, and multilingual text rendering. The official Qwen release highlights inputs up to 4.5K, small text rendering down to 10 pixels, and native support for twelve languages.
Z.ai
GLM-5.3
CurrentLatestGLM-5.3 is Z.ai's post-trained coding and agent model built on the same base model as GLM-5.2. Z.ai positions it for complex coding, long-horizon tasks, and cybersecurity evaluation, with mandatory thinking, low, high, and max reasoning effort, and availability through GLM Coding Plan and ZCode at launch.
Microsoft Research
Mage-Flow
CurrentLatestMage-Flow is Microsoft Research's MIT-licensed 4B image generation and editing family. The official model card provides base, reinforcement-learning-aligned, turbo, and editing variants for teams evaluating downloadable image models and custom serving.
Anthropic
Claude Sonnet 5
CurrentLatestClaude Sonnet 5 is Anthropic's balanced current Sonnet model for agentic coding, tool use, and production knowledge work, offering a stronger cost-performance lane than prior Sonnet releases.
Anthropic
Claude Opus 5
CurrentLatestClaude Opus 5 is Anthropic's advanced Claude model for complex agentic coding and enterprise work, with a 1M-token context window, adaptive thinking, and strong long-horizon task behavior.
Anthropic
Claude Fable 5
CurrentLatestAnthropic's highest-capability widely released Claude model, documented for deep reasoning, codebase-scale work, long-context enterprise workloads, and multimodal inputs.
OpenAI
GPT-5.6 Terra
CurrentLatestOpenAI's GPT-5.6 Terra is the balanced GPT-5.6 tier for teams that want strong reasoning and coding quality at a lower cost than the Sol flagship lane.
Mistral AI
Mistral Medium 3.5
CurrentLatestMistral Medium 3.5 is Mistral's frontier-class multimodal model optimized for agentic and coding use cases, released as open weights under a modified MIT license.
AWS
Amazon Nova
CurrentLatestAmazon Nova is AWS’s multimodal foundation model family for text, image, and video workloads delivered through Amazon Bedrock with enterprise IAM, VPC, and governance patterns. Model IDs and modalities vary by region—verify Bedrock model access lists.
Stability AI
Stable Diffusion XL
CurrentLatestStable Diffusion XL (SDXL) 1.0 is Stability AI's latent diffusion text-to-image model for native 1024x1024 generation. The base model can run standalone or feed an optional refiner for the final denoising steps, and the published weights support self-hosted Diffusers workflows.
Recommended
Top current models by information quality score — good defaults when you are not sure where to start.
Sarvam AI
Sarvam 30B
CurrentLatestSarvam 30B is a 30B parameter Mixture-of-Experts chat and reasoning model from Sarvam AI, optimized for Indian languages, real-time conversation, high-throughput voice-agent pipelines, coding, and practical deployment. Sarvam documents 2.4B active parameters per token, 16T tokens of pre-training data, a 64K context window, Grouped Query Attention, Apache 2.0 open weights, and OpenAI-compatible chat completions.
MiniMax
MiniMax M3
CurrentLatestMiniMax M3 is a June 2026 open-weight multimodal model for coding, agentic workflows, computer use, and long-context work. MiniMax documents a one-million-token context window, native image and video understanding, and deployment through hosted or downloadable model paths.
Alibaba Qwen
Qwen3.8-2.4T-A95B
CurrentLatestQwen3.8-2.4T-A95B is Qwen's downloadable 2.4-trillion-parameter sparse mixture-of-experts model with 95B activated parameters. The official model card documents text-only input and output, mandatory thinking, 262,144 native context extensible to roughly 1.01M, configurable reasoning effort, and a dedicated Qwen3.8-Max license.
Missing a frontier release? Add a model (editors)