Structured cards
Model database
Filter by provider, architecture family, or full-text search across descriptions.
Frontier models
Verified flagships and this month’s launches—Claude Opus 5.5, GPT-6 Sol and Luna, GPT-6 Astra, Claude Fable 5.1, Gemini 3.8, and more. Scroll sideways for the full shelf.
By company
Latest verified models from each frontier company, summarized on one page.
All models
Filter and paginate the full catalog. Tabs control lifecycle scope.
Meta
Muse Spark 1.3
CurrentLatestMuse Spark 1.3 is Meta's September 2, 2026 hosted model for long-horizon agentic and coding work. The API ID is muse-spark-1.3, with a 1,048,576-token context window, text/image/video/PDF input, and text output. Meta documents it in Muse Code and Meta Model API, with reasoning effort including max on the Standard tier. Official docs still list muse-spark-1.2 as the Muse Code default while recommending 1.3 for new API work.
Meta
Muse Glimmer 30B
CurrentLatestMuse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.
Meta
Muse Voice Transcribe
CurrentLatestMuse Voice Transcribe is Meta Superintelligence Labs' September 1, 2026 real-time audio perception model. The API ID is muse-voice-transcribe-1.0. It does streaming ASR with adaptive delay, speaker diarization for 20+ speakers, and native endpointing. Access is via Meta Model API (WebSocket realtime and file POST /v1/asr/transcribe), Meta AI for Mac, and Muse Code dictation.
Meta
Muse Image 1.0
CurrentLatestMuse Image 1.0 is Meta's hosted image generation and editing model on Meta Model API. The API ID is muse-image-1.0. One model handles text-to-image, image editing, and multi-image composition. Meta documents an agentic loop that can search the web for visual references and run code to plan layouts before rendering. Reach it through the Responses API or POST /v1/images/generations and /v1/images/edits.
Meta
Llama 3.1 405B Instruct
CurrentLatestMeta’s largest open-weights instruct checkpoint in the Llama 3.1 family, aimed at strong reasoning and coding quality with a permissive license for research and customization. It is typically served on dedicated GPU clusters or via partners (cloud inference, on-prem) rather than a single vendor API.
Meta
Llama 3.1 70B Instruct
CurrentLatestLlama 3.1 70B Instruct is a mid-size open-weights instruct model balancing quality and deployability on a single large GPU or small multi-GPU nodes. Common for private assistants, on-prem pilots, and fine-tunes where 405B is impractical.
Meta
Llama 3.1 8B Instruct
CurrentLatestLlama 3.1 8B Instruct is a small open-weights model for edge laptops, single-GPU servers, and ultra-low-latency assistants. Quality per dollar is competitive for simple tasks but not for frontier reasoning.
Meta
Muse Spark 1.2
CurrentMuse Spark 1.2 is Meta's hosted coding and agent model for code generation, complex debugging, codebase understanding, long-horizon workflows, and tool use. It is available through Muse Code and Meta Model API with a 1M-token context window and separate standard and contributor data-use tiers.
Meta
Llama 3.2 1B Instruct
LegacyLlama 3.2 1B Instruct is among the smallest Llama instruct checkpoints for extreme latency and footprint constraints. Use for routing, tagging, and toy assistants—not for complex reasoning without retrieval augmentation.
Meta
Llama 3.2 3B Instruct
LegacyLlama 3.2 3B Instruct is a compact instruct model in Meta’s 3.2 generation aimed at mobile and edge scenarios with multilingual support on supported checkpoints. Verify hardware targets and license terms for your distribution channel.
Meta
Llama 3 8B
LegacyCatalog entry for this named release; see the provider’s official documentation for modalities, pricing, and context limits.
Meta
Llama 3 70B
LegacyCatalog entry for this named release; see the provider’s official documentation for modalities, pricing, and context limits.
Recommended
Top current models by information quality score — good defaults when you are not sure where to start.
Anthropic
Claude Opus 5.5
CurrentLatestClaude Opus 5.5 is Anthropic's September 22, 2026 Opus-class model for long-running agentic coding and knowledge work, and the first model in the Claude 5.5 family. The Claude API ID is claude-opus-5-5, with a 1M-token context window, 128K synchronous max output, adaptive thinking that cannot be disabled, and default effort medium. Anthropic prices it at $4 input and $20 output per million tokens and says typical workloads cost about 40% less to run than Opus 5.
OpenAI
GPT-Image-2.5 Sunburst
CurrentLatestGPT-Image-2.5 Sunburst is OpenAI's September 8, 2026 most capable image generation and editing model. The API ID is gpt-image-2.5-sunburst, with dated snapshot gpt-image-2.5-sunburst-2026-09-08. It takes text and image input and outputs images. OpenAI positions it for workflows where editing precision matters most, with quality settings low, medium, high, xhigh, max, and auto. Call it on the Images API or as the model of the Responses API image generation tool.
Alibaba Qwen
Qwen3.8-27B
CurrentLatestQwen3.8-27B is Qwen's deployment-oriented dense multimodal model for coding, professional work, research, and long-horizon agents. The official repository documents 27B parameters, native image and video understanding, flexible reasoning effort, a 262,144-token native context extensible to 1M, and compatibility with Transformers, vLLM, SGLang, and TokenSpeed.
Meta
Muse Glimmer 30B
CurrentLatestMuse Glimmer 30B is Meta Superintelligence Labs' open-weight multimodal model for local agents, coding, tool use, long-horizon reasoning, and image understanding. Meta's model card documents a dense 29.6B-parameter architecture with a dedicated perception encoder, 131,072+ context, text-and-image input, text output, controllable reasoning effort, and Apache 2.0 weights.
OpenAI
GPT-6 Sol
CurrentLatestGPT-6 Sol is OpenAI's September 22, 2026 model for complex coding and agentic workflows, positioned as a faster and cheaper GPT-6 lane than GPT-6 Astra. The API model ID is gpt-6-sol, with a 1,050,000-token context window, 128,000 max output tokens, text and image input, and reasoning.effort from none through max (default medium). Standard text pricing is $2 input and $10 output per million tokens.
Missing a frontier release? Add a model (editors)