Structured cards
Model database
Filter by provider, architecture family, or full-text search across descriptions.
Frontier models
Verified flagships and recent launches across major providers—scroll sideways for the full shelf.
All models
Filter and paginate the full catalog. Tabs control lifecycle scope.
AWS
Amazon Titan Text Premier
CurrentLatestTitan Text Premier is AWS’s managed text model for Bedrock workloads emphasizing integration with guardrails, knowledge bases, and private data patterns. It targets enterprise RAG and internal assistants rather than frontier creative writing.
Meta
Llama 3.1 8B Instruct
CurrentLatestLlama 3.1 8B Instruct is a small open-weights model for edge laptops, single-GPU servers, and ultra-low-latency assistants. Quality per dollar is competitive for simple tasks but not for frontier reasoning.
Gemma 2 27B
CurrentLatestGemma 2 27B is Google’s open-weights Gemma family checkpoint balancing quality and deployability for research and product teams that need permissive terms without Vertex-only APIs. It is often fine-tuned for domain tasks on TPU or GPU clusters.
Mistral AI
Mixtral 8x22B
CurrentLatestCatalog entry for this named release; see the provider’s official documentation for modalities, pricing, and context limits.
Community
LLaVA
CurrentLatestCatalog entry for this named release; see the provider’s official documentation for modalities, pricing, and context limits.
xAI
Grok 1.5
CurrentLatestCatalog entry for this named release; see the provider’s official documentation for modalities, pricing, and context limits.
DeepSeek
DeepSeek Coder V2
CurrentLatestCatalog entry for this named release; see the provider’s official documentation for modalities, pricing, and context limits.
Databricks
DBRX
CurrentLatestCatalog entry for this named release; see the provider’s official documentation for modalities, pricing, and context limits.
Gemini 2.5 Pro
CurrentGoogle's advanced Gemini model for complex tasks, with official Gemini API documentation calling out deep reasoning and coding capabilities.
xAI
Grok 4.3
CurrentxAI's documented default for general chat workloads, described in xAI docs as the most intelligent and fastest Grok model for non-specialized use cases.
Recommended
Top current models by information quality score — good defaults when you are not sure where to start.
Sarvam AI
Sarvam 30B
CurrentLatestSarvam 30B is a 30B parameter Mixture-of-Experts chat and reasoning model from Sarvam AI, optimized for Indian languages, real-time conversation, high-throughput voice-agent pipelines, coding, and practical deployment. Sarvam documents 2.4B active parameters per token, 16T tokens of pre-training data, a 64K context window, Grouped Query Attention, Apache 2.0 open weights, and OpenAI-compatible chat completions.
MiniMax
MiniMax M3
CurrentLatestMiniMax M3 is a June 2026 open-weight multimodal model for coding, agentic workflows, computer use, and long-context work. MiniMax documents a one-million-token context window, native image and video understanding, and deployment through hosted or downloadable model paths.
Alibaba Qwen
Qwen3.8-2.4T-A95B
CurrentLatestQwen3.8-2.4T-A95B is Qwen's downloadable 2.4-trillion-parameter sparse mixture-of-experts model with 95B activated parameters. The official model card documents text-only input and output, mandatory thinking, 262,144 native context extensible to roughly 1.01M, configurable reasoning effort, and a dedicated Qwen3.8-Max license.
Missing a frontier release? Add a model (editors)