AI Models Catalog
Discover and compare the latest AI models with detailed information about their capabilities, technical specifications, and real-world applications.
Language Models
Claude Fable 5.1
ProprietaryAnthropic's most capable widely released model: same $10/$50 pricing as Fable 5, cache reads cut 75%, 1M context and always-on adaptive thinking.
Claude Haiku 4.5
ProprietaryAnthropic's fastest, most cost-efficient Claude model, with near-frontier coding at one-third the cost and more than twice the speed of Claude Sonnet 4.
Claude Opus 5
ProprietaryAnthropic's Opus-tier flagship — near-Fable capability at half the price, with adaptive thinking on by default and a 1M token context window.
Claude Sonnet 5
ProprietaryAnthropic's best mix of speed and intelligence — the most agentic Sonnet yet, reaching near-Opus quality on coding and agentic work at Sonnet cost.
Command A+
Apache 2.0Cohere's first Mixture-of-Experts model, released May 2026. 218B total / 25B active parameters, vision input, reasoning, and agentic tool use under Apache 2.0.
DBRX (Discontinued)
Open Model LicenseDatabricks' open Mixture-of-Experts LLM, released March 2024 with 132B total / 36B active. Fully retired from Databricks hosting by December 19, 2025.
DeepSeek V4
MITDeepSeek's open-weight V4 family: V4-Pro (1.6T/49B), V4-Flash (284B/13B) and the experimental V4-Flash-Vision-Exp. MIT weights, 1M context.
ERNIE 5.1
ProprietaryBaidu's closed-weight flagship, released May 9, 2026. A sparse Mixture-of-Experts model with a 128K context — the first ERNIE whose weights stay unpublished.
Gemini 3.8 Flash
ProprietaryGoogle's Gemini 3.8 Flash, GA September 2, 2026. 1M-token context, $0.75/$3.75 per 1M introductory, plus a gated Flash Cyber variant.
Gemma 4
Apache 2.0Google's open-weight Gemma 4 family — Apache 2.0 models from a 2.3B-effective edge model to a 30.7B dense flagship, with QAT builds for local deployment.
GLM-5.3
MIT for GLM-5.3-Flash; GLM-5.3 weights not yet releasedZhipu AI's current line: GLM-5.3, a post-trained coding and security flagship, and GLM-5.3-Flash, a 320B/18B multimodal MoE with hybrid attention.
GPT-6 Astra
ProprietaryOpenAI's flagship since September 4, 2026. A 1,050,000-token reasoning model for long-horizon agentic work, with native computer use.
Grok 4.6
ProprietaryxAI's agent-focused flagship. 500K context at $2.00 in / $6.00 out per 1M — but a prompt of 200K tokens doubles the rate on the entire request.
Inkling
Apache 2.0Thinking Machines Lab's first open-weights model: a 975B-parameter multimodal MoE with 41B active, 1M-token context and controllable thinking effort.
Kimi K3
LicenseMoonshot AI's flagship: a 2.8T-parameter MoE with 104B active, 1M context and native vision. Open weights shipped July 27, 2026.
Ling-2.6-1T
MITAnt Group's open-weight, non-thinking flagship: ~1T parameters with 63B active, a hybrid MLA and Linear Attention stack, a 1M-token context, and an MIT license.
Llama 4
Community LicenseThe last release in Meta's Llama line: Scout and Maverick, natively multimodal Mixture-of-Experts models with 17B active and up to a 10M token context.
LongCat-2.0
MITMeituan's open-weight flagship, announced June 30, 2026. A 1.6T-parameter MoE with ~48B active parameters, a 1M-token context window, and an MIT license.
MiMo-V2.5-Pro
MITXiaomi's open-weight flagship, released April 22, 2026. A 1.02-trillion-parameter MoE with 42B active, a 1M-token context, and an MIT license.
MiniMax-M3
Community LicenseMiniMax's multimodal flagship, released June 1, 2026. A 428B-parameter MoE with ~23B active, a 1M-token context, and open weights under a community license.
Mistral Medium 3.5
Modified MITMistral AI's multimodal model for agentic and coding work: a dense 128B model with a 256K context window, released April 2026 under a modified MIT license.
Muse Glimmer
Apache 2.0Meta's Apache 2.0 open-weight 30B multimodal agent model, distilled from Muse Spark and built to run offline on a single 24 GB consumer GPU.
Muse Spark 1.3
ProprietaryMeta's frontier reasoning model, shipped September 2, 2026: a 1M-token multimodal agent model at $1.25/$4.25 per million tokens.
Nemotron 3 Ultra
OpenMDW-1.1NVIDIA's open Nemotron line: the 550B/55B-active Nemotron 3 Ultra flagship and Nemotron 3.5 Lightning, a 30B/3B-active agent-execution model, both OpenMDW.
Qwen3.8-Max
LicenseAlibaba's flagship Qwen model, released August 3, 2026. A 2.4T-parameter MoE with 95B active, 1M token context, text/image/video input — and open weights.
Seed 2.1 (Doubao)
ProprietaryByteDance's flagship Seed 2.1: Pro for deep-thinking agent work, Turbo for scale, plus the weekly-updated Doubao-Seed-Evolving. Closed weights.
Step-3.7-Flash
Apache 2.0StepFun's open-weight vision-language flagship, released May 29, 2026. A 198B MoE activating ~11B per token, served over 400 tokens/second under Apache 2.0.
Tencent Hy4 preview
Apache 2.0Tencent's open-weight flagship, released August 28, 2026. A 770B MoE with 49B active, a 1M-token context, Apache 2.0 — and explicitly a preview.
Code Models
Composer 2.5
ProprietaryCursor's in-house agentic coding model, released May 18, 2026. Built by post-training Moonshot's Kimi K2.5, delivering frontier-adjacent coding at low cost.
Doubao Seed 2.0 Code
ProprietaryByteDance's dedicated coding model, released February 14, 2026 with the Seed 2.0 family. Combines programming and reasoning with multimodal perception.
Video Generation Models
HappyHorse 1.1
ProprietaryAlibaba's video generation model, released June 23, 2026 by the Alibaba Token Hub team. Generates up to 15 seconds of 1080p video with synchronized audio.
Kling 3.0
ProprietaryKuaishou's closed-weight video family, launched February 5, 2026. Four SKUs across video and image, clips of 3–15 seconds with natively generated audio.
MiniMax H3 (Hailuo 3.0)
H3 Community LicenseMiniMax's omni-modal video model, announced July 31, 2026: 33B open weights, 2K video up to 15 seconds, with stereo audio generated jointly.
Seedance 2.5
ProprietaryByteDance Seed video model, launched July 31, 2026: 30-second audio-video in one pass and 50 reference inputs. No technical report or rate card yet.
Sora 2 (Discontinued)
ProprietaryOpenAI is retiring Sora: the apps closed April 26, 2026 and the API shuts down September 24, 2026, with no OpenAI successor.
Veo 3.1
ProprietaryGoogle DeepMind's video model with native audio. Three tiers from $0.05 to $0.60 per second, 4–8 second clips up to 4K, plus video extension.
Vidu Q3 Pro
ProprietaryShengShu's flagship Vidu Q3 Pro video model, announced January 30, 2026. Generates up to 16 seconds of 1080p video with native audio in a single pass.
Wan 3.0
Proprietary (Wan 3.0); Apache 2.0Alibaba's video family. Wan 3.0 (August 2026) generates 30-second 1080p clips with audio, API-only. Wan 2.2 is still the last Apache 2.0 release.
Image Generation Models
HiDream-O1-Image
MITHiDream.ai's 8B pixel-level unified transformer, open-sourced under MIT on May 8, 2026. No VAE, no separate text encoder, and generates up to 2,048x2,048.
HunyuanImage 3.0
Hunyuan Community LicenseTencent's 80B-parameter MoE image model, released September 28, 2025. Natively multimodal and autoregressive, under a restricted community license.
Nano Banana
ProprietaryGoogle's Gemini image family — Pro, 2, 2 Lite and the legacy original. Prices run $0.0336 to $0.24 per image, and the cheapest one outranks the flagship.
Qwen-Image 3.0
ProprietaryAlibaba's third-generation image model, announced July 21, 2026: 4,500-token prompts, 10px text, 12 languages — and a second flagship with no weights.
Seedream 5.0
ProprietaryByteDance's unified multimodal image model. Seedream 5.0 Lite (Feb 13, 2026) added deep thinking and web search grounding; a Pro tier followed in June 2026.
Stable Diffusion 3.5
Community LicenseStability AI's flagship open-weights image model, released October 2024. Three MMDiT variants, still billed as its most powerful image model in September 2026.
Z-Image
Apache 2.0Alibaba Tongyi-MAI's 6B open-source image model. Z-Image-Turbo, released November 26, 2025, generates 8-step images inside 16 GB of VRAM under Apache 2.0.
Can't Find a Model?
We're constantly expanding our models catalog. Let us know what AI models you'd like us to add!