---
source: 'https://howaiworks.ai/models'
section: models
count: 46
---

# Models

> All 46 pages in the Models section. Each link is the Markdown twin; drop the `.md` for the HTML page.

- [Claude Fable 5](https://howaiworks.ai/models/claude-fable.md) — updated 2026-07-24
  Anthropic's most capable widely released model, built for demanding reasoning and long-horizon agentic work, with adaptive thinking and a 1M token context.
- [Claude Haiku 4.5](https://howaiworks.ai/models/claude-haiku.md) — updated 2026-07-24
  Anthropic's fastest, most cost-efficient Claude model, with near-frontier coding at one-third the cost and more than twice the speed of Claude Sonnet 4.
- [Claude Opus 5](https://howaiworks.ai/models/claude-opus.md) — updated 2026-07-24
  Anthropic's Opus-tier flagship — near-Fable capability at half the price, with adaptive thinking on by default and a 1M token context window.
- [Claude Sonnet 5](https://howaiworks.ai/models/claude-sonnet.md) — updated 2026-07-24
  Anthropic's best mix of speed and intelligence — the most agentic Sonnet yet, reaching near-Opus quality on coding and agentic work at Sonnet cost.
- [Command A+](https://howaiworks.ai/models/command.md) — updated 2026-07-08
  Cohere's first Mixture-of-Experts model, released May 2026. 218B total / 25B active parameters, vision input, reasoning, and agentic tool use under Apache 2.0.
- [Composer 2.5](https://howaiworks.ai/models/composer.md) — updated 2026-07-24
  Cursor's in-house agentic coding model, released May 18, 2026. Built by post-training Moonshot's Kimi K2.5, delivering frontier-adjacent coding at low cost.
- [DBRX (Discontinued)](https://howaiworks.ai/models/dbrx.md) — updated 2026-07-08
  Databricks' open Mixture-of-Experts LLM, released March 2024 with 132B total / 36B active. Fully retired from Databricks hosting by December 19, 2025.
- [DeepSeek V4](https://howaiworks.ai/models/deepseek.md) — updated 2026-07-08
  DeepSeek's open-weight MoE flagship: V4-Pro (1.6T total / 49B active) and V4-Flash (284B / 13B), with hybrid sparse attention, a 1M context, and an MIT license.
- [Seed 2.1 (Doubao)](https://howaiworks.ai/models/doubao-pro.md) — updated 2026-07-08
  ByteDance's current flagship family (Seed 2.1, Doubao), released June 23, 2026. A general agent with multimodal understanding, coding, and GUI task execution.
- [Doubao Seed 2.0 Code](https://howaiworks.ai/models/doubao-seed-code.md) — updated 2026-07-08
  ByteDance's dedicated coding model, released February 14, 2026 with the Seed 2.0 family. Combines programming and reasoning with multimodal perception.
- [Seed1.5-VL](https://howaiworks.ai/models/doubao-vision.md) — updated 2026-07-08
  ByteDance Seed's standalone vision-language model: a 532M vision encoder with a 20B active-parameter MoE, top results on 38 of 60 public VLM benchmarks.
- [ERNIE 5.1](https://howaiworks.ai/models/ernie.md) — updated 2026-07-08
  Baidu's closed-weight flagship, released May 9, 2026. A sparse Mixture-of-Experts model with a 128K context — the first ERNIE whose weights stay unpublished.
- [Gemini 3.5](https://howaiworks.ai/models/gemini.md) — updated 2026-09-02
  Google's Gemini 3.5 generation. Gemini 3.8 Flash, shipped September 2, 2026, is the flagship at $0.75/$3.75 per 1M introductory, with a gated Cyber variant.
- [Gemma 4](https://howaiworks.ai/models/gemma.md) — updated 2026-07-08
  Google's open-weight Gemma 4 family — five Apache 2.0 models from 2.3B to a 30.7B dense flagship, spanning 140+ languages and native audio, vision, and text.
- [GLM-5.3](https://howaiworks.ai/models/glm.md) — updated 2026-08-27
  Zhipu AI's current line: GLM-5.3, a post-trained coding and security flagship, and GLM-5.3-Flash, a 320B/18B multimodal MoE with hybrid attention.
- [GPT-5.6](https://howaiworks.ai/models/gpt.md) — updated 2026-07-24
  OpenAI's flagship family — Sol, Terra and Luna — generally available since July 9, 2026. All three share a 1,050,000 token context and a February 2026 cutoff.
- [Grok 4.5](https://howaiworks.ai/models/grok.md) — updated 2026-07-24
  xAI's flagship model, a mixture-of-experts system trained jointly with Cursor. A 500K token context at $2.00 in / $6.00 out per million tokens.
- [Hailuo 2.3](https://howaiworks.ai/models/hailuo.md) — updated 2026-07-08
  MiniMax's flagship video model, released October 28, 2025. Closed weights, text-to-video and image-to-video up to 1080p, no audio — still its newest.
- [HappyHorse 1.1](https://howaiworks.ai/models/happyhorse.md) — updated 2026-07-08
  Alibaba's video generation model, released June 23, 2026 by the Alibaba Token Hub team. Generates up to 15 seconds of 1080p video with synchronized audio.
- [HiDream-O1-Image](https://howaiworks.ai/models/hidream.md) — updated 2026-07-08
  HiDream.ai's 8B pixel-level unified transformer, open-sourced under MIT on May 8, 2026. No VAE, no separate text encoder, and generates up to 2,048x2,048.
- [Tencent Hy3](https://howaiworks.ai/models/hunyuan.md) — updated 2026-07-08
  Tencent's open-weight flagship, released July 6, 2026. A 295B-parameter MoE with 21B active, hybrid fast-and-slow thinking, a 256K context, and Apache 2.0.
- [HunyuanImage 3.0](https://howaiworks.ai/models/hunyuan-image.md) — updated 2026-07-08
  Tencent's 80B-parameter MoE image model, released September 28, 2025. Natively multimodal and autoregressive, under a restricted community license.
- [Inkling](https://howaiworks.ai/models/inkling.md) — updated 2026-07-18
  Thinking Machines Lab's first open-weights model: a 975B-parameter multimodal MoE with 41B active, 1M-token context and controllable thinking effort.
- [Kimi K3](https://howaiworks.ai/models/kimi.md) — updated 2026-07-17
  Moonshot AI's flagship, announced July 16, 2026. A 2.8T-parameter MoE with 1M context, native vision, always-on thinking, and weights promised by July 27.
- [Kling 3.0](https://howaiworks.ai/models/kling.md) — updated 2026-07-08
  Kuaishou's closed-weight video family, launched February 5, 2026. Four SKUs across video and image, clips of 3–15 seconds with natively generated audio.
- [Ling-2.6-1T](https://howaiworks.ai/models/ling.md) — updated 2026-07-08
  Ant Group's open-weight, non-thinking flagship: ~1T parameters with 63B active, a hybrid MLA and Linear Attention stack, a 1M-token context, and an MIT license.
- [Llama 4](https://howaiworks.ai/models/llama.md) — updated 2026-07-08
  Meta's last open-weight frontier release: Scout and Maverick, natively multimodal Mixture-of-Experts models with 17B active and up to a 10M token context.
- [LongCat-2.0](https://howaiworks.ai/models/longcat.md) — updated 2026-07-08
  Meituan's open-weight flagship, announced June 30, 2026. A 1.6T-parameter MoE with ~48B active parameters, a 1M-token context window, and an MIT license.
- [MiMo-V2.5-Pro](https://howaiworks.ai/models/mimo.md) — updated 2026-07-08
  Xiaomi's open-weight flagship, released April 22, 2026. A 1.02-trillion-parameter MoE with 42B active, a 1M-token context, and an MIT license.
- [MiniMax-M3](https://howaiworks.ai/models/minimax-m3.md) — updated 2026-07-08
  MiniMax's multimodal flagship, released June 1, 2026. A 428B-parameter MoE with ~23B active, a 1M-token context, and open weights under a community license.
- [Mistral Medium 3.5](https://howaiworks.ai/models/mistral-medium.md) — updated 2026-07-08
  Mistral AI's multimodal model for agentic and coding work: a dense 128B model with a 256K context window, released April 2026 under a modified MIT license.
- [MOSS-TTS](https://howaiworks.ai/models/moss-tts.md) — updated 2026-07-08
  OpenMOSS's open-source speech synthesis model (Feb 6, 2026). An 8B autoregressive TTS system under Apache 2.0, with zero-shot voice cloning and 20 languages.
- [Muse Spark](https://howaiworks.ai/models/muse-spark.md) — updated 2026-07-24
  Meta's frontier model, released April 8, 2026 by Meta Superintelligence Labs. A natively multimodal reasoning model with tool use and multi-agent orchestration.
- [Nano Banana](https://howaiworks.ai/models/nano-banana.md) — updated 2026-08-02
  Google's Gemini image family — Pro, 2, 2 Lite and the legacy original. Prices run $0.0336 to $0.24 per image, and the cheapest one outranks the flagship.
- [Nemotron 3 Ultra](https://howaiworks.ai/models/nemotron.md) — updated 2026-07-18
  NVIDIA's open-weight Nemotron 3 flagship: a 550B / 55B-active hybrid Mamba-Transformer MoE with a 1M-token context, NVFP4 pretraining, and an OpenMDW license.
- [Qwen3.7-Max](https://howaiworks.ai/models/qwen.md) — updated 2026-07-08
  Alibaba Cloud's flagship Qwen model, announced May 20, 2026. A closed-weight, API-only agent with a 1M token context window and hybrid thinking on by default.
- [Qwen-Image 2.0](https://howaiworks.ai/models/qwen-image.md) — updated 2026-07-08
  Alibaba's unified image generation and editing model, launched February 10, 2026. Native 2K output, and the line's first flagship shipped with closed weights.
- [Seedance 2.5](https://howaiworks.ai/models/seedance.md) — updated 2026-08-02
  ByteDance Seed video model, launched July 31, 2026: 30-second audio-video in one pass and 50 reference inputs. No technical report or rate card yet.
- [Seedream 5.0](https://howaiworks.ai/models/seedream.md) — updated 2026-07-08
  ByteDance's unified multimodal image model. Seedream 5.0 Lite (Feb 13, 2026) added deep thinking and web search grounding; a Pro tier followed in June 2026.
- [Sora 2](https://howaiworks.ai/models/sora.md) — updated 2026-07-08
  OpenAI's video generation model with synced audio, released September 2025. The consumer app shut down April 2026; the API runs until September 24, 2026.
- [Stable Diffusion 3.5](https://howaiworks.ai/models/stable-diffusion.md) — updated 2026-07-08
  Stability AI's flagship open-weights image model, released October 2024. Three MMDiT variants, still billed as its most powerful image model yet in mid-2026.
- [Step-3.7-Flash](https://howaiworks.ai/models/step.md) — updated 2026-07-08
  StepFun's open-weight vision-language flagship, released May 29, 2026. A 198B MoE activating ~11B per token, served over 400 tokens/second under Apache 2.0.
- [Veo 3.1](https://howaiworks.ai/models/veo.md) — updated 2026-08-02
  Google DeepMind's video model with native audio. Three tiers from $0.05 to $0.60 per second, 4–8 second clips up to 4K, plus video extension.
- [Vidu Q3 Pro](https://howaiworks.ai/models/vidu.md) — updated 2026-07-08
  ShengShu's flagship Vidu Q3 Pro video model, announced January 30, 2026. Generates up to 16 seconds of 1080p video with native audio in a single pass.
- [Wan 2.2](https://howaiworks.ai/models/wan.md) — updated 2026-07-08
  Alibaba's open-weights video model, released July 28, 2025 under Apache 2.0. A 27B MoE diffusion transformer with 14B active, plus a 5B variant for an RTX 4090.
- [Z-Image](https://howaiworks.ai/models/z-image.md) — updated 2026-07-08
  Alibaba Tongyi-MAI's 6B open-source image model. Z-Image-Turbo, released November 26, 2025, generates 8-step images inside 16 GB of VRAM under Apache 2.0.
