MiniMax H3
Omni-modal video generation with native stereo audio.
Only published profiles are counted. More models appear after source review and final editorial approval.
Omni-modal video generation with native stereo audio.
A source-based guide to choosing between GPT-5.6 Sol, Terra, and Luna, with current API token prices, reasoning options, caching rules, and practical workload boundaries.
A documentation-based profile of Qwen3.8-27B, including its native vision-language capabilities, 262K context, thinking controls, supported runtimes, Apache 2.0 license, and official FP8 variant.
A source-based profile of Tencent’s 27B open-weight GUI agent, including its screenshot-driven action loop, official vLLM deployment path, runtime requirements, and computer-use safety boundaries.
A source-based profile of Google DeepMind’s Gemma 4 12B instruction-tuned model, covering its encoder-free multimodal design, 256K context, Transformers run path, and documented evidence boundaries.
A source-based profile of Qwen’s experimental Flash-Next architecture, including its sparse-attention design, 125B/6B-active model structure, long-context claims, supported runtimes, and deployment boundaries.
A source-based profile of Tencent’s Hy4 Preview MoE model, covering its 770B/49B-active architecture, one-million-token context, official vLLM and SGLang recipes, FP8 path, and documented limitations.
An evidence-based guide to NVIDIA?s 64B distilled image-to-video model, its four-step inference paths, supported runtimes, and official multi-GPU hardware boundary.
A source-based overview of Nemotron 3 Nano Omni, including its multimodal architecture, official precision variants, minimum GPU guidance, and supported inference runtimes.