EXPLORE / 01
Models index
Foundation models, releases, and capability boundaries
RECORDS
Entity records
28 records · 1/1
- 001
Qwen3.8-27B
An open-weights dense 27B vision-language model open-sourced by Alibaba Cloud's Qwen team. Built on the Qwen3.5 architecture, it features a 64-layer hybrid attention design (48 linear and 16 gated), native MTP draft head, and thinking mode, supporting native 262K context (scalable to 1M tokens) for single-GPU local deployment and long-horizon agentic tasks.
huggingface.co - 002
GLM-5.3
A 743B-parameter flagship model released by Zhipu AI, enhanced through post-training reinforcement learning scaling with IndexShare, SAO, and Slime for complex terminal execution, coding, and cybersecurity evaluation.
z.ai - 003
Gemini 3.7 Flash
Google's efficient workhorse multimodal model featuring a 1M-token context window and tunable thinking, optimized for long-horizon coding, complex reasoning, and agentic workflows.
blog.google - 004
Qwen3.8-Max
A flagship multimodal foundation model by Alibaba, built on a 2.4-trillion-parameter sparse MoE architecture with 95 billion activated parameters, supporting 1-million-token context and long-horizon autonomous task execution.
qwen.ai - 005
Seedance 2.5
A next-generation multimodal audio-visual generation model by ByteDance, featuring 30-second single-clip generation, 50-slot multimodal reference conditioning, and timestamp-based local editing.
bytedance.com - 006
DeepSeek-V4-Flash
An efficiency-optimized Mixture-of-Experts (MoE) model by DeepSeek, featuring 284B total and 13B active parameters, supporting a 1M context window, and tailored for high throughput, low latency, and agentic workflows.
deepseek.com - 007
Claude Opus 5
Anthropic's frontier reasoning model approaching Claude Fable 5 intelligence, featuring dynamic effort settings and default thinking.
anthropic.com - 008
Gemini 3.5 Flash Cyber
A domain-specific model fine-tuned by Google for cybersecurity, integrated into CodeMender for multi-agent automated vulnerability discovery and remediation.
blog.google - 009
Gemini 3.5 Flash-Lite
A high-throughput, low-latency model released by Google in July 2026, reaching 350 output tokens/sec with built-in computer use tools and configurable thinking levels.
blog.google - 010
Gemini 3.6 Flash
An upgraded workhorse model released by Google in July 2026. It reduces output token usage by 17% compared to Gemini 3.5 Flash while boosting DeepSWE coding benchmarks by up to 65%, optimized for agentic tool-calling.
blog.google - 011
Qwen-Image-3.0
A multimodal image generation base model released by Alibaba's Tongyi Lab in July 2026. Focusing on real-world productivity, it features strong long-text comprehension, complex layout control, precise micro-text rendering, and multilingual UI simulation.
qwen.ai - 012
Qwen-Audio-3.0-TTS
A hosted, production-oriented text-to-speech (TTS) system released by Alibaba's Tongyi Lab in July 2026. It offers a Plus tier for high-quality dubbing and a Flash tier for ultra-low latency real-time voice agents, topping independent quality leaderboards.
bailian.console.aliyun.com - 013
Qwen3.8-Max-Preview
Alibaba Qwen's 2.4-trillion-parameter flagship preview model, first available through Qwen Cloud Token Plan, Qoder, and QoderWork for coding, data analysis, Office workflows, and complex long-horizon agentic tasks. Qwen has previewed an open-weight release, while the date, license, and full evaluations remain undisclosed.
qwen.ai - 014
Kimi K3
Moonshot AI's flagship multimodal reasoning model released in July 2026. It has 2.8 trillion total parameters and combines an MoE architecture with Kimi Delta Attention and Attention Residuals, native vision, and a 1-million-token context window for long-horizon coding, knowledge work, and deep reasoning.
kimi.com - 015
Qwen-Audio-3.0-Realtime
Alibaba's real-time voice interaction model featuring millisecond-level low latency, full-duplex interaction, and proactive Agent tool execution. It is available in Plus (reasoning optimized) and Flash (speed optimized) versions.
qwen.ai - 016
Gemini 3.5 Pro
Google's flagship large language model expected to launch in July 2026, featuring a 2-million-token context window, deep retraining, and optimization for complex logical reasoning and coding tasks.
deepmind.google - 017
Kimi K2.7 Code
Large language model provided by Moonshot AI.
Website pending - 018
Gemini 3.5 Flash
Large language model provided by Google.
Website pending - 019
GLM 5.2
Large language model provided by Zhipu AI.
Website pending - 020
GPT-5.4
Large language model provided by OpenAI.
Website pending - 021
GPT-5.5
Large language model provided by OpenAI.
Website pending - 022
Claude Opus 4.8
Large language model provided by Anthropic.
Website pending - 023
Claude Fable 5
Large language model provided by Anthropic.
Website pending - 024
GPT-5.6 Luna
Large language model provided by OpenAI.
Website pending - 025
GPT-5.6 Terra
Large language model provided by OpenAI.
Website pending - 026
GPT-5.6 Sol
Large language model provided by OpenAI.
Website pending - 027
Hy3
**Hy3** is a 295B-parameter Mixture-of-Experts (MoE) model with 21B active parameters and 3.8B MTP layer parameters, developed by the Tencent Hy Team. Following the Hy3 Preview launch in late April, we gathered feedback from 50+ products and scaled up post-training with higher quality data. Today, we introduce Hy3, which outperforms similar-size models and rivals flagship open-source models with 2-5x parameters. It also shows significant gains in utility across various products and productivity tasks.
Website pending - 028
Claude Sonnet 5
Large language model provided by Anthropic.
Website pending