🧰AI tool library
included 405 AI models · 22 manufacturers · Continuously updated
405 models in total
🧠 Model
AI CodingAI Writing Alibaba (general meaning)
qwen3.8-2.4t-a95b
2.4T parameter open source weight, SMoE architecture, 95B active parameters, 1M context
📊 2.4T total parameters, 95B active parameters 💰/Token per million 📐 1,048,576 Token
🧠 Model
AI CodingAI Productivity DeepSeek
deepseek-v4-pro-0813
Large-scale MoE architecture, 1M context, 384K output, bash.435/bash.87 per million Tokens
📊 Large-scale MoE (parameters are not disclosed) 💰 bash.435/bas 📐 1,048,576 Token
🧠 Model
AI CodingAI Writing Anthropic
Claude Sonnet 5
Anthropic's latest mid-range model, the agent capability is close to Opus 4.8 but the price is 60% lower, 1M context, programming score 63.2%
📊 Undisclosed (guessed hundreds of B levels) 💰/million input tokens, 📐 1M tokens
🧠 Model
AI Coding Sakana AI
Sakana Fugu Ultra
Japan's Sakana AI multi-agent orchestration system, single API dynamic scheduling cutting-edge model pool, SWE-Bench Pro 73.7 points
📊 Multi-agent orchestration architecture, replaceable underlying model pool 💰 Fugu Ultra 0 📐 Depends on the underlying calling model
🧠 Model
AI Coding GLM (GLM)
glm-5-2
Zhipu open source programming model, 753B parameters, 1M context, MIT license, many benchmarks surpass GPT-5.5, the cost is only one-sixth
📊 753B MoE, about 40B active parameters 💰 API $1.40/$4 📐 1,000,000 Token
🧠 Model
AI Coding MiniMax
MiniMax M3
The first open weight model integrating 1 million Token context, native multi-modality and cutting-edge programming capabilities, SWE-Bench Pro surpasses GPT-5.5 by 59%
📊 Open weight, MSA sparse attention architecture 💰 Enter bash.20/million 📐 1 million Token
🧠 Model
AI Video Kling
keye-vl-2
Kuaishou open source 30B MoE video understanding model, 256K ultra-long context, Apache 2.0 free for commercial use, native Agent capabilities
📊 30B MoE, 3B activation parameters, DeepS 💰 Apache 2.0 open source 📐 256K tokens
🧠 Model
AI Coding MiMo
mimo-code
Xiaomi open source terminal AI programming agent, persistent project memory, multi-agent collaboration, MIT protocol, one-click installation
📊 MiMo V2.5 model, million tokens 💰 Free model access, V2.5 📐 1M tokens
🧠 Model
AI Coding Dark Side of the Moon (Kimi)
Kimi K2.7 Code
The Dark Side of the Moon open source programming special model, 384 expert MoE architecture, programming benchmark increased by 21.8%, reasoning token consumption reduced by 30%, compatible with OpenAI SDK
📊 384 Expert MoE (8 active + 1 shared) 💰 API: ~bash.4 📐 128K tokens
🧠 Model
AI Coding Anthropic
claude-fable-5
Anthropic's first public Mythos-level model, qualitative improvement in programming capabilities, 80% Magenta Code benchmark, 1M context window
📊 Undisclosed (Mythos architecture) 💰 0/0 per million tokens ( 📐 1M tokens
🧠 Model
AI Image Gen Ideogram
ideogram-4.0
Open source weighted image generation model, DesignArena is the first open source, native 2K resolution + JSON layout control + transparent background
📊 9.3B parameters, 34-layer DiT architecture, Qwen 💰 API: bash.03 📐 Native 2K resolution output
🧠 Model
AI Coding NVIDIA
cosmos-3
The world's first full-modal world model supports text/image/video/audio/action understanding and generation, a basic physical AI model.
📊 Super/Nano/Edge three versions 💰 Open source weight, recommend DGX C 📐 Support long video sequence understanding and generation
🧠 Model
AI Coding NVIDIA
Nemotron-3-Ultra
NVIDIA 550B parameter open source MoE model is specially built for long-term AI agent workflow and is the most powerful open source AI in the United States.
📊 550B total parameters, MoE architecture, 64 experts each time 💰 Open source and free (you need to bring your own GPU set 📐 Support long sequence reasoning
🧠 Model
AI CodingAI Writing Anthropic
Claude Opus 4.8
Anthropic flagship model, 1 million token contexts, agent programming leads the industry with 69.2%, and honesty optimization reduces illusions
📊 Undisclosed (MoE architecture, estimated trillion-level parameters) 💰 Enter /M token and enter 📐 1 million token input, 12.
🧠 Model
AI CodingAI Productivity Alibaba (general meaning)
Qwen3.7-Plus
Alibaba Tongyi multi-modal agent model, the first in China in Vision Arena, supports GUI operation/CLI call/code self-verification
📊 Undisclosed parameter amount 💰 Open source and free 📐 Long context
🧠 Model
AI CodingAI Writing Anthropic
claude-haiku-4-5
Released in October 2025, Claude Haiku 4.5 is Anthropic’s fastest and most economical model.
📊 Undisclosed (Ultimate Efficiency Architecture) 💰 Enter $1/million tokens 📐 200K tokens
🧠 Model
AI Writing Anthropic
claude-haiku-3-5
Claude Haiku 3.5, released in October 2024, is a lightweight model of the Claude 3.5 family, designed for fast, high-volume tasks.
📊 Undisclosed (3.5th generation lightweight architecture) 💰 Enter $0.80/million to 📐 200K tokens
🧠 Model
AI CodingAI Writing Anthropic
claude-opus-4-7
Claude Opus 4.7 is Anthropic's flagship model released in April 2026. It has 1M context, high-resolution visual understanding, self-verification and xhigh reasoning effort control, and is at the forefront of the industry in coding, agents and complex multi-step tasks.
📊 Undisclosed, Anthropic flagship architecture (estimated 💰 Enter $5/million tokens 📐 1M tokens
🧠 Model
AI CodingAI Writing Anthropic
claude-opus-4-6
Claude Opus 4.6 will be released in February 2026. Adaptive Thinking will be introduced to replace manual effort parameters and can independently determine the depth of reasoning.
📊 Unpublished (same as Opus 4.5/4.7 level frame) 💰 Enter $5/million tokens 📐 1M tokens (beta)
🧠 Model
AI CodingAI Writing Anthropic
claude-sonnet-4-6
Claude Sonnet 4.6, released in February 2026, is close to Opus-level coding and agent capabilities, but has more advantages in speed and cost.
📊 Undisclosed (balanced performance architecture) 💰 Enter $3/million tokens 📐 1M tokens
🧠 Model
AI CodingAI Writing Anthropic
claude-sonnet-4-5
Claude Sonnet 4.5 was released in September 2025 and is recommended by Anthropic as the default model.
📊 Undisclosed (balanced optimization architecture) 💰 Enter $3/million tokens 📐 200K tokens(1M
🧠 Model
AI Image Gen Black Forest Labs
flux-1.1-pro
This pioneering model easily transforms text and images into stunning, detailed, vividly colored visuals, especially for creators and marketers who require high-quality content
📊 12B (Flow Transformer 💰 About $0.04/image
🧠 Model
AI Image Gen Black Forest Labs
flux.1-kontext-pro
Flux AI Vincent graph model, the effect is comparable to Midjourney, crushing StableDiffusion
📊 12B (Flow Transformer 💰 Starting from about $0.04/image
🧠 Model
AI Writing DeepSeek
deepseek-chat
The most powerful open source MoE model DeepSeek-V3.2
📊 DeepSeek-V3.2, 671B general parameters
🧠 Model
AI Writing DeepSeek
deepseek-r1-250528
DeepSeek latest model.
📊 DeepSeek-R1-250528 (same as
🧠 Model
AI Writing DeepSeek
deepseek-r1-250120
deepseek-r1 is a deep thinking model launched by deepseek.
📊 DeepSeek-R1-250120 (Initial
🧠 Model
AI Writing DeepSeek
deepseek-r1-distill-llama-70b
DeepSeek-R1 based on Llama-3.3-70B, which uses reinforcement learning technology on a large scale during the training phase
📊 70B parameters, Dense architecture (based on Llam
🧠 Model
AI Writing DeepSeek
deepseek-r1-distill-qwen-32b
DeepSeek-R1 based on Qwen2.5-32B, which uses reinforcement learning technology on a large scale during the training phase
📊 32B parameters, Dense architecture (based on Qwen
🧠 Model
AI Writing DeepSeek
deepseek-r1-distill-qwen-7b
DeepSeek-R1 based on Qwen2.5-Math-7B, which uses reinforcement learning technology on a large scale during the training phase
📊 7B parameters, Dense architecture (based on Qwen2
🧠 Model
AI Writing DeepSeek
deepseek-r1-searching
Deepseek-r1 is deployed in the open source version and can be searched online.
📊 Based on DeepSeek-R1 open source version deployment + connection
🧠 Model
AI Writing DeepSeek
deepseek-reasoner
deepseek-r1
📊 API alias for DeepSeek-R1 (point to
🧠 Model
AI Writing DeepSeek
deepseek-v3-250324
deepseek's latest large model, improved reasoning task performance, enhanced front-end development capabilities, upgraded Chinese writing, and optimized Chinese search capabilities
📊 DeepSeek-V3-250324 (same as
🧠 Model
AI Writing DeepSeek
deepseek-v3-0324
deepseek's latest large model, improved reasoning task performance, enhanced front-end development capabilities, upgraded Chinese writing, and optimized Chinese search capabilities
📊 DeepSeek-V3-0324, 671
🧠 Model
AI Writing DeepSeek
deepseek-v3-1-think-250821
The default points to the latest deepseek-v3-1-terminus
📊 DeepSeek-V3.1-Think(
🧠 Model
AI Writing DeepSeek
deepseek-v3.1
Deepseek's newly launched hybrid reasoning model supports two reasoning modes: thinking and non-thinking, and is more efficient than deepseek-r1-0528.
📊 DeepSeek-V3.1, 685B general parameters
🧠 Model
AI Writing DeepSeek
deepseek-v3.1-fast
DeepSeek-v3.1-fast is an extremely exciting breakthrough product designed to solve the problem of "speed and depth" of large models.
📊 DeepSeek-V3.1-Fast, base
🧠 Model
AI CodingAI Writing DeepSeek
deepseek-v3.2
DeepSeek-V3.2 is the first model we launched that integrates thinking into tool use, and supports both thinking mode and non-thinking mode tool invocation.
📊 DeepSeek-V3.2, 671B general parameters
🧠 Model
AI Writing DeepSeek
deepseek-v3.2-fast
DeepSeek-V3.2-Fast is a high-performance, low-latency variant based on the DeepSeek V3.2 architecture
📊 DeepSeek-V3.2-Fast, base
🧠 Model
AI CodingAI Writing DeepSeek
deepseek-v4-flash
DeepSeek-V4-Flash is a lightweight version of the DeepSeek V4 series. It focuses on high cost performance and high throughput efficiency. It is suitable for general conversations and basic text tasks. It also supports millions of Token long contexts and efficient reasoning.
📊 DeepSeek-V4-Flash, 28
🧠 Model
AI CodingAI Writing DeepSeek
deepseek-v4-pro
DeepSeek-V4-Pro is a high-performance open source large model launched by DeepSeek. It has top reasoning and agent capabilities, supports ultra-long context, adapts to domestic Ascend chips, and is extremely cost-effective.
📊 DeepSeek-V4-Pro, 1.6T
🧠 Model
AI Writing Google
gemini-2.0-flash
Google 2.0 launches latest model that outperforms state-of-the-art (O1) models in multi-modal AI capabilities
📊 Context window: 1M tokens | Max
🧠 Model
AI Writing Google
gemini-2.5-flash-lite-preview-09-2025
Gemini 2.5 Flash-Lite-Preview-09-2025 is a sub-model of the Gemini model family that focuses on ultra-low latency, high concurrency and the highest cost performance.
📊 Context window: 1M tokens | Max
🧠 Model
AI Writing Google
gemini-2.5-flash-preview-09-2025
Gemini 2.5 Flash 09-2025 is the latest public preview version of the Gemini 2.5 Flash model launched by Google. It focuses on significantly improving the intelligence and application capabilities of the model while maintaining high efficiency and low cost.
📊 Context window: 1M tokens | Max
🧠 Model
AI Writing Google
gemini-3-pro-preview-11-2025
Gemini 3 is Google’s smartest model series yet, powered by advanced inference capabilities.
📊 Context window: 1M tokens | Max
🧠 Model
AI CodingAI Writing Google
gemini-3-flash-preview
gemini-3-flash-preview The smartest model we've ever built, combining speed and cutting-edge intelligence with exceptional search and grounding capabilities.
📊 Context window: 1M tokens | Max
🧠 Model
AI Image Gen Google
gemini-3.1-flash-image-preview
Nano Banana 2 delivers high-quality image generation and conversational editing with low latency at a mainstream price.
📊 Codename: Nano Banana 2 | Base
🧠 Model
AI CodingAI Writing Google
gemini-3.1-flash-lite
The most cost-effective model from Google, optimized for high-volume agent tasks, translation, and simple data processing.
📊 Context window: 1M tokens | Max
🧠 Model
AI CodingAI Writing Google
gemini-3.5-flash
Gemini 3.5 Flash has been officially released (GA), has stable performance and can be used in large-scale production environments.
📊 Context window: 1M tokens | Max
🧠 Model
AI Productivity Google
gemini-embedding-001
The Gemini API provides text embedding models for generating embeddings for words, phrases, sentences, and code.
📊 Embedding dimension: 3072 (default) | Matr
🧠 Model
AI Writing Google
gemini-flash-latest
Gemini Flash Latest is an artificial intelligence model provided by google-vertex.
📊 Dynamic aliases: always point to the latest Flash version
🧠 Model
AI Productivity Google
gemini-embedding-2-preview
The latest model gemini-embedding-2-preview is the first multi-modal embedding model in the Gemini API.
📊 Embedding dimension: 128-3072 adjustable (recommended 76
🧠 Model
AI Productivity Google
gemma-2b-it
gemma is a series of lightweight and advanced open source models launched by Google, built based on the research and technology of Google Gemini model.
📊 Parameters: 2.5B | Architecture: Decoder
🧠 Model
AI Writing Google
gemma-7b-it
gemma is a series of lightweight and advanced open source models launched by Google, built based on the research and technology of Google Gemini model.
📊 Parameters: 8.5B | Architecture: Decoder
🧠 Model
AI VideoAI Audio Google
omni-flash
Gemini Omni Flash is a multi-modal AI video model launched by Google.
📊 Codename: Gemini Omni Flash
🧠 Model
AI Video Google
veo3.1
Google's latest advanced artificial intelligence model, veo3.1 supports automatic audio generation for videos, with high quality and low price, the most cost-effective option. It supports first and last frames and Vincent videos (8 seconds in duration)
📊 Resolution: 1080p | Duration: 8 seconds |
🧠 Model
AI Video Google
veo3.1-pro
Google's latest advanced artificial intelligence model, veo3.1 high-quality mode, supports automatic audio generation for videos. The quality is extremely high and the price is extremely high. Please be careful when using it. It supports first and last frames and Vincent videos (8 seconds in duration).
📊 Resolution: 1080p (4K upscaling possible) |
🧠 Model
AI Video Google
veo3.1-4k
Google's latest advanced artificial intelligence model, veo3.1 4k mode, supports automatic audio generation for videos, with high quality and low price, the most cost-effective option (8 seconds duration)
📊 Resolution: Native 4K (3840×2160)
🧠 Model
AI Video Google
veo3.1-fast-components
Google's latest advanced artificial intelligence model veo3.1-fast-components supports three cushion image inputs, which is faster, more stable and more controllable, redefining the creative freedom of Tusheng videos (8 seconds in duration)
📊 Resolution: 1080p | Duration: 8 seconds |
🧠 Model
AI Video Google
veo3.1-pro-4k
Google's latest advanced artificial intelligence model, veo3.1 4k high-quality mode, supports automatic audio generation for videos. The quality is extremely high and the price is extremely high. Please be careful when using it. It supports first and last frames and Vincent videos (8 seconds in duration).
📊 Resolution: Native 4K (3840×2160)
🧠 Model
AI Video Google
veo_3_1-components
Google's latest advanced artificial intelligence model, veo_3_1 4k mode, supports automatic audio generation for videos, supports first frame delivery, does not support last frames, high quality and low price, the most cost-effective option (8 seconds duration)
📊 Resolution: 1080p | Duration: 8 seconds |
🧠 Model
AI Video Google
veo_3_1-components-4K
Google's latest advanced artificial intelligence model, veo_3_1 4k mode, supports automatic audio generation for videos, supports first frame delivery, does not support last frames, high quality and low price, the most cost-effective option (8 seconds duration)
📊 Resolution: Native 4K (3840×2160)
🧠 Model
AI Video Google
veo_3_1-fast-4K
Google's latest advanced artificial intelligence model, veo3.1 fast +4k mode, supports automatic video and audio generation, high quality and low price, the most cost-effective option, supports first and last frames and Vincent videos (8 seconds in duration)
📊 Resolution: Native 4K (3840×2160)
🧠 Model
AI Video Google
veo_3_1-fast-components-4K
Google's latest advanced artificial intelligence model veo3.1-fast-components supports three cushion image inputs, which is faster, more stable, and more controllable. It redefines the creative freedom of Tusheng videos, 4k model (8 seconds in duration)
📊 Resolution: Native 4K (3840×2160)
🧠 Model
AI Writing Meta
llama-2-13b
High-capacity Llama models suitable for complex analysis and forecasting tasks
📊 13B (standard Dense architecture)
🧠 Model
AI Writing Meta
llama-2-70b
High-capacity Llama models suitable for complex analysis and forecasting tasks
📊 70B (standard Dense architecture)
🧠 Model
AI Writing Meta
llama-2-7b
Llama-2-7B is the least parameterized but still powerful version of the entire Llama 2 family
📊 7B (standard Dense architecture)
🧠 Model
AI Writing Meta
llama-3-70b
The latest Meta Llama 3 model with 7 billion parameters
📊 70B (Dense architecture, please note "" in desc
🧠 Model
AI Writing Meta
llama-3-8b
The latest Meta Llama 3 model with 700 million parameters
📊 8B (Dense architecture, pay attention to "7" in desc
🧠 Model
AI Writing Meta
llama-3-sonar-large-32k-chat
Llama 3rd generation 32K context model
📊 70B (based on Llama 3 70B, extended
🧠 Model
AI Writing Meta
llama-3-sonar-small-32k-chat
Llama 3rd generation 32K context model
📊 8B (based on Llama 3 8B, extended by 32
🧠 Model
AI Writing Meta
llama-3.1-405b
Meta’s 405B parameter Llama 3.1 model has powerful general capabilities
📊 405B (Dense architecture, open source largest Den
🧠 Model
AI Writing Meta
llama-3.1-405b-instruct
Meta's 405B parameter Llama 3.1 command fine-tuned version, more suitable for dialogue scenes
📊 405B (Dense architecture, instruction fine-tuned version)
🧠 Model
AI Writing Meta
llama-3.1-70b
Meta’s 70B parameter Llama 3.1 model balances performance and efficiency
📊 70B (Dense architecture)
🧠 Model
AI Writing Meta
llama-3.1-70b-instruct
Meta's 70B parameter Llama 3.1 fast instruction version, faster response
📊 70B (Dense architecture, instruction fine-tuned version)
🧠 Model
AI Writing Meta
llama-3.1-8b
Meta's 8B parameter Llama 3.1 model, a lightweight version suitable for resource-constrained scenarios
📊 8B (Dense architecture)
🧠 Model
AI Writing Meta
llama-3.2-11b-vision-instruct
Meta's 11B parameter Llama 3.2 instruction version, lightweight dialogue model
📊 11B (Dense architecture + visual encoder)
🧠 Model
AI Writing Meta
llama-3.2-1b-instruct
Meta's 1B parameter Llama 3.2 instruction version, ultra-lightweight dialogue model
📊 1B (Dense architecture, ultra-lightweight)
🧠 Model
AI Writing Meta
llama-3.2-3b-instruct
@cf/meta/llama-3.2-3b-instruct is an artificial intelligence model provided by cloudflare-workers-ai.
📊 3B (Dense architecture, edge deployment optimization)
🧠 Model
AI Writing Meta
llama-3.2-90b-vision-instruct
Meta's 90B parameter Llama 3.2 command fine-tuned version, more suitable for dialogue scenes
📊 90B (Dense architecture + visual encoder)
🧠 Model
AI Writing Meta
llama-3.3-70b-instruct
Meta-Llama-3-70B-Instruct is a fine-tuned version of the 70B parameter instruction. It is suitable for dialogue scenarios and performs better in understanding language details, context and performing complex tasks.
📊 70B (Dense architecture, Llama 3 series
🧠 Model
AI Writing Meta
meta-llama/llama-4-maverick
llama's best-in-class native multi-modal model delivers superior text for seamless long document analysis.
📊 About 400B total parameters/17B activation (MoE architecture,
🧠 Model
AI Writing Meta
meta-llama/llama-4-scout
llama's best-in-class native multi-modal model delivers superior text, single H100 GPU efficiency, and 10M context windows for seamless long document analysis.
📊 About 109B total parameters/17B activation (MoE architecture,
🧠 Model
AI DesignAI Image Gen Midjourney
mj_blend
Midjourney blend mode
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_custom_zoom
Midjourney custom scaling
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_describe
Midjourney description mode
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_high_variation
Midjourney high variation mode
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_imagine
Midjourneyimagination mode
💰 Subscription starting from $10-$60/month
🧠 Model
AI DesignAI Image Gen Midjourney
mj_inpaint
Midjourney partial redraw
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_low_variation
Midjourney low variation mode
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_modal
Midjourney modal pattern
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_pan
Midjourney pan mode
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_reroll
Midjourney regenerated
💰 Midjourney subscription
🧠 Model
AI Productivity Midjourney
mj_upload
Midjourney upload service
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_upscale
Midjourney zoom mode
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_variation
Midjourney variant mode
💰 Midjourney subscription
🧠 Model
AI DesignAI Image Gen Midjourney
mj_zoom
Midjourney zoom mode
💰 Midjourney subscription
🧠 Model
AI Video MiniMax
MiniMax-Hailuo-02
Conch AI 02 brings your images to life through advanced AI motion synthesis technology.
💰 About $0.28/10 seconds HD
🧠 Model
AI Video MiniMax
MiniMax-Hailuo-2.3
MiniMax video model Hailuo 2.3 further upgrades the dynamic expressiveness based on the Hailuo 02 model, making the picture more realistic and stable.
💰 Same price as version 02, more cost-effective
🧠 Model
AI Writing MiniMax
MiniMax-M2
MiniMax M2 lets the model generate dialogue content and tool calls based on the input context.
📊 230B total parameters/10B activation (MoE) 💰 About $0.30/M input to 📐 205K tokens
🧠 Model
AI Writing MiniMax
MiniMax-M2.5
MiniMax-M2.5 has reached or refreshed the industry's SOTA in productivity scenarios such as programming, tool calling and searching, and office work.
📊 230B total parameters/10B activation (MoE) 💰 About $0.27/M input, $ 📐 205K tokens
🧠 Model
AI CodingAI Writing MiniMax
MiniMax-M2.7
MiniMax-M2.7 has reached or refreshed the industry's SOTA in productivity scenarios such as programming, tool calling and searching, and office work.
📊 230B total parameters/10B activation (MoE) 💰 About $0.30/M input, $ 📐 204.8K tokens
🧠 Model
AI Audio MiniMax
MiniMax-Voice-Design
Conch sound design
💰 Billed by character or number of calls
🧠 Model
AI Writing MiniMax
mimo-v2-pro
Xiaomi MiMo-V2-Pro is designed for demanding real-world agency workflows.
📊 1T+general parameter/42B activation (MoE) 💰 About $1.00/M input, $ 📐 1M tokens
🧠 Model
AI CodingAI Writing MiniMax
mimo-v2.5
MiMo-V2.5, a native full-modal model, supports text, image, video and audio understanding, and has powerful Agent capabilities
📊 1T+general parameter/42B activation (MoE) 💰 About $0.45/M input, $ 📐 1M tokens
🧠 Model
AI CodingAI Writing MiniMax
mimo-v2.5-pro
MiMo-V2.5-Pro is oriented to complex task scenarios and is deeply adapted to Agent and Coding applications. It ranks first in the world's open source models on the GDPVal-AA and ClawEval lists.
📊 1.02T total parameter/42B activation (MoE) 💰 About $0.44/M input, $ 📐 1M tokens
🧠 Model
AI Writing MiniMax
minimax-m2.1
Powerful multi-language programming capabilities, comprehensively upgrade the programming experience
📊 230B total parameters/10B activation (MoE) 💰 Already adopted by M2.5/M2.7 📐 205K tokens
🧠 Model
AI Writing Mistral
Dolphin3.0-R1-Mistral-24B
Mistral Large is Mistral AI’s flagship model with top-level reasoning capabilities.
📊 24B (Dense) 💰 About $0.03/M input, $ 📐 32.8K tokens
🧠 Model
AI Writing Mistral
mistral-large-latest
Mistral Large is Mistral AI’s flagship model with top-level reasoning capabilities.
📊 675B General Parameter/41B Activation (MoE) 💰 About $0.50/M input, $ 📐 262K tokens
🧠 Model
AI Writing Mistral
mistral-small-latest
Mistral Small is an optimized model that focuses on low latency and cost-effectiveness.
📊 119B General Parameter/6B Activation (MoE) 💰 About $0.15/M input to 📐 256K tokens
🧠 Model
AI Productivity OpenAI
babbage-002
Replacement for GPT-3 ada and babbage base models.
📊 Unpublished (about 1.3B) 💰 Very low, enter $0.4/M 📐 16K
🧠 Model
AI Writing OpenAI
chatgpt-4o-latest
Dynamic models in ChatGPT are continuously updated to the current version of GPT-4o.
📊 Undisclosed (~200B) 💰 $2.5/$10 per 📐 128K
🧠 Model
AI CodingAI Writing OpenAI
codex-mini-2025-05-16
codex-mini-2025-05-16 is a lightweight model-specific snapshot optimized for code generation and completion. It has both low latency and high cost performance, and is suitable for real-time programming assistance and high-concurrency code processing scenarios.
📊 Based on o4-mini 💰 Via Codex CLI 📐 128K
🧠 Model
AI Image Gen OpenAI
dall-e-3
Generate a new version of the image DALL-E.
📊 Unpublished (image generated) 💰 Charged according to image size, about $0. 📐 N/A
🧠 Model
AI Writing OpenAI
davinci-002
Replacement for GPT-3 curie and davinci base models.
📊 Unpublished (~175B original version) 💰 High, $2/$2 per 📐 16K
🧠 Model
AI Writing OpenAI
gpt-3.5-turbo-1106
Pure official high-speed GPT3.5 series, supports tools_call
📊 Undisclosed (~20B) 💰 $1/$2 per M 📐 16K
🧠 Model
AI Productivity OpenAI
gpt-3.5-turbo-0613
Pure official high-speed GPT3.5 series, supports function_call
📊 Undisclosed (~20B) 💰 $1.5/$2 per 📐 4K
🧠 Model
AI Writing OpenAI
gpt-3.5-turbo-16k-0613
June 13, 2023 version of OpenAI GPT-3.5 Turbo 16K
📊 Undisclosed (~20B) 💰 $3/$4 per M 📐 16K
🧠 Model
AI Writing OpenAI
gpt-3.5-turbo-16k
Pure official high-speed GPT3.5 16K series, supports function_call
📊 Undisclosed (~20B) 💰 $3/$4 per M 📐 16K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5
GPT-5 is our flagship model for cross-domain coding, reasoning, and agent tasks.
📊 Undisclosed (trillions) 💰$1.25/$10pe 📐 400K (272K in +128K
🧠 Model
AI Writing OpenAI
gpt-4-0613
Purely official GPT4 series, 0613 series models support function_call
📊 Undisclosed (~1.8T MoE) 💰 $30/$60 per 📐 8K
🧠 Model
AI Productivity OpenAI
gpt-4-0125-preview
The latest gpt-4-0125-preview, an upgraded version of gpt-4-1106-preview, has stronger code generation capabilities, reduces model "laziness", and fixes non-English UTF-8 generation problems.
📊 Undisclosed (~1.8T MoE) 💰$10/$30 per 📐 128K
🧠 Model
AI Productivity OpenAI
gpt-4-1106-preview
The latest gpt-4-1106-preview, also known as gpt-4-turbo, is 67% cheaper than gpt-4, supports 128k context, supports tools, and the knowledge deadline is April 2023
📊 Undisclosed (~1.8T MoE) 💰$10/$30 per 📐 128K
🧠 Model
AI Writing OpenAI
gpt-4-32k-0613
32K context window version of OpenAI GPT-4, updated June 13, 2023
📊 Undisclosed (~1.8T MoE) 💰 $60/$120 per 📐 32K
🧠 Model
AI Writing OpenAI
gpt-4-32k
Pure official GPT4 32K series
📊 Undisclosed (~1.8T MoE) 💰 $60/$120 per 📐 32K
🧠 Model
AI Writing OpenAI
gpt-4-all
GPT All model, integrating official GPT-4, networking, image reading, drawing functions, and code interpreter
📊 Undisclosed (~1.8T MoE) 💰 Pricing based on GPT-4 + additional features 📐 8K-32K
🧠 Model
AI Writing OpenAI
gpt-4-turbo-preview
gpt-4-turbo-preview upgraded version, with stronger code generation capabilities, reducing model "laziness" and fixing non-English UTF-8 generation problems
📊 Undisclosed (~1.8T MoE) 💰$10/$30 per 📐 128K
🧠 Model
AI Writing OpenAI
gpt-4-gizmo-*
All GPTs on the official website can be called
📊 Unpublished (based on GPT-4) 💰 According to GPT-4 price + GPT 📐 8K-128K
🧠 Model
AI CodingAI Writing OpenAI
gpt-4.1
GPT-4.1 is an artificial intelligence model provided by openai.
📊 Undisclosed (~1T MoE Est) 💰$2/$8 per M 📐 1M
🧠 Model
AI CodingAI Writing OpenAI
gpt-4.1-mini
GPT-4.1 mini is an artificial intelligence model provided by openai.
📊 Unpublished (medium) 💰 $0.4/$1.6 per 📐 1M
🧠 Model
AI CodingAI Writing OpenAI
gpt-4.1-nano
GPT-4.1 nano is an artificial intelligence model provided by openai.
📊 Unpublished (~8B Est) 💰 $0.1/$0.4 per 📐 1M
🧠 Model
AI Productivity OpenAI
gpt-4.5-preview
openai latest model, gpt-4.5
📊 Undisclosed (very large amount of calculation) 💰$75/$150 per 📐 128K
🧠 Model
AI Writing OpenAI
gpt-4o
GPT-4o is an artificial intelligence model provided by openai.
📊 Undisclosed (~200B Est) 💰 $2.5/$10 per 📐 128K
🧠 Model
AI Writing OpenAI
gpt-4o-all
GPT All model, integrating official GPT-4o, networking, image reading, drawing functions, and code interpreter
📊 Unpublished (based on GPT-4o) 💰 Pricing based on GPT-4o + add-ons 📐 128K
🧠 Model
AI Writing OpenAI
gpt-4o-mini
GPT-4o mini is an artificial intelligence model provided by openai.
📊 Unpublished (~8B Est) 💰 $0.15/$0.6p 📐 128K
🧠 Model
AI Audio OpenAI
gpt-4o-mini-audio-preview
The Audio API allows developers to build speech-to-text and text-to-speech functionality.
📊 Unpublished (based on GPT-4o-mini) 💰 Additional billing for audio tokens 📐 128K
🧠 Model
AI Writing OpenAI
gpt-5-all
gpt-5-all is the latest flagship model of openai
📊 Unpublished (based on GPT-5) 💰 Pricing based on GPT-5 + additional features 📐 400K
🧠 Model
AI Writing OpenAI
gpt-5-chat-latest
GPT-5 Chat points to the GPT-5 snapshot version currently used by ChatGPT.
📊 Unpublished (based on GPT-5 snapshot) 💰 Use via ChatGPT 📐 400K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5-codex
GPT-5-Codex is a version of GPT-5 optimized for proxy encoding in Codex.
📊 Unpublished (based on GPT-5) 💰 Via Codex CLI/ 📐 400K
🧠 Model
AI Writing OpenAI
gpt-5-mini-2025-08-07
GPT-5 mini is a faster and more cost-effective version of GPT-5, especially suitable for processing clearly defined tasks and precise instructions.
📊 Unpublished (Lightweight) 💰 $0.25/$2 per 📐 272K (in) + 128K (out)
🧠 Model
AI Writing OpenAI
gpt-5-nano-2025-08-07
GPT-5 Nano is our fastest and most affordable version of GPT-5, especially suited for summarization and classification tasks.
📊 Unpublished (extremely lightweight) 💰 $0.1/$0.8 per 📐 272K (in) + 128K (out)
🧠 Model
AI Writing OpenAI
gpt-5-search-api
gpt-5 dedicated search model
📊 Unpublished (based on GPT-5) 💰 $1.25/$10 + 📐 400K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.1
GPT-5.1: Smarter, more conversational chat GPT, the most commonly used model, is now warmer, smarter, and better able to follow your commands.
📊 Undisclosed (trillions) 💰$1.25/$10pe 📐 400K
🧠 Model
AI Writing OpenAI
gpt-5.1-all
GPT-5.1: Smarter, more conversational chat GPT, the most commonly used model, is now warmer, smarter and better at following your commands.
📊 Unpublished (based on GPT-5.1) 💰 Pricing based on GPT-5.1 + attached 📐 400K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.1-codex
GPT-5.1-codex is a well-equipped heavy-duty main force.
📊 Unpublished (based on GPT-5.1) 💰 Via Codex CLI/ 📐 400K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.1-codex-max
GPT-5.1-codex-max breaks through the limits of context understanding and logical reasoning, and is designed to coordinate system-level reconstruction across warehouses, independently design and implement the full-stack ecosystem.
📊 Undisclosed (the strongest computing Codex) 💰 Extremely high, used through Codex 📐 400K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.1-codex-mini
GPT-5.1-codex-mini An efficient model for rapid code generation and lightweight development tasks
📊 Unpublished (Lightweight Codex) 💰 Lower than GPT-5.1-Co 📐 400K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.2
GPT-5.2 is the best model for coding and intelligence tasks in a wide range of industries
📊 Undisclosed (trillions, higher than 5.1) 💰$1.75/$14pe 📐 400K
🧠 Model
AI Writing OpenAI
gpt-5.2-all
The best model for coding and intelligence tasks in every industry
📊 Unpublished (based on GPT-5.2) 💰 Pricing based on GPT-5.2 📐 400K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.2-codex
It is a god-level programming AI that can understand extremely complex system architectures, program entire software projects with zero errors, and even self-correct errors.
📊 Unpublished (based on GPT-5.2) 💰 Via Codex CLI/ 📐 400K
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.3-chat-latest
GPT-5.3-chat-latest is a high-speed dialogue model optimized by OpenAI. The response is more direct and natural, reducing preaching and refusal, and is suitable for daily lightweight tasks.
📊 Undisclosed, estimated to be about 1 trillion parameters (MoE mixing expert 💰 Input about $1.75/million tons 📐 400K tokens
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.3-codex
GPT-5.3-Codex redefines the role of AI in programming and pan-productivity through performance improvements, functional generalization, and security upgrades.
📊 Undisclosed, estimated to be about 1 trillion parameters (MoE mixing expert 💰 Input about $1.75/million tons 📐 400K tokens
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.3-codex-spark
GPT-5.3-Codex-Spark Research Preview.
📊 Unpublished, lightweight version, parameter size is smaller than GPT- 💰 Input about $1.75/million tons 📐 400K tokens
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.4
GPT-5.4 is our cutting-edge model for complex professional work.
📊 Undisclosed, estimated to be about 1.5 trillion parameters (MoE hybrid 💰 Enter $2.50/million to 📐 1M tokens (experimental, required
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.4-pro
GPT-5.4pro uses more computing resources to think deeper and provide always better answers.
📊 Based on GPT-5.4 architecture, using more inference calculations 💰 Enter about $15/million tok 📐 1M tokens
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.4-mini
GPT-5.4mini combines the advantages of GPT-5.4 into a faster, more efficient model designed for high-load workloads.
📊 Undisclosed, estimated parameters are about 100B-200B 💰 Input about $0.75/million tons 📐 272K tokens
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.4-nano
GPT‑5.4 nano is the lightest and fastest version of GPT‑5.4, designed for tasks where speed and cost are critical.
📊 Undisclosed, estimated parameters are about 20B-40B 💰 Input about $0.20/million tons 📐 272K tokens
🧠 Model
AI Writing OpenAI
gpt-5.5
GPT-5.5 is the flagship large language model released by OpenAI on April 24, 2026. It is positioned as a new type of intelligence for practical work and agents. The core breakthrough lies in autonomous planning and execution of multi-step complex tasks. It is good at programming, computer operations, scientific research analysis and other fields, and does it with less T
📊 Undisclosed, estimated to exceed 2 trillion parameters (latest generation MoE 💰 Enter $5/million tokens 📐 1M tokens
🧠 Model
AI CodingAI Writing OpenAI
gpt-5.5-pro
GPT-5.5pro is now available for processing Responses API requests, including operating through the BBatch API to support multiple rounds of model interaction capabilities before responding to API requests, and will implement other advanced API features in the future.
📊 Based on GPT-5.5 architecture, using maximum inference calculation 💰 Enter about $30/million tok 📐 1M tokens
🧠 Model
AI Image Gen OpenAI
gpt-image-1-all
Supports the reverse model of charging for all parameters of /v1/images/generations and /v1/images/edits
📊 Same as gpt-image-1 💰 Third-party channel prices are usually lower than 📐 Same as gpt-image-1
🧠 Model
AI Image Gen OpenAI
gpt-image-1-mini
gpt-image-1-mini is our lightweight image generation model
📊 A lightweight version of the image generation model, the parameter size is smaller than gpt- 💰 Approximately gpt-image- 📐 Support common image sizes
🧠 Model
AI Image Gen OpenAI
gpt-image-1.5
GPT Image 1.5 is our latest image generation model, with better instruction tracking and prompt following.
📊 Undisclosed, second-generation image generation architecture 💰 About $0.05-0.15/ 📐 Supports flexible size, up to 4K resolution
🧠 Model
AI Image Gen OpenAI
gpt-image-1.5-all
Image generation model with better instruction tracking and prompt following.
📊 Same as gpt-image-1.5 💰 Third-party channel prices are usually lower than 📐 Same as gpt-image-1.5
🧠 Model
AI Image Gen OpenAI
gpt-image-2-all
gpt-image-2-all is our most advanced image generation model, supporting fast image generation and editing.
📊 Same as gpt-image-2 💰 Third-party channel prices 📐 Same as gpt-image-2
🧠 Model
AI Writing OpenAI
gpt-oss-120b
GPT OSS 120B is an artificial intelligence model provided by vultr.
📊 About 117B total parameters, 5.1B active parameters (Mo 💰 Open source and free (self-deployment), third 📐 256K tokens
🧠 Model
AI Writing OpenAI
gpt-oss-20b
The Gpt-oss-20b model achieves similar results to the OpenAI o3‑mini model on common benchmarks.
📊 About 20B parameters 💰 Open source and free (self-deployment), third 📐 256K tokens
🧠 Model
AI Audio OpenAI
gpt-realtime-1.5-2026-02-23
GPT-Reatime-1.5 is our flagship audio model for voice agents and customer support.
📊 Unpublished, optimized based on GPT-4o real-time architecture 💰 About $2/minute audio input, $ 📐 Real-time audio streaming to support ongoing conversations
🧠 Model
AI Writing OpenAI
o1
The full version of o1 will take more time to think (form a chain of ideas) to arrive at the answer, making it more suitable for performing complex reasoning tasks, especially scientific and mathematical tasks.
📊 Unpublished, based on GPT-4o level architecture + inference chain 💰 Enter $15/million tokens 📐 200K tokens
🧠 Model
AI Writing OpenAI
o1-all
Spends more time thinking (forming chains of ideas) to arrive at an answer, making it better suited for complex reasoning tasks, especially science and math tasks
📊 Unpublished, based on GPT-4o level architecture + inference layer 💰 Third-party channel price (official o1 📐 200K tokens
🧠 Model
AI Productivity OpenAI
o1-mini
o1-mini is a fast, cost-effective inference model tailored for coding, math and science use cases.
📊 Undisclosed, lightweight reasoning architecture 💰 Enter $1.10/million to 📐 200K tokens
🧠 Model
AI Writing OpenAI
o1-mini-all
<div>The o1-mini-all model integrates the basic capabilities of o1-mini, networking, image reading, drawing functions, and code interpreter. The file link can be placed anywhere in the prompt.</div>
📊 Same as o1-mini 💰 Third-party channel prices 📐 200K tokens
🧠 Model
AI Writing OpenAI
o3
o3 is an artificial intelligence model provided by openai.
📊 Undisclosed, second generation reasoning architecture 💰 Enter $2/million tokens 📐 200K tokens
🧠 Model
AI Writing OpenAI
o3-all
OpenAI's most advanced reasoning model is good at multi-modal analysis and supports thinking functions
📊 Unpublished, the latest reasoning architecture 💰 Third-party channel price (official o3 📐 200K tokens
🧠 Model
AI Writing OpenAI
o3-deep-research
o3-deep-research is an artificial intelligence model provided by openai.
📊 Based on o3 architecture, integrating in-depth search and research capabilities 💰 About $5/million tokens 📐 200K tokens
🧠 Model
AI Writing OpenAI
o3-mini
o3-mini is OpenAI’s latest small-scale inference model, optimized for coding, math, and scientific tasks.
📊 Undisclosed, lightweight reasoning architecture 💰 Enter $1.10/million to 📐 200K tokens
🧠 Model
AI Writing OpenAI
o3-mini-all
o3-mini is openai’s latest thinking model
📊 Same as o3-mini 💰 Third-party channels, about $0.55 📐 200K tokens
🧠 Model
AI Writing OpenAI
o3-mini-high-all
o3-mini-high is the latest thinking model of openai, which is slightly smarter than o3-mini
📊 Same as o3-mini, using high inference setting 💰 The price of third-party channels is slightly higher than o 📐 200K tokens
🧠 Model
AI Writing OpenAI
o4-mini
o4-mini is OpenAI’s latest efficient inference model, with low latency and high quality, and supports multi-modal input
📊 Unpublished, the latest lightweight inference architecture 💰 Enter $1.10/million to 📐 200K tokens
🧠 Model
AI Writing OpenAI
o4-mini-all
o4-mini is OpenAI’s latest efficient inference model, with low latency and high quality, and supports multi-modal input
📊 Same as o4-mini 💰 Third-party channel prices 📐 200K tokens
🧠 Model
AI Writing OpenAI
o4-mini-deep-research
o4-mini-deep-research is our faster and more affordable deep research model, ideal for handling complex multi-step research tasks.
📊 Based on o4-mini architecture, integrated deep search 💰 About $1/million tokens 📐 200K tokens
🧠 Model
AI Video OpenAI
sora-2
Now available in Sora 2, OpenAI's latest video generation model is more physically accurate, realistic, and easier to control than previous systems.
📊 Unpublished, based on diffusion+Transformer 💰 Billed based on video duration and resolution ( 📐 Support 4/8/12 seconds video, 720
🧠 Model
AI Video OpenAI
sora-2-all
Reverse of sora-2, supports 10s, 15s, both are 720p
📊 Same as sora-2 💰 Third-party channel prices are usually lower than 📐 Support 10/15 second video, 720p
🧠 Model
AI Video OpenAI
sora-2-pro
Sora 2 pro version, supports 4, 8, 12s, 720p and 1080p
📊 Enhanced Sora 2 architecture 💰 Billed by resolution and duration, 10 📐 Support 4/8/12 seconds video, 720
🧠 Model
AI Audio OpenAI
speech-02-hd
The MiniMax speech large model can intelligently predict the emotion, intonation and other information of the text based on the context, and generate supernatural, high-fidelity, personalized speech.
📊 MiniMax voice model HD version 💰 Charged by character, please consult us for specific price 📐 Support long text speech synthesis
🧠 Model
AI Audio OpenAI
speech-02-turbo
The MiniMax speech large model can intelligently predict the emotion, intonation and other information of the text based on the context, and generate supernatural, high-fidelity, personalized speech.
📊 MiniMax voice model Turbo version 💰 Pay per character, Turbo version 📐 Support long text speech synthesis
🧠 Model
AI Audio OpenAI
speech-2.6-hd
The MiniMax speech large model can intelligently predict the emotion, intonation and other information of the text based on the context, and generate supernatural, high-fidelity, personalized speech.
📊 MiniMax Voice 2.6 HD version 💰 Billed by character (third-party channels) 📐 Support long text speech synthesis
🧠 Model
AI Audio OpenAI
speech-2.6-turbo
The MiniMax speech large model can intelligently predict the emotion, intonation and other information of the text based on the context, and generate supernatural, high-fidelity, personalized speech.
📊 MiniMax Voice 2.6 Turbo version 💰 Billed by character (third-party channels) 📐 Support long text speech synthesis
🧠 Model
AI Audio OpenAI
speech-2.8-hd
The new generation voice HD model accurately restores the details of real tone and comprehensively improves the similarity of timbre.
📊 MiniMax Voice 2.8 HD version (latest) 💰 Billed by character (third-party channels) 📐 Support long text speech synthesis
🧠 Model
AI Audio OpenAI
speech-2.8-turbo
New generation voice Turbo model, extremely fast response, vivid and natural tone expression
📊 MiniMax Voice 2.8 Turbo version ( 💰 Billed by character (third-party channels) 📐 Support long text speech synthesis
🧠 Model
AI Writing OpenAI
text-ada-001
OpenAI’s most basic text generation model, the fastest but the weakest
📊 About 350M parameters 💰 $0.20/million tokens 📐 2049 tokens
🧠 Model
AI Writing OpenAI
text-babbage-001
OpenAI’s Babbage model for basic text processing tasks
📊 About 1.3B parameters 💰 $0.25/million tokens 📐 2049 tokens
🧠 Model
AI Writing OpenAI
text-curie-001
A medium-capacity model of the OpenAI GPT-3 series, which is relatively balanced in speed and capability.
📊 About 6.7B parameters 💰 $1/million tokens( 📐 2049 tokens
🧠 Model
AI Writing OpenAI
text-davinci-edit-001
Based on GPT-3 text editing model, text can be modified according to instructions
📊 Based on 175B parameters (GPT-3 Davin 💰 $10/million tokens 📐 2049 tokens
🧠 Model
AI Productivity OpenAI
text-embedding-3-large
OpenAI third-generation embedding model, more powerful than the small version, suitable for tasks requiring the highest performance
📊 Based on large embedded architecture, native 3072 dimensions (can be reduced to 💰 $0.13/million tokens 📐 Maximum 8191 tokens at a time
🧠 Model
AI Productivity OpenAI
text-embedding-3-small
OpenAI third-generation embedding model, faster and cheaper than second-generation models, suitable for most tasks
📊 Lightweight embedded architecture, default 1536 dimensions 💰 $0.02/million tokens 📐 Maximum 8191 tokens at a time
🧠 Model
AI Productivity OpenAI
text-embedding-ada-002
Ada-based text embedding model optimized for various NLP tasks.
📊 Based on Ada architecture, 1536-dimensional embedding 💰 $0.10/million tokens 📐 Maximum 8191 tokens at a time
🧠 Model
AI Productivity OpenAI
text-embedding-v1
text-embedding-v1 is an efficient text vectorization basic model suitable for building semantic search, text clustering and recommendation systems.
📊 First generation embedded architecture 💰 $0.05/million tokens 📐 Maximum 2048 tokens at a time
🧠 Model
AI Productivity OpenAI
text-moderation-latest
The latest text moderation model for content security checks
📊 Lightweight audit model 💰 Free 📐 Maximum 2000 characters at a time
🧠 Model
AI ProductivityAI Writing OpenAI
text-moderation-stable
Stable version text review model, providing reliable content review
📊 Stable version review model 💰 Free 📐 Maximum 2000 characters at a time
🧠 Model
AI Audio OpenAI
tts-1-1106
<div>Text-to-speech model TTS supports setting timbre. The standard tts-1 model provides the lowest latency, but the quality is lower than the tts-1-hd model. </div>
📊 Version 1106 of tts-1 updated 💰 $15/million characters 📐 Maximum 4096 characters at a time
🧠 Model
AI Audio OpenAI
tts-1
<div>Text-to-speech model TTS supports setting timbre. The standard tts-1 model provides the lowest latency, but the quality is lower than the tts-1-hd model. </div>
📊 Based on TTS architecture, 6 preset sounds 💰 $15/million characters 📐 Maximum 4096 characters at a time
🧠 Model
AI Audio OpenAI
tts-1-hd-1106
Text-to-speech model TTS, supports setting timbre
📊 Version 1106 of tts-1-hd updated 💰 $30/million characters 📐 Maximum 4096 characters at a time
🧠 Model
AI Audio OpenAI
tts-1-hd
Text-to-speech model TTS, supports setting timbre
📊 High-fidelity TTS architecture 💰 $30/million characters 📐 Maximum 4096 characters at a time
🧠 Model
AI Audio OpenAI
whisper-1
Whisper can transcribe speech to text and translate multiple languages into English
📊 About 1.5B parameters 💰 $0.006/minute audio 📐 Maximum audio file size 25MB
🧠 Model
AI VideoAI Audio PixVerse
pixverse-image-template
Different from video-based templates, this is a set of AI template gameplay with pictures as the core - just upload your photos to quickly generate an immersive AI content experience.
💰 Free quota + paid templates
🧠 Model
AI Video PixVerse
pixverse-lipsync
The Lipsync interface is specially designed to solve lip-syncing problems in videos.
💰 Billed based on generation time
🧠 Model
AI VideoAI Audio PixVerse
pixverse-mask-selection
Pixverse Mask Selection (or Swap Mask Generation) is an AI intelligent mask generation tool used by the PixVerse platform in the Modify function. It is specially used to accurately select the required mask in the video.
💰 Included in video editing fee
🧠 Model
AI VideoAI Audio PixVerse
pixverse-mimic
The Mimic function supports migrating the actions in the reference video to the target person picture to achieve action imitation and reconstruction.
💰 Billed based on generation time
🧠 Model
AI Video PixVerse
pixverse-modify
The video editing (Modify) function supports editing any part of the existing video, including adding, replacing, deleting, modifying elements, or changing the overall style.
💰 Billed based on the number of edits or duration
🧠 Model
AI VideoAI Audio PixVerse
pixverse-multi-transition
The multi-frame (Multi-transition) function allows you to generate a 1-30 second video by providing 2-7 key frames, ensuring a consistent style and smooth lens connection.
💰 Billed based on generation time and number of frames
🧠 Model
AI VideoAI Audio PixVerse
pixverse-restyle
The Restyle feature instantly transforms your video clips into a new visual style—such as 3D, anime, cinematic, or augmented reality.
💰 Billed by video duration
🧠 Model
AI VideoAI Audio PixVerse
pixverse-sound-effect
The sound effect generation (background music) interface can intelligently generate sound effects and environmental sounds that adapt to the video screen, and can also be input through text commands.
💰 Billed based on the number of sound effects generated
🧠 Model
AI VideoAI Audio PixVerse
pixverse-swap
SWAP is a powerful video editing capability.
💰 Charged based on the number of substitutions
🧠 Model
AI Video PixVerse
pixverse-video
It belongs to the cutting-edge AI video foundation model. Its core capabilities include Text-to-Video (text to generate video) and Image-to-Video (picture to generate video). It supports uploading pictures or
💰 Billed based on generation time and resolution
🧠 Model
AI Writing xAI
grok-4
The Grok 4 model is capable of deep reasoning and has been trained on xAI’s Colossus supercomputer, promising greater logical reasoning and text generation capabilities.
📊 About 1T+ (MoE architecture, activation parameters about 100B
🧠 Model
AI Writing xAI
grok-3
Grok 3 is an artificial intelligence model provided by xai.
📊 About 314B (MoE hybrid expert architecture, activation parameters
🧠 Model
AI Writing xAI
grok-3-deepsearch
grok-3-deepsearch deep network search model
📊 About 314B (based on Grok 3 MoE architecture
🧠 Model
AI Image Gen xAI
grok-3-image
Use grok3 to generate images
📊 Unpublished (image generation module based on Grok 3)
🧠 Model
AI Writing xAI
grok-3-reasoning
grok-3-reasoning is a model for grok to enhance reasoning capabilities
📊 About 314B (based on Grok 3 MoE architecture
🧠 Model
AI Writing xAI
grok-3-mini
Grok 3 mini, which represents a new frontier in cost-effective reasoning capabilities.
📊 About 100B (MoE architecture lightweight version, activation parameters are about
🧠 Model
AI Writing xAI
grok-3-reasoner
grok-3-reasoning is a model for grok to enhance reasoning capabilities
📊 About 314B (based on Grok 3 MoE architecture
🧠 Model
AI Writing xAI
grok-4-1-fast-non-reasoning
grok-4-1-fast-non-reasoning is an AI model developed by xAI that is optimized for maximum speed when generating responses and executing agent tasks.
📊 About 1T+ (MoE architecture fast version, skipping the inference chain)
🧠 Model
AI Writing xAI
grok-4-1-fast-reasoning
Grok 4.1 Fast is a top tool calling model under xAI, with 2 million context windows.
📊 About 1T+ (MoE architecture, 2M context window)
🧠 Model
AI Writing xAI
grok-4.1
Grok 4.1 excels in creative, emotional and collaborative interactions, is able to capture subtle intentions more keenly, has a more engaging conversation experience, and maintains a high degree of consistency in personality traits, while fully inheriting the sharp intelligent performance and reliable performance of its predecessor.
📊 About 1T+ (MoE architecture, enhanced emotional interaction)
🧠 Model
AI CodingAI Writing xAI
grok-4-20-non-reasoning
Grok 4.20 is xai's latest flagship model, with industry-leading speed and proxy tool calling capabilities.
📊 About 1T+ (MoE architecture flagship version, hallucination rate optimization)
🧠 Model
AI CodingAI Writing xAI
grok-4-20-reasoning
Grok 4.20 is xai's latest flagship model, with industry-leading speed and proxy tool calling capabilities.
📊 About 1T+ (MoE architecture flagship version + inference enhancement)
🧠 Model
AI Writing xAI
grok-4-fast
grok-4-fast, the latest advancement in cost-effective inference models.
📊 About 1T+ (MoE architecture fast version)
🧠 Model
AI Writing xAI
grok-4-fast-reasoning
📊 About 1T+ (MoE architecture fast inference version)
🧠 Model
AI CodingAI Writing xAI
grok-4-fast-non-reasoning
Grok 4 Fast (Non-Reasoning) is an artificial intelligence model provided by xai.
📊 About 1T+ (MoE architecture fast version, no inference chain)
🧠 Model
AI Writing xAI
grok-4.2
Grok-4.2 is an AI productivity system with trillions of parameters, 16-Agent cluster collaboration, supports real-time data stream processing, and can self-evolve through high-frequency weekly updates. It is specially optimized for complex decision-making, business automation, and personalized cognitive enhancement.
📊 Trillions of parameters (16-Agent cluster collaboration architecture
🧠 Model
AI Productivity xAI
grok-4.2-fast
Grok-4.2-Fast is a fast and economical version of the Grok 4 series launched by xAI
📊 Fast version of trillion-level parameters (simplified reasoning chain)
🧠 Model
AI CodingAI Writing xAI
grok-4.3
Our most advanced flagship model, leading the industry in non-hallucination rates, agent tool calling and command following capabilities.
📊 Trillion-level parameters (the latest flagship version of MoE architecture)
🧠 Model
AI Image Gen xAI
grok-imagine-image
Grok-Imagine-Image is a multi-modal AI model launched by the X platform, which can generate high-quality images based on text descriptions.
📊 Undisclosed (independent image generation model)
🧠 Model
AI Image Gen xAI
grok-imagine-image-pro
Grok-Imagine-Image-Pro is an upgraded version of the multi-modal AI model of the X platform. It achieves higher-precision image generation and multiple rounds of optimization through stronger understanding and generation details.
📊 Unpublished (Imagine Image enhanced version
🧠 Model
AI Video xAI
grok-video-3-10s
grok's latest video model, 10s, supports audio and video simultaneous playback
📊 Unpublished (enhanced version of video generation model)
🧠 Model
AI Writing other
MAI-DS-R1
MAI-DS-R1 is a DeepSeek-R1 inference model that was post-trained by the Microsoft AI team to fill information gaps in previous versions of the model and improve its risk signature while maintaining R1 inference capabilities.
📊 671B (based on DeepSeek-R1) 💰 With DeepSeek-R1 📐 128K tokens
🧠 Model
AI CodingAI Writing Bytes (bean bags)
doubao-seed-1-8-251228
doubao-seed-1.8 has stronger multi-modal understanding capabilities and Agent capabilities. The model can perform better in many complex tasks in the real world, further helping enterprises create value.
📊 Undisclosed (MoE architecture upgraded version, estimated total parameter amount is approximately
🧠 Model
AI Writing Bytes (bean bags)
doubao-seed-1-6-251015
Supports adjusting the length of thinking, that is, supporting the reasoning_effort field in the API, which is divided into four modes: minimal, low, medium, and high.
📊 Undisclosed (MoE architecture, estimated total number of parameters is about 200
🧠 Model
AI Writing Bytes (bean bags)
doubao-seed-1-6-flash-250828
The extremely fast response version is the extremely fast version in the Doubao Big Model 1.6 series. It has deep thinking and multi-modal understanding capabilities and supports 256K context.
📊 Undisclosed (MoE architecture lightweight version, estimated activation parameters are approximately
🧠 Model
AI Writing Bytes (bean bags)
doubao-seed-1-6-vision-250815
Doubao-Seed-1.6-vision visual deep thinking model shows stronger general multi-modal understanding and reasoning capabilities in scenarios such as education, image review, inspection and security, and AI search question and answer.
📊 Undisclosed (MoE architecture, supports visual encoder, estimated
🧠 Model
AI CodingAI Writing Bytes (bean bags)
doubao-seed-2-0-code-preview-260215
Doubao-Seed-2.0-Code is optimized for enterprise-level programming needs. Based on the excellent Agent and VLM capabilities of Seed 2.0, it has especially enhanced coding capabilities. It not only has outstanding front-end capabilities, but also meets the common multi-language coding needs of enterprises.
📊 Undisclosed (MoE architecture flagship level, estimated total parameter amount is approximately
🧠 Model
AI CodingAI Writing Bytes (bean bags)
doubao-seed-2-0-lite-260428
The first full-modal understanding model in the Doubao model family supports native unified understanding of video, images, audio, and text, while upgrading Agent, Coding, and GUI capabilities.
📊 Undisclosed (Full-modal version of MoE architecture, estimated total number of parameters
🧠 Model
AI Writing Bytes (bean bags)
doubao-seed-2-0-lite-260215
Doubao-Seed-2.0-lite is a balanced model that balances performance and cost for high-frequency enterprise scenarios. Its comprehensive capabilities surpass the previous generation Doubao-Seed-1.8.
📊 Undisclosed (balanced version of MoE architecture, estimated total parameter amount is approximately
🧠 Model
AI CodingAI Writing Bytes (bean bags)
doubao-seed-2-0-mini-260428
Doubao large model family full-modal understanding model, shorter thinking length, higher tokens efficiency
📊 Undisclosed (MoE architecture lightweight version, estimated activation parameters are approximately
🧠 Model
AI Writing Bytes (bean bags)
doubao-seed-2-0-mini-260215
Doubao-Seed-2.0-mini provides the ultimate model inference speed for low-latency, high-concurrency and cost-sensitive scenarios.
📊 Undisclosed (MoE architecture lightweight version, estimated activation parameters are approximately
🧠 Model
AI Video Bytes (bean bags)
doubao-seedance-1-0-lite-i2v-250428
Seedance 1.0 lite has powerful semantic understanding and command following capabilities, and can finely control the character's appearance, clothing style, and facial expressions. It also shows absolute advantages in multi-subject action analysis, embedded text response, degree adverbs, and lens switching response.
📊 Undisclosed (Video generation DiT architecture, estimated parameters are approximately
🧠 Model
AI Video Bytes (bean bags)
doubao-seedance-1-0-lite-t2v-250428
Seedance 1.0 lite has powerful semantic understanding and command following capabilities, and can finely control the character's appearance, clothing style, and facial expressions. It also shows absolute advantages in multi-subject action analysis, embedded text response, degree adverbs, and lens switching response.
📊 Undisclosed (Video generation DiT architecture, estimated parameters are approximately
🧠 Model
AI Video Bytes (bean bags)
doubao-seedance-1-0-pro-fast-251015
Seedance 1.0 pro fast is a comprehensive model with bottom-end price and top-end performance, achieving an excellent balance between video generation quality, speed, and price.
📊 Undisclosed (Video generation DiT architecture accelerated version, estimated parameters
🧠 Model
AI Video Bytes (bean bags)
doubao-seedance-1-0-pro-250528
Seedance 1.0 is the latest basic video generation model launched by the ByteDance Doubao model team.
📊 Undisclosed (Video generates an upgraded version of DiT architecture, estimated parameters
🧠 Model
AI Video Bytes (bean bags)
doubao-seedance-1-5-pro-251215
Supports audio and video generated by Vincent, audio and video generated by first frame pictures, audio and video generated by first and last frames of pictures
📊 Undisclosed (Video generation DiT architecture, estimated parameters are approximately
🧠 Model
AI Image Gen Bytes (bean bags)
doubao-seedream-3-0-t2i-250415
Seedream 3.0 is a basic model for Chinese and English bilingual image generation that supports native high resolution. Its comprehensive capabilities are comparable to GPT-4o, ranking first in the world.
📊 Undisclosed (image generation DiT architecture, estimated parameters are approximately
🧠 Model
AI Image Gen Bytes (bean bags)
doubao-seedream-5-0-260128
Doubao-Seedream-5.0-lite is the latest image creation model released by ByteDance.
📊 Undisclosed (the latest generation of DiT architecture, the estimated number of parameters is about 1
🧠 Model
AI Image Gen Bytes (bean bags)
doubao-seedream-4-5-251128
Seedream 4.5 natively supports text, single image and multi-image input, enabling diverse gameplay methods such as multi-image fusion creation, image editing, and group image generation based on subject consistency, making image creation more free and controllable.
📊 Undisclosed (multi-modal DiT architecture, estimated number of parameters is about 1
🧠 Model
AI Video Kling
kling-advanced-custom-elements
Kling's official subject-related services have been upgraded to a new version (supports 3s-8s)
💰 Billed based on the number of subject creations
🧠 Model
AI Video Kling
kling-advanced-lip-sync
Advanced lip-syncing model, including face recognition and lip-syncing functions.
💰 Face recognition press + lip sync press 5
🧠 Model
AI Audio Kling
kling-audio
AI audio generation model supports Wensheng sound effects, video dubbing and other functions.
💰 Billed based on the number of sound effects generated
🧠 Model
AI Image Gen Kling
kling-image
The AI image generation model supports functions such as text-generated images, image-generated images, multi-image reference images, and image editing.
💰 Billed based on the number of photos generated and version
🧠 Model
AI Video Kling
kling-avatar-image2video
Digital human video generation model that converts static human pictures into dynamic talking videos.
💰 Billed by second, std/pro
🧠 Model
AI VideoAI Audio Kling
kling-custom-elements
kling's creation subject (supports 3s-8s)
💰 Billed based on the number of subject creations
🧠 Model
AI Audio Kling
kling-custom-voices
The custom timbre creation model allows users to create their own personalized voice timbre by uploading audio files or using audio from existing videos.
💰 Billed based on tone creation + number of calls
🧠 Model
AI Video Kling
kling-effects
The AI special effects center provides 136+ creative video special effects, including holiday themes, dynamic effects, style conversion, etc.
💰 1-7 yuan/time, depending on the type of special effects
🧠 Model
AI Image Gen Kling
kling-image-recognize
The intelligent image recognition model can perform content analysis and recognition on uploaded images, and extract key information, objects, scenes and other elements in the image.
💰 Billed based on the number of recognitions
🧠 Model
AI Video Kling
kling-motion-control
Kling 2.6 Motion Control AI motion transfer technology - Real motion, precise control Transfer precise motions in reference videos to static character images.
💰 Billed based on generation time
🧠 Model
AI Video Kling
kling-multi-elements
The multi-element synthesis model supports the intelligent fusion of multiple elements in videos to achieve complex creative video production.
💰 Billed based on synthesis complexity
🧠 Model
AI Image Gen Kling
kling-omni-image
Kling-omni-image A complete creative suite designed for video and image creation KlingOmni combines video and image generation capabilities in a complete creative suite designed for professionals who need consistent high-quality output every time.
💰 Professional suite pricing, based on functional combinations
🧠 Model
AI Video Kling
kling-omni-video
Keling O1, the product adheres to the Muti-modal language (MVL) concept, uses natural language as the semantic skeleton, and cooperates with multi-modal descriptions such as videos, pictures, and subjects to accurately understand your intentions and operate more intuitively.
💰 Flagship-level pricing
🧠 Model
AI Video Kling
kling-video-extend
The video extension model can intelligently extend existing videos to maintain picture coherence and style consistency, and is suitable for scenarios where video duration needs to be extended.
💰 Billed by extended seconds
🧠 Model
AI Productivity Wisdom Source (BAAI)
BAAI/bge-reranker-v2-m3
BAAI/bge-reranker-v2-m3 is a lightweight multi-language reranking model.
📊 ~568M (based on XLMRoberta) 💰 Open source and free 📐 8192 tokens
🧠 Model
AI Productivity Wisdom Source (BAAI)
Pro/BAAI/bge-reranker-v2-m3
📊 ~568M (same as base version) 💰 Paid API, billed based on call volume 📐 8192 tokens
🧠 Model
AI Writing GLM (GLM)
glm-3-turbo
Zhipu AI general large model is suitable for scenarios that require high knowledge, reasoning ability, and creativity, such as advertising copywriting, novel writing, knowledge writing, code generation, etc.
📊 About 130B (Dense architecture, GLM-3 series
🧠 Model
AI CodingAI Writing GLM (GLM)
glm-5
GLM-5 is the new generation flagship base model of Zhipu. It is built for Agentic Engineering and can provide reliable productivity in complex system engineering and long-range Agent missions.
📊 About 1T + total parameters (MoE architecture, Agentic
🧠 Model
AI Writing GLM (GLM)
glm-4
Zhipu AI's general large model provides more powerful question answering and text generation capabilities.
📊 About 130B (Dense architecture, early GLM-4
🧠 Model
AI Writing GLM (GLM)
glm-4-air
High cost performance: the most balanced model between reasoning power and price, context 128K, maximum output 4K
📊 About 30B (Dense architecture lightweight version)
🧠 Model
AI Writing GLM (GLM)
glm-4-airx
Extremely fast inference: ultra-fast inference speed and powerful inference effect, 8K context, maximum output 4K
📊 About 30B (Dense architecture speed version)
🧠 Model
AI Writing GLM (GLM)
glm-4-flash
Free call: Zhipu AI’s first free API, calling large models at zero cost.
📊 About 14B (Dense architecture lightweight version)
🧠 Model
AI Writing GLM (GLM)
glm-4-long
Ultra-long input: specially designed for processing ultra-long text and memory-based tasks, context 1M, maximum output 4K
📊 About 130B (Dense architecture, 1M context optimization
🧠 Model
AI CodingAI Writing GLM (GLM)
glm-4.5
GLM-4.5 and GLM-4.5-Air are the latest flagship model series, basic models designed for intelligent agent applications.
📊 355B general parameters/32B activation (MoE architecture, A
🧠 Model
AI Writing GLM (GLM)
glm-4.5-flash
GLM-4.5 and GLM-4.5-Air are the latest flagship model series, basic models designed for intelligent agent applications.
📊 106B general parameters/12B activation (MoE architecture free
🧠 Model
AI Writing GLM (GLM)
glm-4.5-air
GLM-4.5 and GLM-4.5-Air are the latest flagship model series, basic models designed for intelligent agent applications.
📊 106B general parameters/12B activation (MoE architecture lightweight
🧠 Model
AI Writing GLM (GLM)
glm-4.5-airx
GLM-4.5 and GLM-4.5-Air are the latest flagship model series, basic models designed for intelligent agent applications.
📊 106B general parameters/12B activation (MoE architecture extremely fast
🧠 Model
AI Writing GLM (GLM)
glm-4.5-x
GLM-4.5 and GLM-4.5-Air are the latest flagship model series, basic models designed for intelligent agent applications.
📊 355B general parameters/32B activation (MoE architecture enhancement
🧠 Model
AI Writing GLM (GLM)
glm-4.6
GLM-4.6 is an artificial intelligence model provided by zhipuai-coding-plan.
📊 About 400B+ total parameters (MoE architecture, Codin
🧠 Model
AI Writing GLM (GLM)
glm-4.6v
The GLM-4.6V series is an important iteration of the GLM series in the multi-modal direction. It increases the context window during training to 128k tokens, achieves the same parameter scale SOTA in visual understanding accuracy, and for the first time integrates Function in the model architecture.
📊 About 400B + total parameters (MoE architecture + visual encoder
🧠 Model
AI Writing GLM (GLM)
glm-4.7
GLM-4.7 is the latest flagship model of GLM. GLM-4.7 strengthens coding capabilities, long-term task planning and tool collaboration for agentic coding scenarios, and has achieved leading performance among open source models in the current lists of multiple public benchmarks.
📊 About 400B+ total parameters (MoE architecture, Agent
🧠 Model
AI Writing GLM (GLM)
glm-5.1
GLM-5.1 is a model designed for Long Horizon Tasks.
📊 About 1T + total parameters (MoE architecture, ultra-long-range mission optimization
🧠 Model
AI Writing Dark Side of the Moon (Kimi)
kimi-k2-250905
Kimi-K2 is a MoE architecture basic model launched by Moonshot AI with super strong code and agent capabilities. It has a total parameter of 1T and an activation parameter of 32B.
📊 Total parameters 1T (trillion), activation parameters 32B, 💰 $0.60/M input, 📐 128K tokens (Part No.
🧠 Model
AI Writing Dark Side of the Moon (Kimi)
kimi-k2-0905
Kimi K2 is a breakthrough hybrid expert model designed for superior performance in cutting-edge knowledge, reasoning and programming tasks.
📊 Total parameters 1T (trillion), activation parameters 32B, 💰 $0.60/M input, 📐 128K tokens
🧠 Model
AI Writing Dark Side of the Moon (Kimi)
kimi-k2-0711-preview
kimi-k2 is a MoE architecture basic model with strong code and agent capabilities.
📊 Total parameters 1T (trillion), activation parameters 32B, 💰 Same price as kimi-k2 📐 128K tokens
🧠 Model
AI CodingAI Writing Dark Side of the Moon (Kimi)
kimi-k2.5
kimi-k2.5 is Kimi's most versatile model to date, with a native multi-modal architecture design that supports both visual and text input, thinking and non-thinking modes, dialogue and agent tasks.
📊 Total parameters 1T (trillion), activation parameters 32B, 💰 $0.60/M input, 📐 256K tokens (262
🧠 Model
AI Writing Dark Side of the Moon (Kimi)
kimi-k2.6
Kimi K2.6 is Kimi's latest and most intelligent model. It has stronger and more stable long-range code writing capabilities, significantly improved command following and self-correction capabilities, and supports text, image and video input, thinking and non-thinking modes, dialogue and Agent tasks.
📊 Total parameters 1T (trillion), activation parameters 32B, 💰 $0.95/M input, 📐 256K tokens (262
🧠 Model
AI Writing Dark Side of the Moon (Kimi)
moonshot-v1-128k
Moonshot AI has launched a language model with hundreds of billions of parameters, which has excellent semantic understanding, instruction following and text generation capabilities.
📊 About 100B (hundred billion parameters), non-MoE thick 💰 $2.00/M input, 📐 128K tokens (131
🧠 Model
AI Writing Dark Side of the Moon (Kimi)
moonshot-v1-32k
Moonshot AI has launched a language model with hundreds of billions of parameters, which has excellent semantic understanding, instruction following and text generation capabilities.
📊 About 100B (hundred billion parameters), non-MoE thick 💰 $1.00/M input, 📐 32K tokens (32,7
🧠 Model
AI Writing Dark Side of the Moon (Kimi)
moonshot-v1-8k
Moonshot AI has launched a language model with hundreds of billions of parameters, which has excellent semantic understanding, instruction following and text generation capabilities.
📊 About 100B (hundred billion parameters), non-MoE thick 💰 $0.20/M input, 📐 8K tokens (8,192
🧠 Model
AI Audio Vidu
audio1.0
This interface is used to input text to generate corresponding reading audio, and can control the speaking speed, volume, and emotion of the reading.
💰 Billed by character
🧠 Model
AI Audio Vidu
vidu-tts
This interface is used to input text to generate corresponding reading audio, and can control the speaking speed, volume, and emotion of the reading.
💰 Billed by character
🧠 Model
AI Video Vidu
vidu2.0
vidu2.0 has a fast generation speed, supports 360p, 720p, 1080p, supports picture-based videos, reference videos, first and last frames
💰 Billed based on generation seconds + resolution
🧠 Model
AI Video Vidu
viduq1
The video generated by viduq1 has a clear picture, supports 1080p, supports Vincent videos, Picture videos, Reference videos, first and last frames
💰 Billed by seconds + resolution gradient
🧠 Model
AI Video Vidu
viduq1-classic
viduq1-classic, rich and stable camera operation, supports graphic videos and first and last frames, and supports 1080p
💰 Same pricing as Q1 series
🧠 Model
AI Image GenAI Video Vidu
viduq2
viduq2 supports Vincent videos and reference videos, with good dynamic effects and rich generated details. It also supports Vincent pictures, picture editing, and reference pictures.
💰 Mid-range pricing, billed by the second
🧠 Model
AI Video Vidu
viduq2-pro
viduq2-pro supports reference video, video editing, 540p, 720p, 1080p
💰 The professional version is priced higher than the basic version
🧠 Model
AI Video Vidu
viduq3
viduq3 has strong picture quality, supports intelligent lens cutting, good dynamic effects, and the strongest balance. It supports 540p, 720p, and 1080p.
💰 Flagship-level pricing
🧠 Model
AI Video Vidu
viduq3-pro
viduq3-pro supports audio and video simultaneous playback, supports intelligent mirror cutting, and supports 720p, 1080p
💰 Top flagship pricing
🧠 Model
AI Video Vidu
viduq3-mix
viduq3-mix supports intelligent lens cutting, and the consistency of multiple cameras is better. Supports 720p, 1080p
💰 Professional grade pricing
🧠 Model
AI Writing Baidu (Wenxin)
ERNIE-3.5-8K
ERNIE 3.5 is Baidu's self-developed flagship large-scale language model, covering a massive amount of Chinese and English corpora. It has powerful general capabilities and can meet the requirements of most dialogue Q&A, creative generation, and plug-in application scenarios; ensuring the timeliness of Q&A information.
📊 About 260 billion (260B) parameters 💰 ¥0.0008/thousand tok 📐 8K tokens
🧠 Model
AI Writing Baidu (Wenxin)
ERNIE-4.0-8K
The most powerful language model in Baidu Wenxin series, its understanding, generation, logic, and memory capabilities have reached the top level in the industry.
📊 About 1 trillion (1T) parameters 💰 ¥0.004/thousand tokens 📐 8K tokens
🧠 Model
AI Writing Baidu (Wenxin)
ERNIE-Character-8K
Baidu's self-developed vertical scene large language model is suitable for application scenarios such as game NPCs, customer service dialogues, and dialogue role-playing. The character style is more distinctive and consistent, the ability to follow instructions is stronger, and the reasoning performance is better.
📊 Undisclosed (based on the ERNIE series, presumably dozens of 💰 ¥0.004/thousand tokens 📐 8K tokens
🧠 Model
AI Writing Baidu (Wenxin)
ERNIE-Functions-8K
ERNIE Functions 8K is a powerful AI model developed by Baidu that is designed for advanced natural language processing tasks and can be integrated with external functions to enhance its functionality.
📊 Unpublished (lightweight domain model, based on ERNIE 💰 ¥0.004/thousand tokens 📐 8K tokens
🧠 Model
AI Writing Baidu (Wenxin)
ERNIE-Lite-8K
Baidu's self-developed lightweight large language model takes into account excellent model effects and reasoning performance, and is suitable for low computing power AI accelerator card reasoning.
📊 Undisclosed (lightweight, guessed number B~tens of B parameters) 💰 Free (original price ¥0.003 📐 8K tokens
🧠 Model
AI Writing Baidu (Wenxin)
ERNIE-Speed-128K
Baidu's self-developed lightweight large language model takes into account excellent model effects and reasoning performance, and is suitable for low computing power AI accelerator card reasoning.
📊 Unpublished (lightweight, parameter size is larger than ERNIE 💰 Free (original price ¥0.016 📐 128K tokens
🧠 Model
AI Writing Baidu (Wenxin)
ERNIE-Speed-8K
Baidu's self-developed lightweight large language model takes into account excellent model effects and reasoning performance, and is suitable for low computing power AI accelerator card reasoning.
📊 Unpublished (lightweight, parameter size is larger than ERNIE 💰 Free (original price ¥0.004 📐 8K tokens
🧠 Model
AI Writing Baidu (Wenxin)
ERNIE-Tiny-8K
ERNIE-Tiny-8K is a lightweight Chinese pre-trained language model developed by the Baidu team.
📊 Undisclosed (minimum model of ERNIE series, estimated number 💰 Free (original price is about ¥0.00 📐 8K tokens
🧠 Model
AI Productivity Baidu (Wenxin)
Embedding-V1
Embedding-V1 is a text representation model based on Baidu Wenxin large model technology. It converts text into a vector form represented by numerical values and can be used in text retrieval, information recommendation, knowledge mining and other scenarios.
📊 Unpublished (dedicated Emb based on Wenxin large model technology 💰 ¥0.002/thousand tokens 📐 The upper limit of single input is about 512 tok
🧠 Model
AI Writing iFlytek
SparkDesk-v1.1
It is a new generation cognitive intelligence large model launched by iFlytek. It has cross-domain knowledge and language understanding capabilities. It can understand and perform tasks based on natural dialogue, and provides language understanding, knowledge Q&A, logical reasoning, mathematical problem solving, and code understanding.
📊 About 170B (early version) 💰 Already obsolete, it is recommended to upgrade to a higher version 📐 8K tokens
🧠 Model
AI Writing iFlytek
SparkDesk-v2.1
It is a new generation cognitive intelligence large model launched by iFlytek. It has cross-domain knowledge and language understanding capabilities. It can understand and perform tasks based on natural dialogue, and provides language understanding, knowledge Q&A, logical reasoning, mathematical problem solving, and code understanding.
📊 About 170B-200B 💰 Medium pricing, general enterprise scenario 📐 16K tokens
🧠 Model
AI Writing iFlytek
SparkDesk-v3.1
It is a new generation cognitive intelligence large model launched by iFlytek. It has cross-domain knowledge and language understanding capabilities. It can understand and perform tasks based on natural dialogue, and provides language understanding, knowledge Q&A, logical reasoning, mathematical problem solving, and code understanding.
📊 About 200B-300B 💰 Mid- to high-end pricing 📐 32K tokens
🧠 Model
AI Writing iFlytek
SparkDesk-v3.5
It is a new generation cognitive intelligence large model launched by iFlytek. It has cross-domain knowledge and language understanding capabilities. It can understand and perform tasks based on natural dialogue, and provides language understanding, knowledge Q&A, logical reasoning, mathematical problem solving, and code understanding.
📊 About 300B+ 💰 High-end pricing, enterprise-level solutions 📐 128K tokens
🧠 Model
AI Productivity NetEase Youdao
netease-youdao/bce-reranker-base_v1
bce-reranker-base_v1 is a bilingual and cross-language reranking model developed by NetEase Youdao, supporting Chinese, English, Japanese and Korean.
📊 ~279M (based on XLMRoberta-b 💰 Open source and free for commercial use 📐 512 tokens
🧠 Model
AI Productivity Ali (Qwen)
Qwen/Qwen3-Reranker-0.6B
Qwen3-Reranker-0.6B is a text reranking model from the Qwen3 series.
📊 0.6B 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Video Ali (Qwen)
happyhorse-1.0-i2v
HappyHorse-1.0-I2V supports Tusheng videos, has highly restored dynamic picture generation capabilities, can accurately understand text semantics, and outputs smooth, natural, high-quality videos with rich details.
📊15B 💰 Alibaba Bailian platform is billed based on the call volume 📐 N/A (video generation)
🧠 Model
AI Video Ali (Qwen)
happyhorse-1.0-r2v
HappyHorse-1.0-R2V supports reference raw videos, more stable subject and scene reference, and supports up to 9 picture references, which can accurately maintain creative intentions and achieve stronger performance capabilities.
📊15B 💰 Alibaba Bailian platform is billed based on the call volume 📐 N/A (video generation)
🧠 Model
AI Video Ali (Qwen)
happyhorse-1.0-t2v
HappyHorse-1.0-T2V supports Wensheng video, has the ability to generate highly restored dynamic pictures, can accurately understand text semantics, and output smooth, natural, and high-quality videos with rich details.
📊15B 💰 Alibaba Bailian platform is billed based on the call volume 📐 N/A (video generation)
🧠 Model
AI Video Ali (Qwen)
happyhorse-1.0-video-edit
HappyHorse-1.0-Video-Edit supports video editing and natural language commands to edit videos. It can refer to up to 5 pictures to edit video elements locally or globally. It can accurately reproduce the dynamic process of the video and achieve stronger performance capabilities.
📊15B 💰 Alibaba Bailian platform is billed based on the call volume 📐 N/A (Video Editing)
🧠 Model
AI Writing Ali (Qwen)
qvq-max
Tongyi Qianwen QVQ visual reasoning model supports visual input and thinking chain output, showing stronger capabilities in mathematics, programming, visual analysis, creation and general tasks.
📊 MoE ~100B+ (general parameters) 💰 Bailian API payment, visual reasoning 📐 128K
🧠 Model
AI Image Gen Ali (Qwen)
qwen-image-edit-2509
qwen-image-edit-2509 is the latest painting model
📊 7B 💰 Alibaba Bailian platform charges per piece, any time 📐 N/A (image editing)
🧠 Model
AI Writing Ali (Qwen)
qwen-max-1201
Tongyi Qianwen's billion-level ultra-large-scale language model supports input in different languages such as Chinese and English.
📊 MoE ~100B+ (general parameters) 💰 Bailian API ¥2.4/M 📐 32K
🧠 Model
AI Image Gen Ali (Qwen)
qwen-image-2.0
The Qwen-Image-2.0 series accelerated version model realizes the integration of image generation and image editing; it has more professional text rendering 1k token command support capabilities, more delicate real texture, delicate depiction of realistic scenes, and stronger semantic compliance capabilities.
📊 7B 💰 Alibaba Bailian platform charges per piece, open 📐 N/A (image generation)
🧠 Model
AI Writing Ali (Qwen)
qwen-long
Tongyi Qianwen long text model
📊 MoE ~100B+ (general parameters) 💰 Bailian API ¥0.5/M 📐 10M
🧠 Model
AI Writing Ali (Qwen)
qwen-max-longcontext
Tongyi Qianwen's billion-level ultra-large-scale language model supports input in different languages such as Chinese and English.
📊 MoE ~100B+ (general parameters) 💰 Bailian API long context ladder 📐 1M
🧠 Model
AI Productivity Ali (Qwen)
qwen-mt-plus
Based on the comprehensively upgraded flagship translation model of Qwen3, it supports translation into 92 languages. The model performance and translation effect are fully upgraded, providing more stable terminology customization, format reduction, and domain prompt capabilities, making the translation more accurate and natural.
📊 ~14B (estimated) 💰 Bailian API ¥0.9/M 📐 32K
🧠 Model
AI Productivity Ali (Qwen)
qwen-mt-turbo
Based on the fully upgraded lightweight text translation model of Qwen3, it supports 92 language translations. The model performance and translation effect are fully upgraded, providing more stable terminology customization, format reduction, and domain prompt capabilities, making the translation more accurate and natural.
📊 ~7B (estimated) 💰 Bailian API ¥0.35/ 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen-omni-turbo
The stable version of Tongyi Qianwen Omni's new multi-modal understanding and generation large model supports text, image, voice and video input, and outputs text, providing 4 natural dialogue sounds.
📊 ~7B (estimated) 💰 Bailian API ¥0.8/M 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen-vl-max
That is, the Tongyi Qianwen ultra-large-scale visual language model.
📊 MoE ~100B+ (general parameters) 💰 Bailian API ¥3/M t 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen-vl-plus
The Tongyi Qianwen VL-PLUS model is relatively balanced in terms of effect and cost. If you are not sure about using a certain model for the time being, you can try the Tongyi Qianwen VL-PLUS model first.
📊 ~70B (estimated) 💰 Bailian API ¥1.5/M 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen1.5-110b-chat
Tongyi Qianwen 1.5 generation 110B parameter dialogue model
📊 110B (Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen1.5-14b-chat
Tongyi Qianwen 1.5 generation 14B parameter dialogue model
📊 14B (Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen1.5-32b-chat
Tongyi Qianwen 1.5 is an open source chat model with 32B scale parameters aligned with human instructions.
📊 32B (Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen1.5-7b-chat
Tongyi Qianwen 1.5 generation 7B parameter dialogue model
📊 7B(Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen2-1.5b-instruct
Tongyi Qianwen 2 is a 1.5B scale model that is open source to the outside world.
📊 1.5B (Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen2-57b-a14b-instruct
Tongyi Qianwen 2 is an open source MOE model with 57B scale and 14B activation parameters.
📊 MoE 57B total/14B activated 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen2-7b-instruct
Tongyi Qianwen 2nd generation 7B parameter instruction model
📊 7B(Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen2-vl-72b-instruct
Tongyi Qianwen 2-VL-72B, extends the context to 32k, enhances image understanding capabilities, and can better recognize multi-language and handwriting in pictures.
📊 72B (Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen2-vl-7b-instruct
Tongyi Qianwen 2-VL-72B, extends the context to 32k, enhances image understanding capabilities, and can better recognize multi-language and handwriting in pictures.
📊 7B(Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-14b-instruct
Qwen2.5 series 14B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge, and has been greatly improved in programming ability and mathematical ability.
📊 14B (Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-14b-instruct-1m
Qwen2.5 series 14B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge, and has been greatly improved in programming ability and mathematical ability.
📊 14B (Dense) 💰 Open source and free (Apache 📐 1M
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-32b-instruct
Tongyi Qianwen 2.5 generation 32B parameter instruction model
📊 32B (Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-3b-instruct
Qwen2.5 series 3B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge, and has been greatly improved in programming ability and mathematical ability.
📊 3B(Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-72b-instruct
Qwen2.5 series 72B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge and has been greatly improved in programming ability and mathematical ability.
📊 72B (Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-7b-instruct
qwen2.5-7b-instruct is an artificial intelligence model provided by aliyun-bailian.
📊 7B(Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-7b-instruct-1m
qwen2.5-7b-instruct-1m is an artificial intelligence model provided by aliyun-bailian.
📊 7B(Dense) 💰 Open source and free (Apache 📐 1M
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen2.5-coder-14b-instruct
Qwen2.5 series programming expert 14B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge and has greatly improved in programming ability and mathematical ability.
📊 14B (Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen2.5-coder-32b-instruct
Qwen2.5 series programming expert 32B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge and has greatly improved in programming ability and mathematical ability.
📊 32B (Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen2.5-coder-7b-instruct
Qwen2.5 series programming expert 7B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge and has greatly improved in programming ability and mathematical ability.
📊 7B(Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-math-72b-instruct
Qwen2.5 series mathematics expert 72B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge, and has been greatly improved in programming ability and mathematical ability.
📊 72B (Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-math-7b-instruct
Qwen2.5 series mathematics expert 7B model, compared with Qwen2, Qwen2.5 has gained significantly more knowledge, and has been greatly improved in programming ability and mathematical ability.
📊 7B(Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-vl-32b-instruct
The flagship visual language model of the Qwen model family has achieved a huge leap forward compared to the previously released Qwen2-VL.
📊 32B (Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-vl-3b-instruct
Command following, mathematics, problem solving, and overall coding have been improved, and the ability to recognize everything has been improved. It supports multiple formats to directly and accurately locate visual elements. It supports the understanding of long video files (up to 10 minutes) and the positioning of events at the second level. It can understand the sequence and speed of time, and supports control based on analysis and positioning capabilities.
📊 3B(Dense) 💰 Open source and free (Apache 📐 128K
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-vl-72b-instruct
qwen2.5-vl-72b-instruct is an artificial intelligence model provided by aliyun-bailian.
📊 72B parameters, Dense architecture, native dynamic resolution
🧠 Model
AI Writing Ali (Qwen)
qwen2.5-vl-7b-instruct
Command following, mathematics, problem solving, and overall coding have been improved, and the ability to recognize everything has been improved. It supports multiple formats to directly and accurately locate visual elements. It supports the understanding of long video files (up to 10 minutes) and the positioning of events at the second level. It can understand the sequence and speed of time, and supports control based on analysis and positioning capabilities.
📊 7B parameters, Dense architecture, ViT+LLM
🧠 Model
AI Writing Ali (Qwen)
qwen3-0.6b
The data set of Qwen3 is significantly expanded compared to Qwen2.5.
📊 0.6B parameters, Dense architecture, pre-training 36
🧠 Model
AI Writing Ali (Qwen)
qwen3-1.7b
The data set of Qwen3 is significantly expanded compared to Qwen2.5.
📊 1.7B parameters, Dense architecture, pre-training 36
🧠 Model
AI Writing Ali (Qwen)
qwen3-14b
Achieve effective integration of thinking mode and non-thinking mode, and switch modes during the conversation.
📊 14B parameters, Dense architecture, pre-training 36T
🧠 Model
AI Writing Ali (Qwen)
qwen3-235b-a22b-instruct-2507
The qwen3-235b-a22b-instruct-2507 model released in July 2025 only supports non-thinking mode and is an upgraded version of qwen3-235b-a22b (non-thinking mode).
📊 235B total parameters/22B activation, MoE architecture (
🧠 Model
AI Writing Ali (Qwen)
qwen3-235b-a22b
The data set of Qwen3 is significantly expanded compared to Qwen2.5.
📊 235B total parameters/22B activation, MoE architecture (
🧠 Model
AI Writing Ali (Qwen)
qwen3-30b-a3b-instruct-2507
The qwen3-30b-a3b-instruct-2507 model released in July 2025 only supports non-thinking mode and is an upgraded version of qwen3-30b-a3b (non-thinking mode).
📊 30B total parameters/3B activation, MoE architecture, 12
🧠 Model
AI Writing Ali (Qwen)
qwen3-30b-a3b-thinking-2507
The qwen3-30b-a3b-thinking-2507 model released in July 2025 only supports thinking mode and is an upgraded version of qwen3-30b-a3b (thinking mode).
📊 30B total parameters/3B activation, MoE architecture, 12
🧠 Model
AI Writing Ali (Qwen)
qwen3-30b-a3b-think
The data set of Qwen3 is significantly expanded compared to Qwen2.5.
📊 30B total parameters/3B activation, MoE architecture, pre-training
🧠 Model
AI Writing Ali (Qwen)
qwen3-32b
The data set of Qwen3 is significantly expanded compared to Qwen2.5.
📊 32B parameters, Dense architecture, pre-training 36T
🧠 Model
AI Writing Ali (Qwen)
qwen3-4b
The data set of Qwen3 is significantly expanded compared to Qwen2.5.
📊 4B parameters, Dense architecture, pre-training 36T
🧠 Model
AI Writing Ali (Qwen)
qwen3-8b
The data set of Qwen3 is significantly expanded compared to Qwen2.5.
📊 8B parameters, Dense architecture, pre-training 36T
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3-coder
Based on the code generation model of Qwen3, it has powerful coding agent capabilities, is good at tool calling and environment interaction, and can achieve independent programming, excellent coding capabilities and general capabilities.
📊 Based on Qwen3 flagship model (guessed ~480B
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3-coder-flash
Based on the code generation model of Qwen3, it inherits the coding agent capabilities of Qwen3-Coder-Plus, supports multiple rounds of tool interaction, and focuses on optimizing warehouse-level understanding capabilities and increasing tool calling stability.
📊 Based on Qwen3-Coder-Plus distillation
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3-coder-30b-a3b-instruct
qwen3-coder-30b-a3b-instruct is a code generation model based on Qwen3. It has powerful Coding Agent capabilities and is good at tool calling and environment interaction. It can achieve independent programming and excellent coding capabilities while being versatile.
📊 30B total parameters/3B activation, MoE architecture, focus
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3-coder-480b-a35b-instruct
Tongyi Qianwen code model open source version.
📊 480B total parameters/35B activation, MoE architecture,
🧠 Model
AI Writing Ali (Qwen)
qwen3-max-2026-01-23
Tongyi Qianwen 3 series Max model, compared with the snapshot on September 23, 2025, this version realizes the effective integration of thinking mode and non-thinking mode, and the overall effect of the model has been greatly improved in all aspects.
📊 Ultra-large scale (estimated ~1T parameter level), closed source model
🧠 Model
AI Productivity Ali (Qwen)
qwen3-rerank
Based on the text ranking model trained by Qwen LLM base, the input Query and candidate Docs are sorted by relevance. It supports 100+ languages and long text input. It is suitable for text retrieval, RAG and other scenarios. The effect is aligned with the open source Qwen3-Rerank series model.
📊 Based on Qwen3 base, designed for text sorting (Rer
🧠 Model
AI Writing Ali (Qwen)
qwen3-max-preview-n
Reverse version of qwen3-max-preview
📊 Reverse/special variant of Qwen3-Max series (
🧠 Model
AI Writing Ali (Qwen)
qwen3-next-80b-a3b-instruct
Qwen3-Next 80B-A3B Instruct is an artificial intelligence model provided by alibaba-cn.
📊 80B total parameters/3B activations, super sparse MoE architecture
🧠 Model
AI Writing Ali (Qwen)
qwen3-next-80b-a3b-thinking
qwen3-next-80b-a3b-thinking, released in September 2025, only supports thinking mode. Compared with qwen3-235b-a22b-thinking-2507, it has improved command following capabilities and the summary reply is more streamlined.
📊 80B total parameters/3B activation, super sparse MoE+ mixing
🧠 Model
AI Writing Ali (Qwen)
qwen3-vl-235b-a22b-instruct
Alibaba Cloud's Tongyi Qianwen VL open source version.
📊 235B total parameters/22B activation, MoE architecture+
🧠 Model
AI Writing Ali (Qwen)
qwen3-vl-235b-a22b-thinking
Alibaba Cloud's Tongyi Qianwen VL open source version.
📊 235B total parameters/22B activation, MoE+Vi
🧠 Model
AI Writing Ali (Qwen)
qwen3-vl-30b-a3b-instruct
The Instruct version of the second largest MoE model in the Qwen3-VL series has fast response speed and supports ultra-long contexts such as long videos and long documents; it has comprehensively upgraded image/video understanding, spatial perception and all-things recognition capabilities; it has visual 2D/3D positioning capabilities and is capable of complex real-world tasks.
📊 30B total parameters/3B activation, MoE+ViT frame
🧠 Model
AI Writing Ali (Qwen)
qwen3-vl-30b-a3b-thinking
Qwen3 - Thinking version of the second largest MoE model in the VL series, with fast response speed, stronger multi-modal understanding and reasoning, visual agent, long video and long document and other ultra-long context support capabilities; comprehensive upgrade of image/video understanding, spatial perception and everything recognition capabilities, capable of complex tasks
📊 30B total parameters/3B activation, MoE+ViT frame
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3-vl-32b-instruct
The non-inference version of the largest Dense model of the Qwen3-VL series, its overall performance is second only to Qwen3-VL-235B-Instruct. It has excellent document recognition and understanding capabilities, strong spatial perception and recognition of all things, and visual 2D detection/spatial reasoning capabilities reaching SO
📊 32B parameters, Dense architecture + ViT visual editing
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3-vl-32b-thinking
The reasoning version of the largest Dense model of the Qwen3-VL series. Its multi-modal reasoning ability is second only to Qwen3-VL-235B-Thinking. It has outstanding STEM & mathematics problem-solving ability, general image and video understanding ability, and its multi-modal Agent ability reaches S
📊 32B parameters, Dense architecture + ViT visual editing
🧠 Model
AI Writing Ali (Qwen)
qwen3-vl-8b-instruct
The Instruct version of the Qwen3-VL series 8B Dense model takes up less video memory, comprehensively upgrades image/video understanding, long video and long document and other ultra-long context support, spatial perception and everything recognition capabilities, and is capable of complex real-world tasks.
📊 8B parameters, Dense architecture + ViT visual coding
🧠 Model
AI Writing Ali (Qwen)
qwen3-vl-8b-thinking
The Thinking version of the Qwen3-VL series 8B Dense model takes up less video memory and can complete multi-modal understanding and reasoning; supports long videos and long documents and other long contexts, visual 2D/3D positioning; comprehensively upgrades image/video understanding, spatial perception and all things recognition capabilities
📊 8B parameters, Dense architecture + ViT visual coding
🧠 Model
AI Writing Ali (Qwen)
qwen3-vl-flash
The Qwen3 series of small-size visual understanding models realize the effective integration of thinking mode and non-thinking mode. The effect is better than the open source version Qwen3-VL-30B-A3B, and the response speed is fast.
📊 Small size VL model (presumably based on 30B-A3B steamer
🧠 Model
AI Writing Ali (Qwen)
qwen3-vl-plus
Tongyi Qianwen VL is a text generation model with visual (image) understanding capabilities. It can not only perform OCR (picture text recognition), but also further summarize and reason, such as extracting attributes from product photos, solving problems based on exercise diagrams, etc.
📊 Medium-scale VL model (presumably based on Qwen3-V
🧠 Model
AI Writing Ali (Qwen)
qwen3.5-122b-a10b
The Qwen3.5 series 122B-A10B native visual language model is based on a hybrid architecture design and integrates a linear attention mechanism and a sparse hybrid expert model to achieve higher reasoning efficiency.
📊 122B total parameters/10B activation, Gated
🧠 Model
AI Writing Ali (Qwen)
qwen3.5-27b
Qwen3.5 series 27B native visual language Dense model incorporates a linear attention mechanism; it has fast response speed and combines inference speed and performance.
📊 27B parameters, Dense architecture + Gated
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3.5-35b-a3b
The Qwen3.5 series 35B-A3B native visual language model is based on a hybrid architecture design and integrates a linear attention mechanism and a sparse hybrid expert model to achieve higher reasoning efficiency.
📊 35B total parameters/3B activation, Gated De
🧠 Model
AI Writing Ali (Qwen)
qwen3.5-397b-a17b
The Qwen3.5 series 397B-A17B native visual language model is based on a hybrid architecture design and integrates a linear attention mechanism and a sparse hybrid expert model to achieve higher reasoning efficiency.
📊 397B total parameters/17B activated, Gated
🧠 Model
AI Writing Ali (Qwen)
qwen3.5-plus
The wen3.5 native visual language series Plus model is based on a hybrid architecture design that integrates a linear attention mechanism and a sparse hybrid expert model to achieve higher reasoning efficiency.
📊 Qwen3.5 series Plus closed source model, Ga
🧠 Model
AI Writing Ali (Qwen)
qwen3.6-27b
Qwen3.6 series 27B native visual language Dense model. Compared with 3.5-27B, the model effect focuses on improving agentic coding capabilities, model STEM and reasoning capabilities. In terms of visual modality, it has spatial intelligence, object positioning and detection capabilities.
📊 27B parameters, Dense architecture + Gated
🧠 Model
AI Writing Ali (Qwen)
qwen3.6-35b-a3b
The Qwen3.6 series 35B-A3B native visual language model is based on a hybrid architecture design and integrates a linear attention mechanism and a sparse hybrid expert model to achieve higher reasoning efficiency.
📊 35B total parameters/3B activation, Gated De
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3.6-max-preview
The Max model Preview version, which is the largest and most comprehensive in the Qwen3.6 series, currently has the pure text model capability open for experience.
📊 Qwen3.6 series largest Max model (recommended
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3.6-plus
The Qwen 3.6 native visual language series Plus model shows excellent performance comparable to the current top cutting-edge models, and the model effect is significantly improved compared to the 3.5 series.
📊 Qwen3.6 series Plus closed source model, Ga
🧠 Model
AI CodingAI Writing Ali (Qwen)
qwen3.7-max
The Max model, which is the largest and most comprehensive model in the Qwen3.7 series, currently has pure text model capabilities open for experience.
📊 Qwen3.7 series’ largest flagship Max model
🧠 Model
AI Writing Ali (Qwen)
qwq-32b
QwQ - open source version.
📊 32B (Dense) 💰 Open source and free (Apache 📐 131K
🧠 Model
AI Writing Ali (Qwen)
qwq-72b-preview
The QwQ-Preview model is an experimental research model developed by the Qwen team in 2024, focusing on enhancing AI reasoning capabilities, especially in the fields of mathematics and programming.
📊 72B (Dense) 💰 Open source and free (Apache 📐 32K
🧠 Model
AI Writing Ali (Qwen)
qwq-plus
The stable version of QwQ is based on the QwQ reasoning model trained by the Qwen2.5 model, which greatly improves the model's reasoning capabilities through reinforcement learning.
📊 MoE ~70B+ (general parameters) 💰 Bailian API ¥0.8/M 📐 128K
🧠 Model
AI Video Ali (Qwen)
wan2.5-i2v-preview
wan2.5-i2v-preview is an artificial intelligence model provided by aliyun-bailian.
📊 Ali Wanxiang 2.5 Tusheng Video (Image-to
🧠 Model
AI Video Ali (Qwen)
wan2.6-i2v
Wanxiang 2.6.
📊 Wanxiang 2.6 Tusheng video version, DiT architecture, 108
🧠 Model
AI Video Ali (Qwen)
wan2.6-i2v-flash
Wanxiang 2.6-Tusheng Video-Flash, faster and more cost-effective generation.
📊 Wanxiang 2.6 Tusheng Video Flash distilled version, Di
🧠 Model
AI Image Gen Ali (Qwen)
wan2.7-image-pro
Wanxiang 2.7 - the ultimate model for image generation and editing, supports Vincentian pictures, Vincentian group pictures, Tusheng group pictures, image editing, multi-image reference generation, interactive editing, and has stronger performance in text rendering, subject consistency, and complex instruction compliance.
📊 Wanxiang 2.7 image flagship version, unified generation + editing framework
🧠 Model
AI Image Gen Ali (Qwen)
z-image-turbo
Z-Image-Turbo is an efficient image generation model that ranked first in the world among Vincentian open source models in the Artificial Analysis evaluation. It can generate photorealistic images comparable to large-scale commercial models with only 6 billion parameters and 8 steps of reasoning. It is also used in both Chinese and English.
📊 6B 💰 Open source and free (Apache 📐 N/A (image generation)
