deepseek-v4.1-flash
DeepSeek
DeepSeek-V4.1-Flash 是 DeepSeek 当前主力模型,官方 API 模型名为 deepseek-flash。支持 1M 上下文与最大 384K 输出,默认可切换思考/非思考模式,具备图像理解、工具调用、JSON Output,并兼容 OpenAI 与 Anthropic API。
¥0.4194/ M
¥1.6775/ M
¥0.0084/ M
—
AI Models
Parcourir et comparer les modèles IA des meilleurs fournisseurs
76 modèles trouvés
DeepSeek
DeepSeek-V4.1-Flash 是 DeepSeek 当前主力模型,官方 API 模型名为 deepseek-flash。支持 1M 上下文与最大 384K 输出,默认可切换思考/非思考模式,具备图像理解、工具调用、JSON Output,并兼容 OpenAI 与 Anthropic API。
¥0.4194/ M
¥1.6775/ M
¥0.0084/ M
—
Qwen
The Flash tier of the Qwen3.8 series: multimodal reasoning for coding and agent workflows, with understanding of documents, codebases and long videos.
¥0.9286/ M
¥2.7859/ M
¥0.0929/ M
¥1.1608/ M
Z.AI
The Flash tier of the GLM-5.3 series: natively multimodal, built for efficient coding and long-chain agents, with stable long-context performance for production applications.
¥0.5528/ M
¥1.8425/ M
¥0.1106/ M
—
Qwen
The flagship reasoning model of the Qwen3.8 series, accepting multimodal input for complex reasoning, visual understanding, coding and agent workflows.
¥8.9732/ M
¥26.9196/ M
¥0.7478/ M
¥1.8722/ M
Moonshot
A multimodal reasoning model for complex coding, knowledge work, and long-running agent tasks. It works across large codebases, calls tools, and debugs, using images, logs, and test output to refine its output.
¥15.8346/ M
¥79.1728/ M
¥1.5835/ M
—
DeepSeek
The DeepSeek V4 flagship, with the series' strongest deep reasoning and code generation, suited to whole-codebase analysis, multi-step automation and demanding instructions.
¥1.9899/ M
¥5.9697/ M
¥0.0663/ M
—
Z.AI
Flagship-tier model for complex software engineering and long-horizon agent tasks, with stable execution across multi-turn tool calls and large-scale codebase refactoring.
¥7.1573/ M
¥25.0504/ M
¥1.7893/ M
—
Z.AI
GLM 5.2 is a large-scale reasoning model with a 1M-token context window, suited for long-horizon agents, repo-level coding, and multi-step automation.
¥3.5628/ M
¥12.4697/ M
¥0.6998/ M
¥0/ M
Moonshot AI
Coding-focused variant in the Kimi K2 family. Accepts text and image inputs, with reasoning on by default. Suited for long-horizon coding and agentic workflows.
¥4.503/ M
¥18.7046/ M
¥0.9006/ M
¥5.2964/ M
ByteDance
Seedream-5.0-pro是字节跳动发布的最新图像创作模型。该模型将图像创作推进到可控生产的新阶段,本次更新的主要亮点是编辑更可控、生产更落地、效果更自然。
—
¥0.1947/ image
—
—
ByteDance
Standard tier of the Seedance family. Supports text-to-video, image-to-video, and reference-to-video, with strong character, style, and camera consistency. Built for ads and short-form creative.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 480p | ¥45.4568/ M | ¥27.669/ M |
| 720p | ¥45.4568/ M | ¥27.669/ M |
| 1080p | ¥50.384/ M | ¥30.6681/ M |
ByteDance
Lightweight tier in the Seed 2.0 family, tuned for high-throughput, low-latency tasks with adjustable reasoning levels. Fits batch processing, moderation, and classification.
¥0.1861/ M
¥1.8606/ M
¥0.0372/ M
—
ByteDance
Seedance 2.0 mini is a new-generation, cost-effective video generation model for a broader range of needs. It keeps competitive quality while bringing video generation to lower-barrier, higher-frequency and larger-scale use cases.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 480p | ¥23.45/ M | ¥14.07/ M |
| 720p | ¥23.45/ M | ¥14.07/ M |
ByteDance
Lightweight Seedance 2.0 variant tuned for speed and cost rather than peak fidelity, with text-to-video, first/last frame control, and multi-reference inputs.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 480p | ¥21.9672/ M | ¥13.0615/ M |
| 720p | ¥21.9672/ M | ¥13.0615/ M |
ByteDance
Lightweight tier in the Seed 2.0 family, tuned for high-throughput, low-latency tasks with adjustable reasoning levels. Fits batch processing, moderation, and classification.
¥1.139/ M
¥4.556/ M
¥0.1139/ M
¥0.0475/ M
ByteDance
Flagship of the Seed 2.0 series for complex reasoning and long-horizon agent workflows with reliable tool use.
¥2.9769/ M
¥14.8845/ M
¥0.5954/ M
—
ByteDance
Flagship of the Seed 2.0 series for complex reasoning and long-horizon agent workflows with reliable tool use.
¥5.695/ M
¥34.17/ M
¥1.139/ M
¥0.0475/ M
Qwen
Flagship agent-centric reasoning model with a 1M context window, excelling at coding, productivity, and long-horizon autonomous tasks.
¥8.3131/ M
¥24.9394/ M
¥1.6626/ M
¥10.394/ M
ByteDance
From ByteDance, supports 文本 et 多模态 and more—call it via the standard API.
¥5.0924/ M
¥25.462/ M
¥1.0185/ M
¥0.0144/ M
ByteDance
From ByteDance, supports 文本 et 多模态 and more—call it via the standard API.
¥5.0924/ M
¥25.462/ M
¥1.0185/ M
¥0.0144/ M
ByteDance
Lightweight tier of the Seed 2.0 family tuned for low-latency agent, coding, and GUI workloads.
¥0.5582/ M
¥3.349/ M
¥0.1116/ M
—
ByteDance
Lightweight tier of the Seed 2.0 family tuned for low-latency agent, coding, and GUI workloads.
¥2.8475/ M
¥22.78/ M
¥0.5695/ M
¥0.0475/ M
Moonshot AI
Latest multimodal model in the K2 series, built for long-horizon coding, code-driven UI/UX generation, and multi-agent orchestration.
¥4.503/ M
¥18.7046/ M
¥0.762/ M
¥5.6288/ M
MiniMax
MiniMax\\'s first multimodal release, accepting image and video input. Long-context inference runs cheaper and faster, and it handles multi-step agent workflows and computer-use scenarios.
¥6.834/ M
¥27.336/ M
¥0.6834/ M
¥0/ M
Z.AI
Long-task autonomous coding model
¥3.5628/ M
¥14.2511/ M
¥0.7719/ M
—
DeepSeek
A faster, lower-cost version of the DeepSeek V4 series that keeps the main model's reasoning and coding foundations, suited to high-concurrency chat, coding assistance and latency-sensitive agents.
¥0.649/ M
¥2.5961/ M
¥0.013/ M
—
ByteDance
Seedance 2.5 是豆包大模型团队推出的新一代多模态视频创作模型,已迈入「长叙事 × 多素材 × 强编辑 × 多语言」的工业级新阶段。
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 480 | ¥0.6888/ s | ¥0.6888/ s |
| 720 | ¥1.5484/ s | ¥1.5484/ s |
Qwen
Qwen 3.7 Plus is a mid-tier multimodal model that reads screens, operates GUIs, and navigates mobile apps end-to-end. Suited for agent workflows and tool calling.
¥1.3958/ M
¥5.5833/ M
¥0.2792/ M
¥1.7448/ M
Tencent
Hy3 preview in high-reasoning mode, built for complex problem solving and multi-step agent workflows that need dependable execution.
¥1.0251/ M
¥3.417/ M
¥0.3417/ M
¥0/ M
Xiaomi
Flagship coding and agent model that sustains thousand-tool-call autonomous workflows for complex software engineering.
¥2.4773/ M
¥4.9546/ M
¥0.0205/ M
¥0/ M

MiniMax
It supports Vincentian and Tusheng videos, can output 6-10 seconds, 768p/1080p images, supports fine camera movement command control, has smooth movement and stable subject, and is suitable for short video creative content.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 1080p | ¥11.8188/ M | ¥11.8188/ M |
| 768p | ¥13.5072/ M | ¥13.5072/ M |
Alibaba
Wan2.7-T2V delivers fully upgraded performances: nuanced, natural emotion in dramatic scenes, hard-hitting action sequences, and more dramatic, rhythmic shot transitions.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720 | ¥0.4275/ s | ¥0.4275/ s |
| 1080 | ¥0.7126/ s | ¥0.7126/ s |

MiniMax
It supports Vincentian and Tusheng videos, can output 6-10 seconds, 512p/768p/1080p multi-level image quality, supports camera movement command control, and flexibly balances image quality and cost.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 1080p | ¥11.8188/ M | ¥11.8188/ M |
| 512p | ¥3.618/ M | ¥3.618/ M |
| 768p | ¥13.5072/ M | ¥13.5072/ M |
Alibaba
Generate 720p/1080p video from text prompts and come with synchronized audio tracks. The picture has movie-level aesthetics and complex motion performance, suitable for creative video creation.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720p | ¥0.603/ s | ¥0.603/ s |
| 1080p | ¥0.9045/ s | ¥0.9045/ s |
Qwen
The first image editing model in the Qwen series, extending Qwen-Image's text rendering to editing tasks. Supports precise bilingual (Chinese and English) text editing, both appearance and semantic edits, and strong results across benchmarks.
—
¥0.1633/ image
—
—
Qwen
Unifies generation and editing in one model; native 2K; professional typography and infographics.
—
¥0.1652/ image
—
—
Fuses generation and editing; more professional text rendering and richer realistic textures and scenes.
—
¥0.2474/ image
—
—
ByteDance
New unified architecture combining generation and editing; knowledge-based output and reference consistency; up to 4K.
—
¥0.1909/ image
—
—
ByteDance
Broad upgrade via model scaling; sharper subject detection, stricter reference fidelity, stronger dense text rendering.
—
¥0.2326/ image
—
—
ByteDance
Adds deep reasoning; better interprets complex prompts with common-sense knowledge; stronger edit consistency.
—
¥0.2047/ image
—
—
MiniMax
Speed-optimized variant of M2.7 with identical outputs, tuned for low-latency coding and agent tool-calling workloads.
¥3.417/ M
¥13.668/ M
¥0.3417/ M
¥2.1356/ M
Qwen
The lightweight vision-language reasoning model of the Qwen3.7 series, strong at object recognition, spatial understanding and real-world visual perception, built for multimodal agents, visual coding, search and computer use.
¥0.201/ M
¥0.871/ M
¥0.0402/ M
¥0.2546/ M
MiniMax
A high-throughput variant of M2.5 with matching quality at roughly triple the speed, tuned for coding and agentic tool use under latency-sensitive workloads.
¥3.417/ M
¥13.668/ M
¥0.1709/ M
¥2.1356/ M
Xiaomi
Mid-tier V2.5 model with native image, audio, and video understanding, built for agent workflows and coding at lower cost than Pro.
¥0.7973/ M
¥1.5946/ M
¥0.2437/ M
¥0/ M
ByteDance
Agent-oriented foundation model tuned for tool calling and complex instruction following, with use cases spanning GUI agents, search agents, and multi-step task orchestration.
¥2.8475/ M
¥22.78/ M
¥0.2848/ M
¥0.0475/ M
ByteDance
Code-focused preview model for agent workflows, with competitive SWE-Bench and LiveCodeBench scores, aimed at autonomous coding agents and multi-language software tasks.
¥5.695/ M
¥34.17/ M
¥0.5695/ M
¥0.0475/ M
Alibaba
Lightweight mid-tier variant tuned for low-latency agent workflows with coding and tool-calling support.
¥0.1652/ M
¥1.6345/ M
¥0/ M
¥0/ M
Alibaba
Open-weight coding model tuned for agentic terminal tasks and repo-scale reasoning with low inference cost.
¥1.4124/ M
¥8.4571/ M
¥0/ M
¥0/ M
Alibaba
Mid-tier Qwen model tuned for multi-step reasoning and agent workflows, compact enough for private and on-prem deployment.
¥0.4898/ M
¥3.9182/ M
¥0/ M
¥0/ M
Alibaba
Thinking-mode variant of the dense open-weight model that outputs step-by-step reasoning; at 27B it surpasses the prior open-weight 397B-A17B flagship across coding, math and multi-step reasoning benchmarks.
¥2.3496/ M
¥14.0973/ M
¥0/ M
¥0/ M
Qwen
Multimodal model from the Qwen3.5 line with toggleable thinking mode, tuned for image and video understanding, document parsing, and multimodal agents.
¥0.6549/ M
¥3.9182/ M
¥0.3246/ M
¥4.0833/ M
Alibaba
Open-weight mid-tier MoE text model with long chain-of-thought reasoning, strong on function-calling and multi-step planning benchmarks for self-hosted agentic workflows.
¥0.6549/ M
¥5.2223/ M
¥0/ M
¥0/ M
Qwen
Plus-tier vision-language model in the Qwen3-VL line, handling text, image and video inputs with strengths in document parsing, video understanding, spatial grounding and agent tool use.
¥0.8201/ M
¥8.1666/ M
¥0.1652/ M
¥3.0696/ M
Z.AI
Text-only model in the GLM-4 family with improved coding benchmarks and tool calls during reasoning, suited for software development and agent workflows.
¥3.2689/ M
¥13.0643/ M
¥0.6549/ M
¥0/ M
Kling
Kling O1 是为创作者打造的全新创意引擎,解锁无限创意可能。基于多模态视觉语言(MVL)概念,使用自然语言结合视频、图片、元素等多模态描述,精准理解您的创作意图。
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720p | ¥0.3888/ s | ¥0.3888/ s |
| 1080p | ¥0.5185/ s | ¥0.5185/ s |
Kling
可灵3.0 Omni是旗舰模型,提供卓越的视频质量、先进的语义理解和增强的伪影修正。基于视频3.0的基础,Omni为专业创作者提供极致品质,具有更精细的细节、更流畅的运动动态和更高的成功率。
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720p | ¥0.3888/ s | ¥0.3888/ s |
| 1080p | ¥0.5185/ s | ¥0.5185/ s |
Kling
可灵视频 3.0 配备革命性智能分镜系统,自动编排多镜头序列,实现电影级转场效果。单次生成最长15秒视频,具备增强的角色一致性、原生级文字叠加和多语言音频混合能力。
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720p | ¥0.3888/ s | ¥0.3888/ s |
| 1080p | ¥0.5185/ s | ¥0.5185/ s |
| 4k | ¥1.9442/ s | ¥1.9442/ s |
| 2k | ¥0.6481/ s | ¥0.6481/ s |
Kling
VIDEO 2.6 模型一次生成视觉画面、自然配音、匹配音效和环境氛围,连接「声音」与「视觉」的世界。无论是输入文字还是上传图片,都能即时创建一个完整的动态视频,包含声音、节奏和沉浸感。
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720p | ¥0.1944/ s | ¥0.1944/ s |
| 1080p | ¥0.324/ s | ¥0.324/ s |
Kling
Kling 2.5 Turbo 针对速度和性价比进行了优化,非常适合大批量视频生成需求。在显著缩短生成时间的同时保持良好的质量。
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720p | ¥0.1944/ s | ¥0.1944/ s |
| 1080p | ¥0.324/ s | ¥0.324/ s |
Kling
Kling 2.1 显著提升了语义响应、动态质量和画面质感。是稳定、高质量视频生成的首选。
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720p | ¥0.2592/ s | ¥0.2592/ s |
| 1080p | ¥0.4536/ s | ¥0.4536/ s |
Kling
Kling 2.0 大师版针对专业创作场景优化,为高要求项目提供电影级输出质量。
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720p | ¥0.2592/ s | ¥0.2592/ s |
| 1080p | ¥0.4536/ s | ¥0.4536/ s |
Alibaba
From Alibaba, supports Video and more—call it via the standard API.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720 | ¥0.335/ s | ¥0.335/ s |
| 1080 | ¥0.5025/ s | ¥0.5025/ s |
Alibaba
From Alibaba, supports Video and more—call it via the standard API.
| Résolution | Sans entrée vidéo | Avec entrée vidéo |
|---|---|---|
| 720 | ¥0.67/ s | ¥0.67/ s |
| 1080 | ¥1.005/ s | ¥1.005/ s |
Qwen
Qwen 3.6 lightweight tier with text, image, and video input. Low latency and low cost, well-suited for high-volume classification, extraction, summarization, and simple agent workflows.
¥0.8297/ M
¥4.9803/ M
¥0.0825/ M
¥1.0374/ M
Moonshot AI
Instruction-tuned for agent workflows with reliable tool calling and multilingual coding.
¥3.2689/ M
¥13.0985/ M
¥0.6549/ M
¥0/ M
Moonshot AI
Open-source thinking model built for long-horizon reasoning and hundreds of sequential tool calls across autonomous research, coding, and agent workflows.
¥3.417/ M
¥14.2375/ M
¥0.8543/ M
¥0/ M
Qwen
Enhanced Qwen model with strong bilingual (Chinese/English) comprehension, excelling at long-document analysis and structured output.
¥1.3855/ M
¥8.3131/ M
¥0.1386/ M
¥1.7319/ M
Xiaomi
Flagship agent model for complex workflows
¥11.39/ M
¥34.17/ M
¥2.278/ M
¥0/ M
MiniMax
Multi-agent model for complex workflows
¥1.4548/ M
¥5.8192/ M
¥0.291/ M
¥0/ M
Z.AI
Balanced model with strong Chinese understanding.
¥3.36/ M
¥13.3833/ M
¥3.1322/ M
¥0/ M
Z.AI
Multimodal model with strong image + text understanding.
¥1.7085/ M
¥5.1255/ M
¥0.3132/ M
¥0/ M
MiniMax
Cost-effective general model for balanced speed and quality.
¥1.7655/ M
¥7.0049/ M
¥0.1765/ M
¥2.2211/ M
Z.AI
Designed for agent-based workflows such as OpenClaw.
¥6.834/ M
¥22.78/ M
¥1.3668/ M
¥0/ M
Moonshot AI
Top Chinese-English bilingual. Strong long context.
¥3.417/ M
¥17.1476/ M
¥0.8543/ M
¥4.089/ M
Z.AI
Competitive general-purpose Chinese model.
¥5.0116/ M
¥18.3949/ M
¥0.9795/ M
¥0/ M
Z.AI
Ultra-fast for Chinese. Very cheap.
¥0.5695/ M
¥2.4489/ M
¥0.057/ M
¥0/ M