deepseek-v4.1-flash
DeepSeek
DeepSeek-V4.1-Flash 是 DeepSeek 当前主力模型,官方 API 模型名为 deepseek-flash。支持 1M 上下文与最大 384K 输出,默认可切换思考/非思考模式,具备图像理解、工具调用、JSON Output,并兼容 OpenAI 与 Anthropic API。
¥0.4187/ M
¥1.675/ M
¥0.0084/ M
—
AI モデル
トッププロバイダーの AI モデルを閲覧・比較
76 件のモデル
DeepSeek
DeepSeek-V4.1-Flash 是 DeepSeek 当前主力模型,官方 API 模型名为 deepseek-flash。支持 1M 上下文与最大 384K 输出,默认可切换思考/非思考模式,具备图像理解、工具调用、JSON Output,并兼容 OpenAI 与 Anthropic API。
¥0.4187/ M
¥1.675/ M
¥0.0084/ M
—
Qwen
The Flash tier of the Qwen3.8 series: multimodal reasoning for coding and agent workflows, with understanding of documents, codebases and long videos.
¥0.9272/ M
¥2.7817/ M
¥0.0927/ M
¥1.159/ M
Z.AI
The Flash tier of the GLM-5.3 series: natively multimodal, built for efficient coding and long-chain agents, with stable long-context performance for production applications.
¥0.8028/ M
¥2.676/ M
¥0.1606/ M
—
Qwen
The flagship reasoning model of the Qwen3.8 series, accepting multimodal input for complex reasoning, visual understanding, coding and agent workflows.
¥8.9598/ M
¥26.8795/ M
¥0.7467/ M
¥1.8694/ M
Moonshot
A multimodal reasoning model for complex coding, knowledge work, and long-running agent tasks. It works across large codebases, calls tools, and debugs, using images, logs, and test output to refine its output.
¥15.8109/ M
¥79.0547/ M
¥1.5811/ M
—
DeepSeek
The DeepSeek V4 flagship, with the series' strongest deep reasoning and code generation, suited to whole-codebase analysis, multi-step automation and demanding instructions.
¥1.9869/ M
¥5.9608/ M
¥0.0662/ M
—
Z.AI
Flagship-tier model for complex software engineering and long-horizon agent tasks, with stable execution across multi-turn tool calls and large-scale codebase refactoring.
¥7.1466/ M
¥25.0131/ M
¥1.7866/ M
—
Z.AI
GLM 5.2 is a large-scale reasoning model with a 1M-token context window, suited for long-horizon agents, repo-level coding, and multi-step automation.
¥3.5575/ M
¥12.4511/ M
¥0.6988/ M
¥0/ M
Moonshot AI
Coding-focused variant in the Kimi K2 family. Accepts text and image inputs, with reasoning on by default. Suited for long-horizon coding and agentic workflows.
¥4.4962/ M
¥18.6767/ M
¥0.8992/ M
¥5.2885/ M
ByteDance
Seedream-5.0-pro是字节跳动发布的最新图像创作模型。该模型将图像创作推进到可控生产的新阶段,本次更新的主要亮点是编辑更可控、生产更落地、效果更自然。
—
¥0.1944/ 枚
—
—
ByteDance
Standard tier of the Seedance family. Supports text-to-video, image-to-video, and reference-to-video, with strong character, style, and camera consistency. Built for ads and short-form creative.
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 480p | ¥45.389/ M | ¥27.6277/ M |
| 720p | ¥45.389/ M | ¥27.6277/ M |
| 1080p | ¥50.3088/ M | ¥30.6223/ M |
ByteDance
Lightweight tier in the Seed 2.0 family, tuned for high-throughput, low-latency tasks with adjustable reasoning levels. Fits batch processing, moderation, and classification.
¥0.1858/ M
¥1.8578/ M
¥0.0372/ M
—
ByteDance
Seedance 2.0 mini is a new-generation, cost-effective video generation model for a broader range of needs. It keeps competitive quality while bringing video generation to lower-barrier, higher-frequency and larger-scale use cases.
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 480p | ¥23.415/ M | ¥14.049/ M |
| 720p | ¥23.415/ M | ¥14.049/ M |
ByteDance
Lightweight Seedance 2.0 variant tuned for speed and cost rather than peak fidelity, with text-to-video, first/last frame control, and multi-reference inputs.
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 480p | ¥21.9344/ M | ¥13.042/ M |
| 720p | ¥21.9344/ M | ¥13.042/ M |
ByteDance
Lightweight tier in the Seed 2.0 family, tuned for high-throughput, low-latency tasks with adjustable reasoning levels. Fits batch processing, moderation, and classification.
¥1.1373/ M
¥4.5492/ M
¥0.1137/ M
¥0.0474/ M
ByteDance
Flagship of the Seed 2.0 series for complex reasoning and long-horizon agent workflows with reliable tool use.
¥2.9725/ M
¥14.8623/ M
¥0.5945/ M
—
ByteDance
Flagship of the Seed 2.0 series for complex reasoning and long-horizon agent workflows with reliable tool use.
¥5.6865/ M
¥34.119/ M
¥1.1373/ M
¥0.0474/ M
Qwen
Flagship agent-centric reasoning model with a 1M context window, excelling at coding, productivity, and long-horizon autonomous tasks.
¥8.3007/ M
¥24.9022/ M
¥1.6601/ M
¥10.3784/ M
ByteDance
ByteDance のモデル。文本、多模态 などに対応。標準 API で呼び出し可能。
¥5.0848/ M
¥25.424/ M
¥1.017/ M
¥0.0144/ M
ByteDance
ByteDance のモデル。文本、多模态 などに対応。標準 API で呼び出し可能。
¥5.0848/ M
¥25.424/ M
¥1.017/ M
¥0.0144/ M
ByteDance
Lightweight tier of the Seed 2.0 family tuned for low-latency agent, coding, and GUI workloads.
¥0.5573/ M
¥3.344/ M
¥0.1115/ M
—
ByteDance
Lightweight tier of the Seed 2.0 family tuned for low-latency agent, coding, and GUI workloads.
¥2.8433/ M
¥22.746/ M
¥0.5687/ M
¥0.0474/ M
Moonshot AI
Latest multimodal model in the K2 series, built for long-horizon coding, code-driven UI/UX generation, and multi-agent orchestration.
¥4.4962/ M
¥18.6767/ M
¥0.7609/ M
¥5.6204/ M
MiniMax
MiniMax\\'s first multimodal release, accepting image and video input. Long-context inference runs cheaper and faster, and it handles multi-step agent workflows and computer-use scenarios.
¥6.8238/ M
¥27.2952/ M
¥0.6824/ M
¥0/ M
Z.AI
Long-task autonomous coding model
¥3.5575/ M
¥14.2298/ M
¥0.7708/ M
—
DeepSeek
A faster, lower-cost version of the DeepSeek V4 series that keeps the main model's reasoning and coding foundations, suited to high-concurrency chat, coding assistance and latency-sensitive agents.
¥0.6481/ M
¥2.5922/ M
¥0.013/ M
—
ByteDance
Seedance 2.5 是豆包大模型团队推出的新一代多模态视频创作模型,已迈入「长叙事 × 多素材 × 强编辑 × 多语言」的工业级新阶段。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 480 | ¥0.6877/ s | ¥0.6877/ s |
| 720 | ¥1.5461/ s | ¥1.5461/ s |
Qwen
Qwen 3.7 Plus is a mid-tier multimodal model that reads screens, operates GUIs, and navigates mobile apps end-to-end. Suited for agent workflows and tool calling.
¥1.3937/ M
¥5.575/ M
¥0.2788/ M
¥1.7422/ M
Tencent
Hy3 preview in high-reasoning mode, built for complex problem solving and multi-step agent workflows that need dependable execution.
¥1.0236/ M
¥3.4119/ M
¥0.3412/ M
¥0/ M
Xiaomi
Flagship coding and agent model that sustains thousand-tool-call autonomous workflows for complex software engineering.
¥2.4736/ M
¥4.9473/ M
¥0.0205/ M
¥0/ M

MiniMax
It supports Vincentian and Tusheng videos, can output 6-10 seconds, 768p/1080p images, supports fine camera movement command control, has smooth movement and stable subject, and is suitable for short video creative content.
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 1080p | ¥11.8012/ M | ¥11.8012/ M |
| 768p | ¥13.487/ M | ¥13.487/ M |
Alibaba
Wan2.7-T2V delivers fully upgraded performances: nuanced, natural emotion in dramatic scenes, hard-hitting action sequences, and more dramatic, rhythmic shot transitions.
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720 | ¥0.4269/ s | ¥0.4269/ s |
| 1080 | ¥0.7115/ s | ¥0.7115/ s |

MiniMax
It supports Vincentian and Tusheng videos, can output 6-10 seconds, 512p/768p/1080p multi-level image quality, supports camera movement command control, and flexibly balances image quality and cost.
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 1080p | ¥11.8012/ M | ¥11.8012/ M |
| 512p | ¥3.6126/ M | ¥3.6126/ M |
| 768p | ¥13.487/ M | ¥13.487/ M |
Alibaba
Generate 720p/1080p video from text prompts and come with synchronized audio tracks. The picture has movie-level aesthetics and complex motion performance, suitable for creative video creation.
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720p | ¥0.6021/ s | ¥0.6021/ s |
| 1080p | ¥0.9032/ s | ¥0.9032/ s |
Qwen
The first image editing model in the Qwen series, extending Qwen-Image's text rendering to editing tasks. Supports precise bilingual (Chinese and English) text editing, both appearance and semantic edits, and strong results across benchmarks.
—
¥0.163/ 枚
—
—
Qwen
Unifies generation and editing in one model; native 2K; professional typography and infographics.
—
¥0.1649/ 枚
—
—
Fuses generation and editing; more professional text rendering and richer realistic textures and scenes.
—
¥0.247/ 枚
—
—
ByteDance
New unified architecture combining generation and editing; knowledge-based output and reference consistency; up to 4K.
—
¥0.1907/ 枚
—
—
ByteDance
Broad upgrade via model scaling; sharper subject detection, stricter reference fidelity, stronger dense text rendering.
—
¥0.2322/ 枚
—
—
ByteDance
Adds deep reasoning; better interprets complex prompts with common-sense knowledge; stronger edit consistency.
—
¥0.2044/ 枚
—
—
MiniMax
Speed-optimized variant of M2.7 with identical outputs, tuned for low-latency coding and agent tool-calling workloads.
¥3.4119/ M
¥13.6476/ M
¥0.3412/ M
¥2.1324/ M
Qwen
The lightweight vision-language reasoning model of the Qwen3.7 series, strong at object recognition, spatial understanding and real-world visual perception, built for multimodal agents, visual coding, search and computer use.
¥0.2007/ M
¥0.8697/ M
¥0.0401/ M
¥0.2542/ M
MiniMax
A high-throughput variant of M2.5 with matching quality at roughly triple the speed, tuned for coding and agentic tool use under latency-sensitive workloads.
¥3.4119/ M
¥13.6476/ M
¥0.1706/ M
¥2.1324/ M
Xiaomi
Mid-tier V2.5 model with native image, audio, and video understanding, built for agent workflows and coding at lower cost than Pro.
¥0.7961/ M
¥1.5922/ M
¥0.2434/ M
¥0/ M
ByteDance
Agent-oriented foundation model tuned for tool calling and complex instruction following, with use cases spanning GUI agents, search agents, and multi-step task orchestration.
¥2.8433/ M
¥22.746/ M
¥0.2843/ M
¥0.0474/ M
ByteDance
Code-focused preview model for agent workflows, with competitive SWE-Bench and LiveCodeBench scores, aimed at autonomous coding agents and multi-language software tasks.
¥5.6865/ M
¥34.119/ M
¥0.5687/ M
¥0.0474/ M
Alibaba
Lightweight mid-tier variant tuned for low-latency agent workflows with coding and tool-calling support.
¥0.1649/ M
¥1.632/ M
¥0/ M
¥0/ M
Alibaba
Open-weight coding model tuned for agentic terminal tasks and repo-scale reasoning with low inference cost.
¥1.4103/ M
¥8.4445/ M
¥0/ M
¥0/ M
Alibaba
Mid-tier Qwen model tuned for multi-step reasoning and agent workflows, compact enough for private and on-prem deployment.
¥0.489/ M
¥3.9123/ M
¥0/ M
¥0/ M
Alibaba
Thinking-mode variant of the dense open-weight model that outputs step-by-step reasoning; at 27B it surpasses the prior open-weight 397B-A17B flagship across coding, math and multi-step reasoning benchmarks.
¥2.346/ M
¥14.0763/ M
¥0/ M
¥0/ M
Qwen
Multimodal model from the Qwen3.5 line with toggleable thinking mode, tuned for image and video understanding, document parsing, and multimodal agents.
¥0.6539/ M
¥3.9123/ M
¥0.3241/ M
¥4.0772/ M
Alibaba
Open-weight mid-tier MoE text model with long chain-of-thought reasoning, strong on function-calling and multi-step planning benchmarks for self-hosted agentic workflows.
¥0.6539/ M
¥5.2145/ M
¥0/ M
¥0/ M
Qwen
Plus-tier vision-language model in the Qwen3-VL line, handling text, image and video inputs with strengths in document parsing, video understanding, spatial grounding and agent tool use.
¥0.8189/ M
¥8.1544/ M
¥0.1649/ M
¥3.065/ M
Z.AI
Text-only model in the GLM-4 family with improved coding benchmarks and tool calls during reasoning, suited for software development and agent workflows.
¥3.2641/ M
¥13.0448/ M
¥0.6539/ M
¥0/ M
Kling
Kling O1 是为创作者打造的全新创意引擎,解锁无限创意可能。基于多模态视觉语言(MVL)概念,使用自然语言结合视频、图片、元素等多模态描述,精准理解您的创作意图。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720p | ¥0.3883/ s | ¥0.3883/ s |
| 1080p | ¥0.5177/ s | ¥0.5177/ s |
Kling
可灵3.0 Omni是旗舰模型,提供卓越的视频质量、先进的语义理解和增强的伪影修正。基于视频3.0的基础,Omni为专业创作者提供极致品质,具有更精细的细节、更流畅的运动动态和更高的成功率。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720p | ¥0.3883/ s | ¥0.3883/ s |
| 1080p | ¥0.5177/ s | ¥0.5177/ s |
Kling
可灵视频 3.0 配备革命性智能分镜系统,自动编排多镜头序列,实现电影级转场效果。单次生成最长15秒视频,具备增强的角色一致性、原生级文字叠加和多语言音频混合能力。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720p | ¥0.3883/ s | ¥0.3883/ s |
| 1080p | ¥0.5177/ s | ¥0.5177/ s |
| 4k | ¥1.9413/ s | ¥1.9413/ s |
| 2k | ¥0.6471/ s | ¥0.6471/ s |
Kling
VIDEO 2.6 模型一次生成视觉画面、自然配音、匹配音效和环境氛围,连接「声音」与「视觉」的世界。无论是输入文字还是上传图片,都能即时创建一个完整的动态视频,包含声音、节奏和沉浸感。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720p | ¥0.1941/ s | ¥0.1941/ s |
| 1080p | ¥0.3236/ s | ¥0.3236/ s |
Kling
Kling 2.5 Turbo 针对速度和性价比进行了优化,非常适合大批量视频生成需求。在显著缩短生成时间的同时保持良好的质量。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720p | ¥0.1941/ s | ¥0.1941/ s |
| 1080p | ¥0.3236/ s | ¥0.3236/ s |
Kling
Kling 2.1 显著提升了语义响应、动态质量和画面质感。是稳定、高质量视频生成的首选。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720p | ¥0.2588/ s | ¥0.2588/ s |
| 1080p | ¥0.453/ s | ¥0.453/ s |
Kling
Kling 2.0 大师版针对专业创作场景优化,为高要求项目提供电影级输出质量。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720p | ¥0.2588/ s | ¥0.2588/ s |
| 1080p | ¥0.453/ s | ¥0.453/ s |
Alibaba
Alibaba のモデル。Video などに対応。標準 API で呼び出し可能。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720 | ¥0.3345/ s | ¥0.3345/ s |
| 1080 | ¥0.5018/ s | ¥0.5018/ s |
Alibaba
Alibaba のモデル。Video などに対応。標準 API で呼び出し可能。
| 解像度 | 動画入力なし | 動画入力あり |
|---|---|---|
| 720 | ¥0.669/ s | ¥0.669/ s |
| 1080 | ¥1.0035/ s | ¥1.0035/ s |
Qwen
Qwen 3.6 lightweight tier with text, image, and video input. Low latency and low cost, well-suited for high-volume classification, extraction, summarization, and simple agent workflows.
¥0.8284/ M
¥4.9729/ M
¥0.0824/ M
¥1.0359/ M
Moonshot AI
Instruction-tuned for agent workflows with reliable tool calling and multilingual coding.
¥3.2641/ M
¥13.079/ M
¥0.6539/ M
¥0/ M
Moonshot AI
Open-source thinking model built for long-horizon reasoning and hundreds of sequential tool calls across autonomous research, coding, and agent workflows.
¥3.4119/ M
¥14.2163/ M
¥0.853/ M
¥0/ M
Qwen
Enhanced Qwen model with strong bilingual (Chinese/English) comprehension, excelling at long-document analysis and structured output.
¥1.3835/ M
¥8.3007/ M
¥0.1383/ M
¥1.7293/ M
Xiaomi
Flagship agent model for complex workflows
¥11.373/ M
¥34.119/ M
¥2.2746/ M
¥0/ M
MiniMax
Multi-agent model for complex workflows
¥1.4526/ M
¥5.8105/ M
¥0.2905/ M
¥0/ M
Z.AI
Balanced model with strong Chinese understanding.
¥3.355/ M
¥13.3633/ M
¥3.1276/ M
¥0/ M
Z.AI
Multimodal model with strong image + text understanding.
¥1.706/ M
¥5.1179/ M
¥0.3128/ M
¥0/ M
MiniMax
Cost-effective general model for balanced speed and quality.
¥1.7628/ M
¥6.9944/ M
¥0.1763/ M
¥2.2177/ M
Z.AI
Designed for agent-based workflows such as OpenClaw.
¥6.8238/ M
¥22.746/ M
¥1.3648/ M
¥0/ M
Moonshot AI
Top Chinese-English bilingual. Strong long context.
¥3.4119/ M
¥17.1221/ M
¥0.853/ M
¥4.0829/ M
Z.AI
Competitive general-purpose Chinese model.
¥5.0041/ M
¥18.3674/ M
¥0.9781/ M
¥0/ M
Z.AI
Ultra-fast for Chinese. Very cheap.
¥0.5687/ M
¥2.4452/ M
¥0.0569/ M
¥0/ M