DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture.
作者DeepSeek
国家/地区🇨🇳 China
发布日期2026年9月10日
概览
费用比率
0.1x
上下文
1M
输入
输出
推理
工具调用
结构化输出
定价/M tokens
输入$0.15
缓存输入$0.00
输出$0.60
上下文
上下文长度1,048,576
最大输出令牌384,000
知识截止日期-
性能
吞吐量219 tok/s
延迟0.99s
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
文档分析#37软件与信息技术服务#104数学#104写作、文学与语言#113
作者Anthropic
国家/地区🇺🇸 United States
发布日期2025年10月15日
概览
费用比率
1x
上下文
200K
输入
输出
推理
工具调用
结构化输出
定价/M tokens
输入$1.00
缓存输入$0.10
输出$5.00
上下文
上下文长度200,000
最大输出令牌64,000
知识截止日期-
性能
吞吐量84 tok/s
延迟0.68s
基准测试
推理、编程和智能体指数的对比得分,以及各领域排名。
智能
53.4
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)
52.7
GPT-6 Astra (max)
50.8