Ling-3.0-flash-fin is inclusionAI's (Ant Group) finance-tuned variant of Ling-3.0-flash, an open-weight 124B-parameter mixture-of-experts model (about 5.1B active per token, 256K-token context). It is optimized for financial research and multi-step investment workflows with heavy tool use, while retaining the base model's general reasoning, coding, and math abilities.
作者InclusionAI
国家/地区🇨🇳 China
发布日期2026年8月27日
概览
费用比率
0.1x
上下文
262K
输入
输出
推理
工具调用
结构化输出
定价/M tokens
输入$0.06
缓存输入$0.01
输出$0.18
上下文
上下文长度262,144
最大输出令牌235,929
知识截止日期-
性能
吞吐量159 tok/s
延迟2.65s
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
文档分析#37软件与信息技术服务#104数学#104写作、文学与语言#113
作者Anthropic
国家/地区🇺🇸 United States
发布日期2025年10月15日
概览
费用比率
1x
上下文
200K
输入
输出
推理
工具调用
结构化输出
定价/M tokens
输入$1.00
缓存输入$0.10
输出$5.00
上下文
上下文长度200,000
最大输出令牌64,000
知识截止日期-
性能
吞吐量88 tok/s
延迟0.68s
基准测试
推理、编程和智能体指数的对比得分,以及各领域排名。
智能
53.4
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)
52.7
GPT-6 Astra (max)
50.8