Ling 3.1 Flash is InclusionAI's (Ant Group) hybrid-reasoning Mixture-of-Experts model (560B total, 25B active parameters, 262K-token context), a much larger successor to the 124B Ling 3.0 Flash. This OpenRouter listing is free of token cost, making it a good fit for trying out strong reasoning, coding, and agentic workloads without spend.
作者InclusionAI
国家/地区🇨🇳 China
发布日期2026年10月2日
概览
费用比率
免费
上下文
262K
输入
输出
推理
工具调用
结构化输出
定价/M tokens
输入$0.00
缓存输入$0.00
输出$0.00
上下文
上下文长度262,144
最大输出令牌32,768
知识截止日期-
性能
吞吐量210 tok/s
延迟1.83s
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
文档分析#37数学#114软件与信息技术服务#116写作、文学与语言#125
作者Anthropic
国家/地区🇺🇸 United States
发布日期2025年10月15日
概览
费用比率
1x
上下文
200K
输入
输出
推理
工具调用
结构化输出
定价/M tokens
输入$1.00
缓存输入$0.10
输出$5.00
上下文
上下文长度200,000
最大输出令牌64,000
知识截止日期-
性能
吞吐量90 tok/s
延迟0.55s
基准测试
推理、编程和智能体指数的对比得分,以及各领域排名。
智能
57.6
Claude Opus 5.5 (Max, Default Fallback)
56.0
Claude Sonnet 5.5 (Max, Default Fallback)
53.4