← 所有模型

Ling-3.0-flash

🇨🇳 InclusionAI · Ling-3.0-flash

输入价格 $0.021 每百万 tokens NT$0.67
输出价格 $0.063 每百万 tokens NT$2.0
Context Window 262K tokens 输出上限 33K
OpenRouter 路由价 请以官方定价页为准
通过 OpenRouter 使用此模型 →

概览

Ling-3.0-flash 是 InclusionAI 推出的大型语言模型 API,属于其 Ling-3.0-flash 系列。输入每百万 token $0.021、输出每百万 token $0.063,定位在预算型区间,是高吞吐或成本敏感工作负载中较便宜的选择之一。输出 token 的成本约为输入的 3 倍,因此以提示为主的工作负载会比以生成为主的明显便宜。较大的 262K token 上下文窗口(约 393 页文字)让它能在单次调用中读入整本书、大型代码库或冗长逐字稿。在 Artificial Analysis 的 Intelligence Index 上得分为 38(B 级),可作为其整体推理能力相对于本站其他模型的参考指标。本页所有价格反映的是 OpenRouter 的路由费率,每日自动同步;正式投入生产前,请以提供商官方定价为准。

维度 单位 价格 (USD) 价格 (TWD) 有效自
输入 每 1M tokens $0.021 NT$0.67
输出 每 1M tokens $0.063 NT$2.0
缓存读取 每 1M tokens $0.0042 NT$0.13

提供商
InclusionAI
模型家族
Ling-3.0-flash
版本字符串
inclusionai/ling-3.0-flash
状态
使用中
模态
文字
Context Window
262,144 tokens
输出上限
32,768 tokens

综合指标

由 Artificial Analysis 评估的跨领域能力指数 — Artificial Analysis

Agentic Index 29 C 测量于 2026-08-26
Coding Index 51 A 测量于 2026-08-26
Intelligence Index 38 B 测量于 2026-08-26

Benchmark 分数

数据来源:Artificial Analysis

AA-LCR 67.0% B 测量于 2026-08-26
GPQA Diamond 85.5% S 测量于 2026-08-26
Non-Hallucination 56.0% 测量于 2026-08-26
Omniscience Accuracy 18.2% 测量于 2026-08-26
SciCode 41.1% B 测量于 2026-08-26

效能指标

实测数据,由 Artificial Analysis 每 72 小时更新 — Artificial Analysis

首 Token 延迟 2.4s 测量于 2026-08-26
输出速度 383 t/s 测量于 2026-08-26
回应时间 9.0s 测量于 2026-08-26

过去 90 天价格走势

输入 / 输出价格(USD per 1M tokens)

过去 90 天记录;每次价格变动会在此呈现

日期 维度 价格 (USD) 来源
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter
输入 $0.021 OpenRouter
缓存读取 $0.0042 OpenRouter
输出 $0.063 OpenRouter

描述

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

重点摘要

以下为本页面的关键数据,供快速参考与引用。

  • Ling-3.0-flash input 价格为 $0.021/M tokens
  • Ling-3.0-flash output 价格为 $0.063/M tokens
  • Context window:262,144 tokens
  • 提供商:InclusionAI
  • 模型家族:Ling-3.0-flash
  • 支持模态:文字
  • 数据来源:OpenRouter,每日自动更新