I

InclusionAI: ling-3.0-flash-fin:free

ling-3.0-flash-fin:free
26.2万 上下文3.3万 输出推理工具缓存结构化
发布日期 Jul 23, 2026更新时间 Aug 31, 2026

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers to complete more useful work within constrained token, latency, and serving-cost budgets.

模式 chat分词器 Other量化 bf16inclusionAI/Ling-3.0-flash

价格

输入价格
$0.00/ 100 万 token
输出价格
$0.00/ 100 万 token
上下文窗口 262.1K Token兼容端点 openai供应商 InclusionAI

在线时长

性能

正在加载性能数据...

使用量与排名

正在加载使用量……

支持的参数

所有提供方:为该模型提供服务的每个上游都支持。部分提供方:取决于处理请求的上游。默认值:未设置时发送的值。

参数提供方默认值
frequency_penalty所有提供方-
include_reasoning所有提供方-
logit_bias所有提供方-
logprobs所有提供方-
max_tokens所有提供方-
min_p所有提供方-
presence_penalty所有提供方-
reasoning所有提供方-
repetition_penalty所有提供方-
response_format所有提供方-
seed所有提供方-
stop所有提供方-
temperature所有提供方-
tool_choice所有提供方-
tools所有提供方-
top_k所有提供方-
top_logprobs所有提供方-
top_p所有提供方-

常见问题

相似模型