I

InclusionAI: ling-3.0-flash-vl:free

ling-3.0-flash-vl:free
26.2万 コンテキスト3.3万 出力推論ツール動画キャッシュ構造化
リリース日 Sep 10, 2026更新日 Sep 11, 2026

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers to complete more useful work within constrained token, latency, and serving-cost budgets.

モード chatトークナイザー Other量子化 fp16inclusionAI/Ling-3.0-flash-VL

このモデルのプロバイダーは現在すべて混雑しています

すべての上流プロバイダーがレート制限に達しました。制限が解除されるとモデルは自動的に復帰します。通常は数時間以内です。しばらくしてから再試行するか、別のモデルに切り替えてください。

このモデルをDiscordでリクエスト

料金

入力料金
$0.00/ 100万トークン
出力料金
$0.00/ 100万トークン
コンテキストウィンドウ 262.1K トークン対応エンドポイント -ベンダー InclusionAI

稼働率

パフォーマンス

パフォーマンスデータを読み込み中...

使用状況とランキング

使用状況を読み込み中...

対応パラメータ

すべてのプロバイダー = このモデルを提供するすべてのアップストリームが対応しています。一部のプロバイダー = リクエストを処理するアップストリームによって異なります。デフォルト = 未設定時に送信される値です。

パラメータプロバイダーデフォルト
frequency_penaltyすべてのプロバイダー-
include_reasoningすべてのプロバイダー-
logit_biasすべてのプロバイダー-
max_tokensすべてのプロバイダー-
min_pすべてのプロバイダー-
presence_penaltyすべてのプロバイダー-
reasoningすべてのプロバイダー-
repetition_penaltyすべてのプロバイダー-
response_formatすべてのプロバイダー-
seedすべてのプロバイダー-
stopすべてのプロバイダー-
structured_outputsすべてのプロバイダー-
temperatureすべてのプロバイダー-
tool_choiceすべてのプロバイダー-
toolsすべてのプロバイダー-
top_kすべてのプロバイダー-
top_pすべてのプロバイダー-

よくあるご質問

類似モデル