I

Inception: mercury-2:free

mercury-2:free
128K הקשר50K פלטחשיבהכליםמטמוןמובנההודעת מערכת
שוחרר Mar 4, 2026עודכן Jul 8, 2026

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving >1,000 tokens/sec on standard GPUs. Mercury 2 is 5x+ faster than leading speed-optimized LLMs like Claude 4.5 Haiku and GPT 5 Mini, at a fraction of the cost. Mercury 2 supports tunable reasoning levels, 128K context, native tool use, and schema-aligned JSON output. Built for coding workflows where latency compounds, real-time voice/search, and agent loops. OpenAI API compatible. Read more in the blog post.

מצב chatמטקנן Other

תמחור

מחיר קלט
$0.00/ מיליון אסימונים
מחיר פלט
$0.00/ מיליון אסימונים
חלון הקשר 128K טוקניםנקודות קצה תואמות openaiספק Inception

זמן פעילות

ביצועים

טוען נתוני ביצועים...

שימוש ודירוג

טוען שימוש...

פרמטרים נתמכים

כל הספקים = נתמך אצל כל upstream שמגיש את המודל הזה. חלק מהספקים = תלוי ב-upstream שמטפל בבקשה. ברירת מחדל = הערך שנשלח כשלא הגדרת דבר.

פרמטרספקיםברירת מחדל
frequency_penaltyחלק מהספקיםלא נשלח כברירת מחדל
include_reasoningכל הספקים-
max_tokensכל הספקים-
presence_penaltyחלק מהספקיםלא נשלח כברירת מחדל
reasoningכל הספקים-
reasoning_effortכל הספקים-
repetition_penaltyחלק מהספקיםלא נשלח כברירת מחדל
response_formatכל הספקים-
stopכל הספקים-
structured_outputsכל הספקים-
temperatureכל הספקים0.75
tool_choiceכל הספקים-
toolsכל הספקים-
top_kחלק מהספקיםלא נשלח כברירת מחדל
top_pחלק מהספקיםלא נשלח כברירת מחדל

שאלות נפוצות

מודלים דומים