I
Inception: mercury-2:free
mercury-2:free
128K 컨텍스트50K 출력추론도구캐시구조화시스템 메시지
출시일 Mar 4, 2026업데이트됨 Jul 8, 2026
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving >1,000 tokens/sec on standard GPUs. Mercury 2 is 5x+ faster than leading speed-optimized LLMs like Claude 4.5 Haiku and GPT 5 Mini, at a fraction of the cost. Mercury 2 supports tunable reasoning levels, 128K context, native tool use, and schema-aligned JSON output. Built for coding workflows where latency compounds, real-time voice/search, and agent loops. OpenAI API compatible. Read more in the blog post.
모드 chat토크나이저 Other
요금
입력 가격
$0.00/ 100만 토큰
출력 가격
$0.00/ 100만 토큰
컨텍스트 윈도우 128K 토큰호환 엔드포인트 openai공급자 Inception
가동 시간
성능
성능 데이터 로딩 중...
사용량 및 순위
사용량 불러오는 중...
지원 파라미터
모든 프로바이더 = 이 모델을 제공하는 모든 업스트림에서 지원됩니다. 일부 프로바이더 = 요청을 처리하는 업스트림에 따라 다릅니다. 기본값 = 설정하지 않았을 때 전송되는 값입니다.
| 파라미터 | 프로바이더 | 기본값 |
|---|---|---|
| frequency_penalty | 일부 프로바이더 | 기본적으로 전송되지 않음 |
| include_reasoning | 모든 프로바이더 | - |
| max_tokens | 모든 프로바이더 | - |
| presence_penalty | 일부 프로바이더 | 기본적으로 전송되지 않음 |
| reasoning | 모든 프로바이더 | - |
| reasoning_effort | 모든 프로바이더 | - |
| repetition_penalty | 일부 프로바이더 | 기본적으로 전송되지 않음 |
| response_format | 모든 프로바이더 | - |
| stop | 모든 프로바이더 | - |
| structured_outputs | 모든 프로바이더 | - |
| temperature | 모든 프로바이더 | 0.75 |
| tool_choice | 모든 프로바이더 | - |
| tools | 모든 프로바이더 | - |
| top_k | 일부 프로바이더 | 기본적으로 전송되지 않음 |
| top_p | 일부 프로바이더 | 기본적으로 전송되지 않음 |