A

Alibaba: qwen3-next-80b-a3b-instruct:free

qwen3-next-80b-a3b-instruct:free
131.1K 컨텍스트65.5K 출력도구캐시구조화
출시일 Sep 11, 2025지식 기준일 Sep 2025업데이트됨 Jun 18, 2026

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual use, while remaining robust on alignment and formatting. Compared with prior Qwen3 instruct variants, it focuses on higher throughput and stability on ultra-long inputs and multi-turn dialogues, making it well-suited for RAG, tool use, and agentic workflows that require consistent final answers rather than visible chain-of-thought. The model employs scaling-efficient training and decoding to improve parameter efficiency and inference speed, and has been validated on a broad set of public benchmarks where it reaches or approaches larger Qwen3 systems in several categories while outperforming earlier mid-sized baselines. It is best used as a general assistant, code helper, and long-context task solver in production settings where deterministic, instruction-following outputs are preferred.

모드 chat토크나이저 Qwen3양자화 fp8Qwen/Qwen3-Next-80B-A3B-Instruct

요금

입력 가격
$0.00/ 100만 토큰
출력 가격
$0.00/ 100만 토큰
컨텍스트 윈도우 131.1K 토큰호환 엔드포인트 openai공급자 Alibaba

가동 시간

성능

성능 데이터 로딩 중...

사용량 및 순위

사용량 불러오는 중...

지원 파라미터

모든 프로바이더 = 이 모델을 제공하는 모든 업스트림에서 지원됩니다. 일부 프로바이더 = 요청을 처리하는 업스트림에 따라 다릅니다. 기본값 = 설정하지 않았을 때 전송되는 값입니다.

파라미터프로바이더기본값
frequency_penalty모든 프로바이더-
logit_bias일부 프로바이더-
logprobs모든 프로바이더-
max_tokens모든 프로바이더-
min_p일부 프로바이더-
presence_penalty모든 프로바이더-
repetition_penalty모든 프로바이더-
response_format모든 프로바이더-
seed모든 프로바이더-
stop모든 프로바이더-
structured_outputs모든 프로바이더-
temperature모든 프로바이더-
tool_choice모든 프로바이더-
tools모든 프로바이더-
top_k모든 프로바이더-
top_logprobs모든 프로바이더-
top_p모든 프로바이더-

자주 묻는 질문

유사 모델