M
Meta: llama-3.2-11b-vision:free
llama-3.2-11b-vision:free
128K 컨텍스트8.2K 출력도구비전
출시일 Sep 25, 2024업데이트됨 Jun 18, 2026
Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and visual question answering, bridging the gap between language generation and visual reasoning. Pre-trained on a massive dataset of image-text pairs, it performs well in complex, high-accuracy image analysis. Its ability to integrate visual understanding with language processing makes it an ideal solution for industries requiring comprehensive visual-linguistic AI applications, such as content creation, AI-driven customer service, and research. Click here for the original model card. Usage of this model is subject to Meta's Acceptable Use Policy.
모드 chat
요금
입력 가격
$0.00/ 100만 토큰
출력 가격
$0.00/ 100만 토큰
컨텍스트 윈도우 128K 토큰호환 엔드포인트 openai공급자 Meta
가동 시간
성능
성능 데이터 로딩 중...
사용량 및 순위
사용량 불러오는 중...