Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
모달리티
텍스트이미지
입력 / 출력 단가
₩183 / ₩730
1M
컨텍스트
13.11만
출시
2025. 10. 23.
성능 데이터가 아직 없습니다.