모델 변형·크기 (HF 컬렉션)

모델 변형·크기 (HF 컬렉션)

Qwen 팀은 Hugging Face에 Qwen3-VL 컬렉션을 운영하면서 크기·아키텍처·정밀도별로 다양한 가중치를 공개해요. 대표 변형은 2B부터 235B-A22B까지 걸쳐 있어요.

각 변형은 Instruct(일반 지시 따르기)와 Thinking(추론 강화)으로 나뉘고, 같은 모델에 FP8 양자화 버전과 GGUF 경량 포맷도 제공돼요. 엣지에서 돌리고 싶다면 GGUF, 고성능 서빙이라면 Instruct/Thinking FP8을 고르면 돼요.

출처: https://huggingface.co/collections/Qwen/qwen3-vl

대표 변형

크기 아키텍처 변형
2B Dense Instruct, Thinking, FP8, GGUF
4B Dense Instruct, Thinking, FP8, GGUF
8B Dense Instruct, Thinking, FP8, GGUF
30B-A3B MoE Instruct, Thinking, FP8, GGUF
32B Dense Instruct, Thinking, FP8, GGUF
235B-A22B MoE Instruct, Thinking, FP8, GGUF

활용 팁

  • 개인 데모·노트북: 작은 Dense(2B~8B) Instruct가 무난해요.
  • 서빙 품질 극대화: 235B-A22B 또는 30B-A3B Thinking.
  • 리소스 제약 있는 엣지: GGUF 포맷.

더 알아보기