Skip to main content
株式会社オブライト
Services
About
Company
Column
Glossary
Pricing
Free Tools
Contact
日本語
日本語
メニューを開く
Column
Qwen 3.8
Articles tagged "Qwen 3.8"
3 articles
AI
2026-08-31
Qwen3.8-Flash-Next Requirements — VRAM 75GB to 354GB by Quantization [125B MoE, Runs on a Single RTX 4090, Aug 2026]
Qwen3.8-Flash-Next needs roughly 75GB (1-bit) to 354GB (BF16) of memory depending on quantization. This 125B-total/6B-active open-weight MoE has been run on a single 24GB RTX 4090 via MoE expert offloading, while the official vLLM/SGLang FP8 recipe needs ~250GB across multiple datacenter GPUs. VRAM and GPU tables inside. Updated August 2026.
Qwen 3.8
Requirements
VRAM
AI
2026-08-15
Qwen3.8-27B System Requirements — VRAM 9–56GB, Apache-2.0 Licensed [2026]
Qwen3.8-27B weights landed on August 15, 2026. This at-a-glance requirements guide maps VRAM needs to the actual published file sizes: Q4_K_M is 17.1GB and won't fit a 16GB GPU, making IQ4_XS (15.7GB) the practical floor. Licensed Apache-2.0 for commercial use.
Qwen 3.8
Requirements
VRAM
AI
2026-08-04
Qwen3.8 Max: 2.4T MoE, 95B Active, Pricing & Open Weights
Qwen3.8 Max: Alibaba's 2.4T sparse MoE, ~95B active/token. SWE-bench 87.3%. API from $2.00/$6.00/M tokens (in/out). Open weights promised. Updated Aug. 2026.
Qwen 3.8
Alibaba
MoE