Qwen3.8-27B FP8: FP8 Quantization, Thinking Control, and 1M Context in Practice
Qwen3.8-27B-FP8 packages a 27B vision-language model as fine-grained FP8 (block size 128) weights. It adds practical controls like `reasoning_effort` and `preserve_thinking`, plus native 262K context with extensibility toward 1M tokens for long-horizon tasks.