Your browser does not seem to support JavaScript. As a result, your viewing experience will be diminished, and you have been placed in read-only mode.
Please download a browser that supports JavaScript, or enable it if it's disabled (i.e. NoScript).
目前用llama.cpp在4090 24G上跑的qwen3.6 27B q4_k_m, q8的kv,120k上下文,跑hermes感觉还行,问下大佬,如果入手4090 48G上FP8版本会有明显提升么?如果提高不大,我就在4090 24G上苟着吧,谢谢……