A user is seeking advice on running the Qwen3.8-27B model across a 24GB VRAM setup (RTX 5080 16GB + RTX 3070 8GB) for lightweight coding tasks such as text edits and style tweaks. They are a beginner looking for the optimal quantization and configuration settings to maximize performance within their hardware constraints.

Read original