A user with an RTX 4080 GPU and 32 GB RAM asks whether this hardware is sufficient for running recent local LLMs such as DeepSeek 4.1 Flash or Qwen 3.8 Flash at usable speeds. They seek advice on model suitability before considering alternative setups.

Read original