The poster complains about misleading content in r/LocalLLM, citing posts that falsely claim high performance from running Qwen 3.8 27B on low‑end hardware with Q1s quantization and exaggerated savings from running ten agents at 80 t/s each. They argue that such “BS” posts, often used to share GitHub links without disclosing quantization details, should be stopped and that moderators should require accurate quant information in titles.
Read original
reddit/r/LocalLLM