The author inserted factual knowledge directly into the weights of Llama‑3.1‑8B by manually setting specific neuron weights in an appended MLP region, avoiding any fine‑tuning, LoRA, or RAG, and created a visualizer that maps each fact’s location in the model. This method preserves the original model weights while enabling precise recall of the injected facts.
Read original
reddit/r/LocalLLM