How bad do you think models like Qwen3.8-27B or GLM-5.3-Flash would be with H-Neurons disabled?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| TL;DR: this paper proposes a method to fix hallucination rates to very low levels or zero by disabling neurons which contribute to hallucination. This discovery has been out for a while now, but it hasn't been that popular, since it kind of lobotomises parts of the LLM. I honestly don't care too much about talking to AI, but instead care about it producing working and good code. I wonder what percentage models would get on e.g. DeepSWE if we found their H-Neurons and disabled them? [link] [comments] |
More from r/LocalLLaMA
-
Apple A20 Pro debuts with 7-core GPU, 32-core Neural Engine and 50% more memory bandwidth (~115 GB/s)
Sep 9
-
Surveillance plagiarism by OpenAI
Sep 9
-
Don't let FOMO win if you're interested in local llm from a hobby/learning aspect
Sep 9
-
Server rebuild to custom loop. 2x RTX Titans 24gb, 1x 22gb 2080ti | T: 70GB VRAM.
Sep 9
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.