FiNE: Fine-grained Neuron-level Model Editing for Reliable and Safe LLMs.

Summary

We introduce Fine-grained Neuron-level Editing (FiNE), a new method for improving Large Language Models (LLMs). FiNE enhances LLM reliability and safety by precisely editing internal network components, leading to more robust and accurate outputs.

Related Concept Videos