Related Experiment Video
Updated: Apr 24, 2026

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Aligning large language models across the lifecycle: A survey on safety-usability trade-offs from pre-training to
Zhiqiang Hao1, Hongming Fei2, Chang Liu3
1State Key Laboratory for Novel Software Technology, No. 163 Xianlin Avenue, Nanjing, 210039, China; Software Institute, Nanjing University, No. 22 Hankou Road, Nanjing, 210039, China; Department of ECE, National University of Singapore, 4 Engineering Drive 3, Singapore, 117583, Singapore.
Abstract:
Large language models (LLMs) are increasingly embedded in search, productivity tools, and autonomous agents, where safety failures or degraded utility can propagate across many applications. Yet most alignment techniques are still designed and evaluated in isolation, making it difficult to see how early choices in data, objectives, and optimization interact with later fine-tuning and adaptation. This survey takes a lifecycle view of LLM alignment with the safety-usability trade-off as the organizing lens. We first examine how pre-training data curation, corpus sanitization, privacy protection, and safety-aware objectives shape baseline behavior and memorization risk. We then compare post-training alignment paradigms, including supervised fine-tuning and both RL-based and RL-free preference optimization, such as Reinforcement Learning from Human Feedback (RLHF), Reinforcement Learning from AI Feedback (RLAIF), Constitutional AI (CAI), and Direct Preference Optimization (DPO). Finally, we analyze lightweight adaptation and model editing, including parameter-efficient fine-tuning (PEFT), adapters, knowledge editing, and machine unlearning, as a second front for both eroding and restoring earlier safety guarantees. Across stages, we provide an operational safety-usability ontology, a quantitative synthesis of reported trends, and minimum evaluation checklists for resource-constrained practice. We conclude with open challenges in multimodal and cross-lingual safety, dynamic value pluralism, and providing clearer guarantees for editing and unlearning in real-world pipelines.
Related Concept Videos
Language Development
The critical period for language acquisition suggests that the ability to acquire language is at its peak early in life. As people age, this proficiency decreases. Language development begins very...
Language and Cognition
Improving Translational Accuracy
Improving Translational Accuracy
Language
Corballis and Suddendorf (2007) and Tomasello and Rakoczy (2003) highlight the role of language in...
Components of Language
