On Parameter Efficiency of Neural Language Models @allenai
On Parameter Efficiency of Neural Language Models  @allenai
Uploaded November 2023 | Updated September 2026, 2 days ago
Abstract: Pre-trained neural language models have demonstrated remarkable generalizability in various downstream tasks, such as natural language understanding and question answering. However, these models have grown to contain hundreds of billions of parameters, making them difficult to be deployed in applications with latency requirements and memory constraints. Furthermore, existing research have demonstrated the existence of significant redundant parameters in neural language models. Such redundancy can further compromise their downstream generalizability. To tackle these challenges, my research focus on training neural language models towards higher parameter efficiency and better model generalizability. In this talk, I will introduce three major directions of my research: 1) designing model training algorithms for better parameter utilization, 2) developing pruning and distillation strategies for reliable model compression, and 3) improving cross-task generalizability of parameter-efficient fine-tuning methods.

Bio: Chen Liang is a final-year PhD student from Georgia Institute of Technology, working with Prof. Tuo Zhao in the FLASH research group. Her research interests broadly lie in deep learning and natural language processing, with a major focus on developing methodologies and algorithms to improve model generalizability and parameter efficiency for pre-trained neural language models. Before starting her PhD, Chen received her BS degree in Electrical Engineering from University of Southern California.
link: cliang1453.github.io
On Parameter Efficiency of Neural Language ModelsEnhancing the Reliability and Continual Improvement of Neural Dialogue SystemsMoving Forward by Moving Backward: Embedding Action Impact over Action Semantics | AI2OlmoEarth: Powerful new foundation models and open infrastructure for planetary insightsWhat Do NLP Researchers Believe? Results of the NLP Community MetasurveyAi2 at Google Cloud Next 2025Advancing Generalist Robots: Models, Representations, and DatasetsEvaluating ethical and social risks from large modelsWildDet3D | iphone app demoGetting started with SERA in Claude CodeBenchmarking Compositionality with Formal LanguagesAuto-Formalization for Trustworthy Planning
Ai2 |

On Parameter Efficiency of Neural Language Models

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER