CoLLiE: Collaborative Training of Large Language Models in an Efficient Way

Kai Lv; Shuo Zhang; Tianle Gu; Shuhao Xing; Jiawei Hong; Keyu Chen; Xiaoran Liu; Yuqing Yang; Honglin Guo; Tengxiao Liu; Yu Sun; Qipeng Guo; Hang Yan; Xipeng Qiu

doi:10.18653/v1/2023.emnlp-demo.48

CoLLiE: Collaborative Training of Large Language Models in an Efficient Way

Kai Lv, Shuo Zhang, Tianle Gu, Shuhao Xing, Jiawei Hong, Keyu Chen, Xiaoran Liu, Yuqing Yang, Honglin Guo, Tengxiao Liu, Yu Sun, Qipeng Guo, Hang Yan, Xipeng Qiu

Abstract

Large language models (LLMs) are increasingly pivotal in a wide range of natural language processing tasks. Access to pre-trained models, courtesy of the open-source community, has made it possible to adapt these models to specific applications for enhanced performance. However, the substantial resources required for training these models necessitate efficient solutions. This paper introduces CoLLiE, an efficient library that facilitates collaborative training of large language models using 3D parallelism, parameter-efficient fine-tuning (PEFT) methods, and optimizers such as Lion, Adan, Sophia, and LOMO. With its modular design and comprehensive functionality, CoLLiE offers a balanced blend of efficiency, ease of use, and customization. CoLLiE has proven superior training efficiency in comparison with prevalent solutions in pre-training and fine-tuning scenarios. Furthermore, we provide an empirical evaluation of the correlation between model size and GPU memory consumption under different optimization methods, as well as an analysis of the throughput. Lastly, we carry out a comprehensive comparison of various optimizers and PEFT methods within the instruction-tuning context. CoLLiE is available at https://github.com/OpenLMLab/collie.

Anthology ID:: 2023.emnlp-demo.48
Volume:: Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing: System Demonstrations
Month:: December
Year:: 2023
Address:: Singapore
Editors:: Yansong Feng, Els Lefever
Venue:: EMNLP
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 527–542
Language:
URL:: https://aclanthology.org/2023.emnlp-demo.48
DOI:: 10.18653/v1/2023.emnlp-demo.48
Bibkey:
Cite (ACL):: Kai Lv, Shuo Zhang, Tianle Gu, Shuhao Xing, Jiawei Hong, Keyu Chen, Xiaoran Liu, Yuqing Yang, Honglin Guo, Tengxiao Liu, Yu Sun, Qipeng Guo, Hang Yan, and Xipeng Qiu. 2023. CoLLiE: Collaborative Training of Large Language Models in an Efficient Way. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing: System Demonstrations, pages 527–542, Singapore. Association for Computational Linguistics.
Cite (Informal):: CoLLiE: Collaborative Training of Large Language Models in an Efficient Way (Lv et al., EMNLP 2023)
Copy Citation:
PDF:: https://aclanthology.org/2023.emnlp-demo.48.pdf
Video:: https://aclanthology.org/2023.emnlp-demo.48.mp4

PDF Cite Search Video