ACL · 2024 · Conference paper · Top venue

SAPT: A Shared Attention Framework for Parameter-Efficient Continual Learning of Large Language Models

Weixiang Zhao, Shilong Wang, Yulin Hu, Yanyan Zhao, Bing Qin, Xuanyu Zhang, Qing Yang, Dongliang Xu, Wanxiang Che

Harbin Institute of Technology · XTC (China)

Published in
Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Date
2024-01-01
Citations
52
arXiv
2401.08295
Cite
@inproceedings{zhao2024sapt,
  title = {SAPT: A Shared Attention Framework for Parameter-Efficient Continual Learning of Large Language Models},
  author = {Weixiang Zhao and Shilong Wang and Yulin Hu and Yanyan Zhao and Bing Qin and Xuanyu Zhang and Qing Yang and Dongliang Xu and Wanxiang Che},
  year = {2024},
  booktitle = {Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)},
  eprint = {2401.08295},
  archivePrefix = {arXiv},
  doi = {10.18653/v1/2024.acl-long.625},
  url = {https://arxiv.org/abs/2401.08295},
}