상세 보기
IM-BERT: Enhancing Robustness of BERT through the Implicit Euler Method
- Kim, Mihyeon;
- Park, Juhyoung;
- Kim, Youngbin
SCOPUS
2초록
Pre-trained Language Models (PLMs) have achieved remarkable performance on diverse NLP tasks through pre-training and fine-tuning. However, fine-tuning the model with a large number of parameters on limited downstream datasets often leads to vulnerability to adversarial attacks, causing overfitting of the model on standard datasets. To address these issues, we propose IM-BERT from the perspective of a dynamic system by conceptualizing a layer of BERT as a solution of Ordinary Differential Equations (ODEs). Under the situation of initial value perturbation, we analyze the numerical stability of two main numerical ODE solvers: the explicit and implicit Euler approaches. Based on these analyses, we introduce a numerically robust IM-connection incorporating BERT's layers. This strategy enhances the robustness of PLMs against adversarial attacks, even in low-resource scenarios, without introducing additional parameters or adversarial training strategies. Experimental results on the adversarial GLUE (AdvGLUE) dataset validate the robustness of IM-BERT under various conditions. Compared to the original BERT, IM-BERT exhibits a performance improvement of approximately 8.3%p on the AdvGLUE dataset. Furthermore, in low-resource scenarios, IM-BERT outperforms BERT by achieving 5.9%p higher accuracy. © 2024 Association for Computational Linguistics.
- 제목
- IM-BERT: Enhancing Robustness of BERT through the Implicit Euler Method
- 저자
- Kim, Mihyeon; Park, Juhyoung; Kim, Youngbin
- 발행일
- 2024-11
- 유형
- Conference paper
- 저널명
- EMNLP 2024 - 2024 Conference on Empirical Methods in Natural Language Processing, Proceedings of the Conference
- 권
- 2024
- 페이지
- 16217 ~ 16229
- 언어
- ENG
- 출판사
- Association for Computational Linguistics (ACL)
- 분량
- 13 페이지