GlitchProber: Advancing Effective Detection and Mitigation of Glitch Tokens in Large Language Models

Large language models (LLMs) have achieved unprecedented success in the field of natural language processing. However, the black-box nature of their internal mechanisms has brought many concerns about their trustworthiness and interpretability. Recent research has discovered a class of abnormal toke...

Celý popis

Uložené v:

Podrobná bibliografia
Vydané v:	IEEE/ACM International Conference on Automated Software Engineering : [proceedings] s. 643 - 655
Hlavní autori:	Zhang, Zhibo, Bai, Wuxia, Li, Yuxi, Meng, Mark Huasong, Wang, Kailong, Shi, Ling, Li, Li, Wang, Jun, Wang, Haoyu
Médium:	Konferenčný príspevok..
Jazyk:	English
Vydavateľské údaje:	ACM 27.10.2024
Predmet:	Feature extraction Glitch token Large language models LLM analysis LLM security Maintenance engineering Prevention and mitigation Principal component analysis Reliability Software engineering Support vector machines Systematics Vocabulary
ISSN:	2643-1572
On-line prístup:	Získať plný text
Tagy:	Pridať tag Žiadne tagy, Buďte prvý, kto otaguje tento záznam!

Buďte prvý, kto okomentuje tento záznam!