A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models
- Dong Shu
- , Xuansheng Wu
- , Haiyan Zhao
- , Daking Rai
- , Ziyu Yao
- , Ninghao Liu
- , Mengnan Du
Research output: Chapter in book / Conference proceeding › Conference article published in proceeding or book › Academic research › peer-review
10
Link opens in a new tab
Citations
(Scopus)