Skip to main navigation Skip to search Skip to main content

A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models

  • Dong Shu
  • , Xuansheng Wu
  • , Haiyan Zhao
  • , Daking Rai
  • , Ziyu Yao
  • , Ninghao Liu
  • , Mengnan Du

Research output: Chapter in book / Conference proceedingConference article published in proceeding or bookAcademic researchpeer-review

Fingerprint

Dive into the research topics of 'A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models'. Together they form a unique fingerprint.
Sort by

Keyphrases

Computer Science