Application of Sparse Autoencoders to Enhance Mechanistic Interpretability of Large Language Models in Medicine
{{output}}
Large language models (LLMs) are being increasingly incorporated into clinical workflows due to their ability to synthesize medical knowledge and support diagnosis and treatment planning. However, their opaque internal decision-making processes limit trust, re... ...