publications

Unveiling Decision-Making in LLMs for Text Classification: Extraction of Influential and Interpretable concepts with Sparse Autoencoders” has been accepted at Findings of EACL 2026

Back to news