description [ICLR 2026][Interpretability][Sparse Autoencoders] Standard Sparse Autoencoders (SAEs) learned on aligned multimodal embeddings (like CLIP/CLAP) produce "split dictionaries"—where most ...