Beyond Decodability: Reconstructing Language Model Representations with an Encoding Probe
1 directory member surfaced this signal.
“Accepted to EMNLP 2026: "Beyond Decodability: Reconstructing Language Model Representations with an Encoding Probe". Probing asks what you can decode from a model's activations. We reverse it: reconstruct activations from interpretable features, then ablate…”