arXiv cs.CLPaper
The First Token Is a Clue: Verbalizing Multi-Token Concepts from the J-lens
This is solid interpretability work but aimed at a narrow audience: researchers building lens methods for LLM analysis. The finding that first tokens carry enough signal to recover multi-token concepts is interesting for mechanistic understanding, but doesn't change how builders or operators use models. Only read if you're actively working on interpretability infrastructure.