Logit-lens and probing analysis show that LLMs encode input and output scripts in early layers but only commit to the target script in the final layers. Intermediate representations frequently default to Latin text.
HOW THIS AFFECTS YOU
●
researcherThis insight suggests that model depth is a critical factor for reliable multilingual script following.