LLM performs tokenization, but if LLM isn't trained on texts from a specific language, it will incorrectly tokenize the text when parsing the prompt and generating the response text. Therefore, in DeepSeek, for example, you might see phrases like:
In the Russian-speaking community, volunteer communities like Ruadaptnaya room and Vikhr models are responsible for adapting models to Russian. However, when retraining models to Russian, they almost inevitably lose some of their skills.