Posts by Jouni Luoma

Enabling Language-specific Reasoning in Multilingual Models with Reinforcement Learning

We introduce the Poro 2 Long family of models, a follow-up to the Poro 2 family that demonstrated exceptional performance in Finnish and English instruction-following and conversational tasks. This model family demonstrates our progress in developing models capable of reasoning in the same language as the user. A prerequisite for an effective reasoning model is the ability to process longer sequences beyond the sequence lengths used in pretraining; therefore, the Poro 2 Long models feature a context window of 128k tokens.

Read more ...


Continued Pretraining: A Practical Playbook for Language-Specific LLM Adaptation

What if you could make a state-of-the-art LLM fluent in a new language—without training from scratch? In this guide, we show how we did just that with Finnish.

Read more ...