arxiv:2502.09056

An Open Recipe: Adapting Language-Specific LLMs to a Reasoning Model in One Day via Model Merging

Published on Feb 13

· Submitted by

akhaliq on Feb 14

Upvote

Authors:

Kunat Pipatanakul ,

Pittawat Taveekitworachai ,

Potsawee Manakul ,

Abstract

This paper investigates data selection and model merging methodologies aimed at incorporating advanced reasoning capabilities such as those of DeepSeek R1 into language-specific large language models (LLMs), with a particular focus on the Thai LLM. Our goal is to enhance the reasoning capabilities of language-specific LLMs while maintaining their target language abilities. DeepSeek R1 excels in reasoning but primarily benefits high-resource languages such as English and Chinese. However, low-resource languages remain underserved due to the dominance of English-centric training data and model optimizations, which limit performance in these languages. This limitation results in unreliable code-switching and diminished effectiveness on tasks in low-resource languages. Meanwhile, local and regional LLM initiatives have attempted to bridge this gap by developing language-specific LLMs that focus on improving local linguistic fidelity. We demonstrate that, with only publicly available datasets and a computational budget of $120, it is possible to enhance the reasoning capabilities of language-specific LLMs to match the level of DeepSeek R1, without compromising their performance on target language tasks.

View arXiv page View PDF Add to collection

Community

akhaliq

Paper submitter 2 days ago

This comment has been hidden

akhaliq

Paper submitter 2 days ago

kunato

Paper author 2 days ago

This paper explore data selection and model merging to enhance language-specific LLMs (e.g., Thai) with DeepSeek R1-level reasoning. Using only public datasets and a $120 budget, we achieve this without compromising performance on language tasks.

librarian-bot

1 day ago

This is an automated message from the Librarian Bot. I found the following papers similar to this paper.

The following papers were recommended by the Semantic Scholar API

Please give a thumbs up to this comment if you found it helpful!

If you want recommendations for any Paper on Hugging Face checkout this Space

You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.

Tap or paste here to upload images

· Sign up or log in to comment

Upvote

An Open Recipe: Adapting Language-Specific LLMs to a Reasoning Model in One Day via Model Merging

Abstract

Community

Models citing this paper 1

Datasets citing this paper 1

Spaces citing this paper 3

Collections including this paper 2