File size: 345 Bytes
aaf278c
81ab75e
aaf278c
 
 
 
52f801e
b6f2f7e
52f801e
b6f2f7e
52f801e
1
2
3
4
5
6
7
8
9
10
11
---
base_model: HuggingFaceTB/SmolLM-360M
language:
- ms
- en
---
# HuggingFaceTB/SmolLM-360M continue pretraining on Malaysian context dataset

Continue pretraining on 50B tokens, dataset prepared at https://github.com/malaysia-ai/pretrain-text-dataset/tree/main/smollm

Wandb at https://wandb.ai/huseinzol05/finetune-HuggingFaceTB-SmolLM-360M/