File size: 345 Bytes
aaf278c 52f801e b6f2f7e 52f801e b6f2f7e 52f801e |
1 2 3 4 5 6 7 8 9 10 11 |
---
language:
- ms
- en
base_model: HuggingFaceTB/SmolLM-360M
---
# HuggingFaceTB/SmolLM-360M continue pretraining on Malaysian context dataset
Continue pretraining on 50B tokens, dataset prepared at https://github.com/malaysia-ai/pretrain-text-dataset/tree/main/smollm
Wandb at https://wandb.ai/huseinzol05/finetune-HuggingFaceTB-SmolLM-360M/ |