huseinzol05's picture
Update README.md
aaf278c verified
|
raw
history blame
345 Bytes
metadata
language:
  - ms
  - en
base_model: HuggingFaceTB/SmolLM-360M

HuggingFaceTB/SmolLM-360M continue pretraining on Malaysian context dataset

Continue pretraining on 50B tokens, dataset prepared at https://github.com/malaysia-ai/pretrain-text-dataset/tree/main/smollm

Wandb at https://wandb.ai/huseinzol05/finetune-HuggingFaceTB-SmolLM-360M/