Model Card for Model ID

AI ์™€ ๋น…๋ฐ์ดํ„ฐ ๋ถ„์„ ์ „๋ฌธ ๊ธฐ์—…์ธ Linkbricks์˜ ๋ฐ์ดํ„ฐ์‚ฌ์ด์–ธํ‹ฐ์ŠคํŠธ์ธ ์ง€์œค์„ฑ(Saxo) ์ด์‚ฌ๊ฐ€ meta-llama/Meta-Llama-3-8B๋ฅผ ๋ฒ ์ด์Šค๋ชจ๋ธ๋กœ GCP์ƒ์˜ H100-80G 8๊ฐœ๋ฅผ ํ†ตํ•ด SFT-DPO ํ›ˆ๋ จํ•œ ํ•œ๊ธ€ ๊ธฐ๋ฐ˜ LLAMA3-8b 4๊ฐœ์˜ MoE(Mixture of Expert)๋ชจ๋ธ. ํ† ํฌ๋‚˜์ด์ €๋Š” ๋ผ๋งˆ3๋ž‘ ๋™์ผํ•˜๋ฉฐ ํ•œ๊ธ€ VOCA ํ™•์žฅ์€ ํ•˜์ง€ ์•Š์€ ๋ฒ„์ „ ์ž…๋‹ˆ๋‹ค. ์ผ๋ฐ˜์งˆ์˜์‘๋‹ต(์ฑ„ํŒ…)-์˜๋ฃŒ-๊ตฐ์‚ฌ-์ฝ”๋”ฉ ํŠนํ™” LLM์„ ํ†ตํ•ฉ

Dr. Yunsung Ji (Saxo), a data scientist at Linkbricks, a company specializing in AI and big data analytics, trained the meta-llama/Meta-Llama-3-8B base model on 8 H100-60Gs on GCP for 4 hours of instructional training (8000 Tokens). Accelerate, Deepspeed Zero-3 libraries were used.

www.linkbricks.com, www.linkbricks.vc

Downloads last month
8
Safetensors
Model size
24.9B params
Tensor type
BF16
ยท
Inference Providers NEW
This model is not currently available via any of the supported Inference Providers.

Model tree for Saxo/Linkbricks-Horizon-AI-Korean-LLAMA3blend-4x8b

Finetuned
(535)
this model

Dataset used to train Saxo/Linkbricks-Horizon-AI-Korean-LLAMA3blend-4x8b