Prototie AI v2 โ ํ๊ตญ์ด ์ถ๋ก ํนํ AI (K-AI ๋ฆฌ๋๋ณด๋ #1 ๋ชฉํ)
Qwen3-32B ๊ธฐ๋ฐ 3๋จ๊ณ(SFTโGRPOโDARE-TIES) ํ์ต ํ์ดํ๋ผ์ธ์ผ๋ก ๊ฐ๋ฐ๋ ํ๊ตญ์ด AI์ ๋๋ค.
ํ์ต ํ์ดํ๋ผ์ธ v2
- STEP1 SFT: KMMLU + CLIcK + ํ๊ตญ์ด ์ํ 25K + GSM8K + ๋ ผ๋ฆฌ์ถ๋ก ํฉ์ฑ (Qwen3-32B, LoRA r=64)
- STEP2 GRPO: 5์ค ๋ณด์ ํจ์ ๊ฐํํ์ต (accuracy / korean / format / reasoning_depth / kmmlu_format)
- STEP3 DARE-TIES: GRPO(w=0.75) + SFT(w=0.25) ๋ชจ๋ธ ๋ณํฉ โ ์ถ๋ก ๋น์ค ๊ฐํ
v2 ์ฃผ์ ๊ฐ์ ์
| ํญ๋ชฉ | v1 | v2 |
|---|---|---|
| ๋ฒ ์ด์ค ๋ชจ๋ธ | Qwen3-14B | Qwen3-32B (+128% ํ๋ผ๋ฏธํฐ) |
| GRPO ๋ณด์ ํจ์ | 3์ค | 5์ค (+reasoning_depth, +kmmlu_format) |
| DARE-TIES ๋น์ค | GRPO 0.70 | GRPO 0.75 (์ถ๋ก ๊ฐํ) |
| ํ์ต ๋ฐ์ดํฐ | ๋ฒ์ฉ | KMMLU-Pro + MuSR-style + HLE-style ์ถ๊ฐ |
์ฌ์ฉ ์์
from transformers import AutoModelForCausalLM, AutoTokenizer
import torch
tokenizer = AutoTokenizer.from_pretrained("prototie/prototie-ai-final-v2", trust_remote_code=True)
model = AutoModelForCausalLM.from_pretrained(
"prototie/prototie-ai-final-v2",
torch_dtype=torch.bfloat16, device_map="auto", trust_remote_code=True,
)
ํ๊ฐ ๋์ ๋ฒค์น๋งํฌ (K-AI ๋ฆฌ๋๋ณด๋)
- KMMLU-Pro: ์ ๋ฌธ์ง ๊ณ ๋๋ ์ง์ (๋ฒ๋ฅ , ์ํ, ๊ณตํ ๋ฑ)
- CLIcK: ํ๊ตญ ๋ฌธํยท์ธ์ดยท์์
- HLE(Ko): ๊ณ ๋๋ ์ถ๋ก
- MuSR(Ko): ๋ค๋จ๊ณ ๋ ผ๋ฆฌ ์ถ๋ก
- Com2-main(ko): ์์ ์ถ๋ก
- Downloads last month
- 1
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support
Model tree for prototie/prototie-ai-final-v2
Base model
Qwen/Qwen3-32B