Dongwei
/

DeepSeek-R1-Distill-Qwen-7B-GRPO_Math

Text Generation

Generated from Trainer

text-generation-inference

Inference Endpoints

Model card Files Files and versions Community

DeepSeek-R1-Distill-Qwen-7B-GRPO_Math / generation_config.json

Dongwei's picture

Model save

6c50d3f verified 12 days ago

186 Bytes

	{
	"_from_model_config": true,
	"bos_token_id": 151646,
	"do_sample": true,
	"eos_token_id": 151643,
	"temperature": 0.6,
	"top_p": 0.95,
	"transformers_version": "4.49.0.dev0"
	}