yufanzh

qwen3-1p7b-rl-checkpoints-20260926

reinforcement-learningtransformers

huggingface.co/yufanzh/qwen3-1p7b-rl-checkpoints-20260926

Updated 2026-09-26 ·Open on Hugging Face →
transformers · safetensors · reinforcement-learning · base_model:Qwen/Qwen3-1.7B-Base · base_model:finetune:Qwen/Qwen3-1.7B-Base · endpoints_compatible · region:us
The README has not been fetched yet (metadata-first ingest). The model page and the metadata below are live.
Mirrored from the Hugging Face Hub and served from the Conceptio Open Knowledge Archive. Read the original card at https://huggingface.co/yufanzh/qwen3-1p7b-rl-checkpoints-20260926.