-
Open-Reasoner-Zero/Open-Reasoner-Zero-32B
Reinforcement Learning • 33B • Updated • 187 • 33 -
Open-Reasoner-Zero/Open-Reasoner-Zero-7B
Reinforcement Learning • 8B • Updated • 1.05k • 34 -
Open-Reasoner-Zero/Open-Reasoner-Zero-1.5B
Reinforcement Learning • 2B • Updated • 968 -
Open-Reasoner-Zero/Open-Reasoner-Zero-0.5B
Reinforcement Learning • 0.5B • Updated • 300
AI & ML interests
Scale up the Reasoner-Zero Training
Organization Card
Welcome to Open-Reasoner-Zero!
Please check our GitHub!
-
Open-Reasoner-Zero/Open-Reasoner-Zero-32B
Reinforcement Learning • 33B • Updated • 187 • 33 -
Open-Reasoner-Zero/Open-Reasoner-Zero-7B
Reinforcement Learning • 8B • Updated • 1.05k • 34 -
Open-Reasoner-Zero/Open-Reasoner-Zero-1.5B
Reinforcement Learning • 2B • Updated • 968 -
Open-Reasoner-Zero/Open-Reasoner-Zero-0.5B
Reinforcement Learning • 0.5B • Updated • 300
models 9
Open-Reasoner-Zero/ORZ-R1-Distill-Qwen-14B
15B • Updated • 5 • 2
Open-Reasoner-Zero/Open-Reasoner-Zero-Critic-32B
Reinforcement Learning • 32B • Updated • 9 • 7
Open-Reasoner-Zero/Open-Reasoner-Zero-Critic-7B
Reinforcement Learning • 7B • Updated • 10 • 1
Open-Reasoner-Zero/Open-Reasoner-Zero-Critic-0.5B
Reinforcement Learning • 0.5B • Updated • 6
Open-Reasoner-Zero/Open-Reasoner-Zero-7B
Reinforcement Learning • 8B • Updated • 1.05k • 34
Open-Reasoner-Zero/Open-Reasoner-Zero-32B
Reinforcement Learning • 33B • Updated • 187 • 33
Open-Reasoner-Zero/Open-Reasoner-Zero-0.5B
Reinforcement Learning • 0.5B • Updated • 300
Open-Reasoner-Zero/Open-Reasoner-Zero-Critic-1.5B
Reinforcement Learning • 2B • Updated • 3 • 1
Open-Reasoner-Zero/Open-Reasoner-Zero-1.5B
Reinforcement Learning • 2B • Updated • 968