Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Edit Models filters
Main
Tasks
Libraries
Languages
Licenses
Other
Tasks
Text Generation
Any-to-Any
Image-Text-to-Text
Image-to-Text
Image-to-Image
Text-to-Image
Text-to-Video
Text-to-Speech
+ 44
Parameters
Reset Parameters
< 1B
6B
12B
32B
128B
> 500B
< 1B
> 500B
Libraries
PyTorch
google-tensorflow
TensorFlow
JAX
Transformers
Diffusers
GGUF
MLX
Transformers.js
Safetensors
+ 45
+ 47
+ 44
Apps
vLLM
llama.cpp
MLX LM
LM Studio
Ollama
Jan
Draw Things
DiffusionBee
JoyFusion
+ 8
+ 10
Inference Providers
Groq
Novita
Cerebras
Nscale
fal
Together AI
Fireworks
Featherless AI
Zai
+ 9
+ 11
+ 10
Hardware
Add your hardware
Apply filters
Models
11
Base only
Inference Available
Inference
Add filters
Sort: Trending
asmitamohanty/reflect-rl-rpo-model
Updated
Dec 1, 2025
Jinhe/ReflectRL-Qwen2.5-Math-7B-GRPO
Text Generation
•
8B
•
Updated
Aug 6
•
17
Jinhe/ReflectRL-Qwen2.5-Math-7B-GRPO-ReflectRL
Text Generation
•
8B
•
Updated
Aug 6
•
24
Jinhe/ReflectRL-Qwen2.5-Math-7B-DAPO
Text Generation
•
8B
•
Updated
Aug 6
•
24
Jinhe/ReflectRL-Qwen2.5-Math-7B-DAPO-ReflectRL
Text Generation
•
8B
•
Updated
Aug 6
•
19
Jinhe/ReflectRL-Llama-3.1-8B-Instruct-GRPO
Text Generation
•
8B
•
Updated
Aug 6
•
17
Jinhe/ReflectRL-Llama-3.1-8B-Instruct-GRPO-ReflectRL
Text Generation
•
8B
•
Updated
Aug 6
•
16
Jinhe/ReflectRL-Qwen2.5-1.5B-Instruct-GRPO
Text Generation
•
2B
•
Updated
Aug 6
•
91
Jinhe/ReflectRL-Qwen2.5-1.5B-Instruct-GRPO-ReflectRL
Text Generation
•
2B
•
Updated
Aug 6
•
17
Jinhe/ReflectRL-Qwen2.5-3B-Instruct-GRPO
Text Generation
•
3B
•
Updated
Aug 6
•
96
Jinhe/ReflectRL-Qwen2.5-3B-Instruct-GRPO-ReflectRL
Text Generation
•
3B
•
Updated
Aug 6
•
19