Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
hanxiao
hahnli
Follow
AI & ML interests
Reinforcement Learning, RLHF, LLM, Human-AI colaboration
Recent Activity
updated
a model
about 13 hours ago
hahnli/gemma-3-4b-Search-Cheating-Agent
published
a model
about 13 hours ago
hahnli/gemma-3-4b-Search-Cheating-Agent
updated
a model
about 13 hours ago
hahnli/gemma-3-4b-AgenticRoles-SelfMonitor-RL
View all activity
Organizations
None yet
models
34
Sort: Recently updated
hahnli/gemma-3-4b-Search-Cheating-Agent
5B
•
Updated
about 13 hours ago
hahnli/gemma-3-4b-AgenticRoles-SelfMonitor-RL
5B
•
Updated
about 13 hours ago
hahnli/gemma-3-4b-AgenticRoles-RLCM
5B
•
Updated
2 days ago
•
9
hahnli/gemma-3-4b-AgenticRoles-RLVM
5B
•
Updated
3 days ago
•
7
hahnli/Qwen3-8B-AgenticRoles-SelfMonitor-RL
8B
•
Updated
about 1 month ago
•
3
hahnli/Qwen3-8B-AgenticRoles-RLVM
8B
•
Updated
about 1 month ago
•
5
hahnli/Qwen3-8B-AgenticRoles-MixedRL
8B
•
Updated
about 1 month ago
•
3
hahnli/Qwen3-8B-AgenticRoles-RL
8B
•
Updated
about 1 month ago
•
3
hahnli/Qwen3-8B-Search-SelfMonitor-RL
8B
•
Updated
Jul 18
•
2
hahnli/Qwen3-4B-Search-SelfMonitor-RL
4B
•
Updated
Jul 18
•
2
View 34 models
datasets
1
hahnli/PKU-SafeRLHF-Filter
Viewer
•
Updated
Aug 8, 2025
•
62.4k
•
8