Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Karthik Iyer
kiyer97
11
9
Follow
AI & ML interests
Reinforcement learning, reward modeling, RLHF, policy optimization, offline RL
Recent Activity
upvoted
a
paper
about 15 hours ago
SenseNova-U1.5: Towards Native Unified Visual Intelligence
upvoted
a
paper
about 15 hours ago
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement
liked
a dataset
1 day ago
jojo0217/korean_rlhf_dataset
View all activity
Organizations
None yet
kiyer97
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
2 datasets
1 day ago
jojo0217/korean_rlhf_dataset
Viewer
•
Updated
Sep 25, 2023
•
107k
•
93
•
33
yitingxie/rlhf-reward-datasets
Viewer
•
Updated
Jan 1, 2023
•
81.4k
•
174
•
70
liked
2 models
1 day ago
RLHFlow/ArmoRM-Llama3-8B-v0.1
Text Classification
•
8B
•
Updated
Sep 23, 2024
•
15.6k
•
192
kaiokendev/superhot-13b-8k-no-rlhf-test
Updated
Jul 1, 2023
•
66
liked
3 datasets
2 days ago
tasksource/oasst1_pairwise_rlhf_reward
Viewer
•
Updated
Jul 4, 2023
•
18.9k
•
233
•
50
liyucheng/zhihu_rlhf_3k
Viewer
•
Updated
Apr 15, 2023
•
3.46k
•
167
•
98
Anthropic/hh-rlhf
Viewer
•
Updated
May 26, 2023
•
169k
•
39.5k
•
2.08k
liked
2 models
2 days ago
RLHFlow/pair-preference-model-LLaMA3-8B
Text Generation
•
8B
•
Updated
Oct 14, 2024
•
245
•
•
41
nvidia/nemotron-3-8b-chat-4k-rlhf
Text Generation
•
Updated
Feb 9, 2024
•
35