Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Karthik Iyer's picture

Karthik Iyer

kiyer97
11 9

AI & ML interests

Reinforcement learning, reward modeling, RLHF, policy optimization, offline RL

Recent Activity

upvoted a paper about 15 hours ago
SenseNova-U1.5: Towards Native Unified Visual Intelligence
upvoted a paper about 15 hours ago
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement
liked a dataset 1 day ago
jojo0217/korean_rlhf_dataset
View all activity

Organizations

None yet

liked 2 datasets 1 day ago

jojo0217/korean_rlhf_dataset

Viewer • Updated Sep 25, 2023 • 107k • 93 • 33

yitingxie/rlhf-reward-datasets

Viewer • Updated Jan 1, 2023 • 81.4k • 174 • 70
liked 2 models 1 day ago

RLHFlow/ArmoRM-Llama3-8B-v0.1

Text Classification • 8B • Updated Sep 23, 2024 • 15.6k • 192

kaiokendev/superhot-13b-8k-no-rlhf-test

Updated Jul 1, 2023 • 66
liked 3 datasets 2 days ago

tasksource/oasst1_pairwise_rlhf_reward

Viewer • Updated Jul 4, 2023 • 18.9k • 233 • 50

liyucheng/zhihu_rlhf_3k

Viewer • Updated Apr 15, 2023 • 3.46k • 167 • 98

Anthropic/hh-rlhf

Viewer • Updated May 26, 2023 • 169k • 39.5k • 2.08k
liked 2 models 2 days ago

RLHFlow/pair-preference-model-LLaMA3-8B

Text Generation • 8B • Updated Oct 14, 2024 • 245 • • 41

nvidia/nemotron-3-8b-chat-4k-rlhf

Text Generation • Updated Feb 9, 2024 • 35
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs