Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
HappyBreadDuck's picture

HappyBreadDuck

HappyBreadDuck
1 3
ยท

AI & ML interests

None yet

Recent Activity

liked a model about 20 hours ago
Kwaipilot/KAT-Coder-V2.5-Dev
reacted to sergiopaniego's post with ๐Ÿ”ฅ about 21 hours ago
Can you do RL over taste? I've spent some time reproducing, in the open, Surya N's idea of training a model to paint with code. It's a coding model that learns to paint watercolours by writing JS code, trained with GRPO. I used TRL and OpenEnv for this, with the whole pipeline running on Hugging Face. The interesting part is that the reward has no correct answer, unlike a math problem. In this case it's based on the artistic preferences of the person who builds the dataset. Everything is published: the environment, the reference pool, the trained adapters, every painting of every run with the code that made it, and a write-up with all the decisions, including the ones that went wrong. Blog post: https://huggingface.co/blog/train-to-paint-with-code
liked a model 3 days ago
gbuzhf/KAT-Coder-V2.5-Dev-MTP-GGUF
View all activity

Organizations

None yet

HappyBreadDuck 's models

None public yet
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs