Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Shufan Shen's picture

Shufan Shen

shufanshen
2 3 4
webbrain-a10's profile picture
·
https://dblp.org/pid/277/0707.html
  • ssfgunner

AI & ML interests

Interpretable machine learning, parameter-efficient fine-tuning.

Recent Activity

upvoted a paper about 16 hours ago
On-Policy Parameter Update Direction Underlies Generalization in LLM Post-Training
authored a paper 2 days ago
On-Policy Parameter Update Direction Underlies Generalization in LLM Post-Training
authored a paper 2 days ago
Edit Less, Achieve More: Dynamic Sparse Neuron Masking for Lifelong Knowledge Editing in LLMs
View all activity

Organizations

ucas's profile picture

shufanshen 's collections 1

OPSFT
  • shufanshen/Qwen3-4B-GRPO-DeepMath-50-steps

    Reinforcement Learning • 4B • Updated 19 days ago • 15
  • shufanshen/Qwen3-4B-GRPO-DeepMath-100-steps

    Reinforcement Learning • 4B • Updated 19 days ago • 13
  • shufanshen/Qwen3-4B-GRPO-DeepMath-150-steps

    Reinforcement Learning • 4B • Updated 19 days ago • 19
  • shufanshen/Qwen3-8B-GRPO-DeepMath-50-steps

    Reinforcement Learning • 8B • Updated 19 days ago • 17
OPSFT
  • shufanshen/Qwen3-4B-GRPO-DeepMath-50-steps

    Reinforcement Learning • 4B • Updated 19 days ago • 15
  • shufanshen/Qwen3-4B-GRPO-DeepMath-100-steps

    Reinforcement Learning • 4B • Updated 19 days ago • 13
  • shufanshen/Qwen3-4B-GRPO-DeepMath-150-steps

    Reinforcement Learning • 4B • Updated 19 days ago • 19
  • shufanshen/Qwen3-8B-GRPO-DeepMath-50-steps

    Reinforcement Learning • 8B • Updated 19 days ago • 17
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs