benchmarks moonshotai/PerceptionBench Viewer • Updated Aug 1 • 3k • 3.89k • 50 SWE-bench/SWE-bench_Verified Benchmark • Updated 21 days ago • 500 • 124k • 152 datacurve/deep-swe Benchmark • Updated Jun 2 • 113 • 1.02k • 65 cais/hle Benchmark • Updated Jan 20 • 2.5k • 40.1k • 948
post-training Qyrou/reasoning-corpus-4K-5M-v1 Preview • Updated 15 days ago • 12k • 204 yannelli/laravel-11-qa Viewer • Updated Oct 4, 2024 • 12.6k • 134 • 6 Manusagents/GPT-5.5-Gemini-3.1-Pro-Grok-4-Claude-Fable-5-Mythos-5-Qwen-3.7-Max-and-more-Distillation-Dataset Viewer • Updated 8 days ago • 18.5M • 18.4k • 193 HuggingFaceH4/ultrachat_200k Viewer • Updated Oct 16, 2024 • 515k • 112k • 885
Manusagents/GPT-5.5-Gemini-3.1-Pro-Grok-4-Claude-Fable-5-Mythos-5-Qwen-3.7-Max-and-more-Distillation-Dataset Viewer • Updated 8 days ago • 18.5M • 18.4k • 193
benchmarks moonshotai/PerceptionBench Viewer • Updated Aug 1 • 3k • 3.89k • 50 SWE-bench/SWE-bench_Verified Benchmark • Updated 21 days ago • 500 • 124k • 152 datacurve/deep-swe Benchmark • Updated Jun 2 • 113 • 1.02k • 65 cais/hle Benchmark • Updated Jan 20 • 2.5k • 40.1k • 948
post-training Qyrou/reasoning-corpus-4K-5M-v1 Preview • Updated 15 days ago • 12k • 204 yannelli/laravel-11-qa Viewer • Updated Oct 4, 2024 • 12.6k • 134 • 6 Manusagents/GPT-5.5-Gemini-3.1-Pro-Grok-4-Claude-Fable-5-Mythos-5-Qwen-3.7-Max-and-more-Distillation-Dataset Viewer • Updated 8 days ago • 18.5M • 18.4k • 193 HuggingFaceH4/ultrachat_200k Viewer • Updated Oct 16, 2024 • 515k • 112k • 885
Manusagents/GPT-5.5-Gemini-3.1-Pro-Grok-4-Claude-Fable-5-Mythos-5-Qwen-3.7-Max-and-more-Distillation-Dataset Viewer • Updated 8 days ago • 18.5M • 18.4k • 193