Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
princeton-nlp
's Collections
RLMT Experiments
SimPO
SWE-bench
ProLong
Sheared Llama
SimCSE
SimPO
updated
Mar 16, 2025
This collections contains a list of SimPO and baseline models.
Upvote
24
+14
Sort: Collection
princeton-nlp/gemma-2-9b-it-SimPO
Text Generation
•
9B
•
Updated
Aug 2, 2024
•
414
•
•
173
princeton-nlp/gemma-2-9b-it-DPO
Text Generation
•
9B
•
Updated
Jul 18, 2024
•
11
•
•
9
princeton-nlp/Llama-3-Base-8B-SFT-IPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
67
•
•
1
princeton-nlp/Llama-3-Base-8B-SFT-DPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
96
•
princeton-nlp/Llama-3-Base-8B-SFT-KTO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
19
•
princeton-nlp/Llama-3-Base-8B-SFT-ORPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
36
•
princeton-nlp/Llama-3-Base-8B-SFT-RDPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
61
•
princeton-nlp/Llama-3-Base-8B-SFT-SimPO
Text Generation
•
8B
•
Updated
May 24, 2024
•
44
•
•
1
princeton-nlp/Llama-3-Base-8B-SFT
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
628
•
•
4
princeton-nlp/Llama-3-Instruct-8B-SimPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
20
•
•
59
princeton-nlp/Llama-3-Instruct-8B-IPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
15
•
princeton-nlp/Llama-3-Instruct-8B-KTO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
17
•
princeton-nlp/Llama-3-Instruct-8B-ORPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
20
•
princeton-nlp/Llama-3-Instruct-8B-RDPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
14
•
princeton-nlp/Llama-3-Instruct-8B-DPO
Text Generation
•
8B
•
Updated
Jun 17, 2024
•
12
•
princeton-nlp/Mistral-7B-Instruct-RDPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
17
princeton-nlp/Mistral-7B-Instruct-DPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
10
princeton-nlp/Mistral-7B-Instruct-IPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
13
princeton-nlp/Mistral-7B-Instruct-KTO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
14
princeton-nlp/Mistral-7B-Instruct-SimPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
12
•
2
princeton-nlp/Mistral-7B-Instruct-ORPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
14
princeton-nlp/Mistral-7B-Base-SFT-IPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
36
princeton-nlp/Mistral-7B-Base-SFT-KTO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
13
princeton-nlp/Mistral-7B-Base-SFT-DPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
18
princeton-nlp/Mistral-7B-Base-SFT-RDPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
24
princeton-nlp/Mistral-7B-Base-SFT-SimPO
Text Generation
•
7B
•
Updated
Jun 17, 2024
•
53
princeton-nlp/llama3-ultrafeedback
Viewer
•
Updated
Jul 18, 2024
•
61.8k
•
108
•
18
princeton-nlp/Mistral-7B-Base-SFT-CPO
Text Generation
•
7B
•
Updated
Sep 30, 2024
•
15
•
1
princeton-nlp/Mistral-7B-Base-SFT-RRHF
Text Generation
•
7B
•
Updated
Sep 30, 2024
•
10
princeton-nlp/Mistral-7B-Base-SFT-SLiC-HF
Text Generation
•
7B
•
Updated
Jul 7, 2024
•
14
princeton-nlp/Mistral-7B-Instruct-CPO
Text Generation
•
7B
•
Updated
Jul 7, 2024
•
13
princeton-nlp/Mistral-7B-Instruct-RRHF
Text Generation
•
7B
•
Updated
Jul 7, 2024
•
17
princeton-nlp/Mistral-7B-Instruct-SLiC-HF
Text Generation
•
7B
•
Updated
Jul 7, 2024
•
11
princeton-nlp/Llama-3-Base-8B-SFT-CPO
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
11
•
princeton-nlp/Llama-3-Base-8B-SFT-RRHF
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
12
•
princeton-nlp/Llama-3-Base-8B-SFT-SLiC-HF
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
42
•
princeton-nlp/Llama-3-Instruct-8B-CPO
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
11
•
princeton-nlp/Llama-3-Instruct-8B-RRHF
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
18
•
princeton-nlp/Llama-3-Instruct-8B-SLiC-HF
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
14
•
princeton-nlp/Llama-3-Instruct-8B-RRHF-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
15
•
princeton-nlp/Llama-3-Instruct-8B-SLiC-HF-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
17
•
princeton-nlp/Llama-3-Instruct-8B-DPO-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
16
•
princeton-nlp/Llama-3-Instruct-8B-IPO-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
13
•
princeton-nlp/Llama-3-Instruct-8B-CPO-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
13
•
princeton-nlp/Llama-3-Instruct-8B-KTO-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
15
•
princeton-nlp/Llama-3-Instruct-8B-ORPO-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
19
•
princeton-nlp/Llama-3-Instruct-8B-RDPO-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
19
•
princeton-nlp/Llama-3-Instruct-8B-SimPO-v0.2
Text Generation
•
8B
•
Updated
Jul 7, 2024
•
65
•
•
8
princeton-nlp/llama3-ultrafeedback-armorm
Viewer
•
Updated
Jul 18, 2024
•
61.8k
•
654
•
20
Upvote
24
+20
Sort: Collection
Share collection
View history
Collection guide
Browse collections