Our models trained to play chess (4B)
Princeton NLP group
princeton-nlp
AI & ML interests
None yet
Recent Activity
upvoted a paper about 13 hours ago
Language Models that Play Chess and Explain Their Moves updated a model 1 day ago
princeton-nlp/queen_hce-4 updated a model 1 day ago
princeton-nlp/queen_pawn-8Organizations
SimPO
This collections contains a list of SimPO and baseline models.
-
princeton-nlp/gemma-2-9b-it-SimPO
Text Generation • 9B • Updated • 447 • • 173 -
princeton-nlp/gemma-2-9b-it-DPO
Text Generation • 9B • Updated • 51 • • 9 -
princeton-nlp/Llama-3-Base-8B-SFT-IPO
Text Generation • 8B • Updated • 57 • • 1 -
princeton-nlp/Llama-3-Base-8B-SFT-DPO
Text Generation • 8B • Updated • 82 •
ProLong
ProLong is a family of long-context models that are continued trained and supervised fine-tuned from Llama-3-8B, with a maximum context window of 512K
-
princeton-nlp/Llama-3-8B-ProLong-64k-Base
Text Generation • 8B • Updated • 8.83k • • 7 -
princeton-nlp/Llama-3-8B-ProLong-64k-Instruct
Text Generation • 8B • Updated • 8.62k • • 13 -
princeton-nlp/Llama-3-8B-ProLong-512k-Base
8B • Updated • 9.21k • 11 -
princeton-nlp/Llama-3-8B-ProLong-512k-Instruct
8B • Updated • 8.36k • 26
SimCSE
-
princeton-nlp/unsup-simcse-bert-base-uncased
Feature Extraction • Updated • 1.02k • 6 -
princeton-nlp/unsup-simcse-bert-large-uncased
Feature Extraction • Updated • 219 • 1 -
princeton-nlp/unsup-simcse-roberta-base
Feature Extraction • Updated • 6.49k • • 9 -
princeton-nlp/unsup-simcse-roberta-large
Feature Extraction • Updated • 802 • 3
RLMT Experiments
The *RLMT* collection. Coming soon!
SWE-bench
SWE-bench is a benchmark for evaluating Language Models and AI Systems on their ability resolve real world GitHub Issues.
Sheared Llama
-
princeton-nlp/Sheared-LLaMA-1.3B
Text Generation • Updated • 8.05k • 98 -
princeton-nlp/Sheared-LLaMA-2.7B
Text Generation • Updated • 1.85k • 61 -
princeton-nlp/Sheared-LLaMA-1.3B-ShareGPT
Text Generation • Updated • 236 • 10 -
princeton-nlp/Sheared-LLaMA-2.7B-ShareGPT
Text Generation • Updated • 215 • 8
QUEEN (Chess) models
Our models trained to play chess (4B)
RLMT Experiments
The *RLMT* collection. Coming soon!
SimPO
This collections contains a list of SimPO and baseline models.
-
princeton-nlp/gemma-2-9b-it-SimPO
Text Generation • 9B • Updated • 447 • • 173 -
princeton-nlp/gemma-2-9b-it-DPO
Text Generation • 9B • Updated • 51 • • 9 -
princeton-nlp/Llama-3-Base-8B-SFT-IPO
Text Generation • 8B • Updated • 57 • • 1 -
princeton-nlp/Llama-3-Base-8B-SFT-DPO
Text Generation • 8B • Updated • 82 •
SWE-bench
SWE-bench is a benchmark for evaluating Language Models and AI Systems on their ability resolve real world GitHub Issues.
ProLong
ProLong is a family of long-context models that are continued trained and supervised fine-tuned from Llama-3-8B, with a maximum context window of 512K
-
princeton-nlp/Llama-3-8B-ProLong-64k-Base
Text Generation • 8B • Updated • 8.83k • • 7 -
princeton-nlp/Llama-3-8B-ProLong-64k-Instruct
Text Generation • 8B • Updated • 8.62k • • 13 -
princeton-nlp/Llama-3-8B-ProLong-512k-Base
8B • Updated • 9.21k • 11 -
princeton-nlp/Llama-3-8B-ProLong-512k-Instruct
8B • Updated • 8.36k • 26
Sheared Llama
-
princeton-nlp/Sheared-LLaMA-1.3B
Text Generation • Updated • 8.05k • 98 -
princeton-nlp/Sheared-LLaMA-2.7B
Text Generation • Updated • 1.85k • 61 -
princeton-nlp/Sheared-LLaMA-1.3B-ShareGPT
Text Generation • Updated • 236 • 10 -
princeton-nlp/Sheared-LLaMA-2.7B-ShareGPT
Text Generation • Updated • 215 • 8
SimCSE
-
princeton-nlp/unsup-simcse-bert-base-uncased
Feature Extraction • Updated • 1.02k • 6 -
princeton-nlp/unsup-simcse-bert-large-uncased
Feature Extraction • Updated • 219 • 1 -
princeton-nlp/unsup-simcse-roberta-base
Feature Extraction • Updated • 6.49k • • 9 -
princeton-nlp/unsup-simcse-roberta-large
Feature Extraction • Updated • 802 • 3