lewtun
·
AI & ML interests
LLMs, LLMs, LLMs
Recent Activity
Organizations
lewtun/sft_openassistant-guanaco
Updated
Text Classification
• 0.5B • Updated • 13
lewtun/pythia-6.9b-deduped-tldr-online-dpo
7B • Updated • 5
lewtun/qwen2-1.5B-ultrafeedback-online-dpo
2B • Updated • 9
lewtun/qwen2-0.5B-ultrafeedback-online-dpo
0.6B • Updated • 5
lewtun/pythia-2.8b-deduped-tldr-online-dpo
3B • Updated • 2
lewtun/qwen2-7B-ultrafeedback-online-dpo-bs-1
Updated
lewtun/qwen2-7B-ultrafeedback-online-dpo-bs-2
Updated
lewtun/qwen2-7B-ultrafeedback-online-dpo
Updated
lewtun/pythia-1b-deduped-tldr-online-dpo
1B • Updated • 1
lewtun/pythia-1b-tldr-online-dpo
Updated
lewtun/qwen2-0.5B-lr-5e-7
Updated
lewtun/qwen2-7B-lr-3e-6-tok-1024
Updated
lewtun/qwen2-0.5B-lr-3e-6-tok-1024
Updated
lewtun/qwen2-1.5B-lr-3e-6-tok-1024
Updated
lewtun/qwen2-1.5B-lr-3e-6-tok-2048
Updated
lewtun/qwen2-0.5B-lr-3e-6-tok-2048
Updated
lewtun/qwen2-1.5B-lr-3e-6
2B • Updated • 3
lewtun/qwen2-0.5B-lr-3e-6
0.5B • Updated • 14
0.5B • Updated • 2
2B • Updated • 6
lewtun/EleutherAI_pythia-1b
1B • Updated • 3
lewtun/kto-aligned-model-lora
lewtun/gemma-7b-dpo-full-openhermes-mix1-beta-0.05
Text Generation
• 9B • Updated • 11