You need to agree to share your contact information to access this model

This repository is publicly accessible, but you have to accept the conditions to access its files and content.

Log in or Sign Up to review the conditions and access this model content.

HIPPO πŸ¦›

This repository contains our 4B model for HIPPO: Harmful Input, Positive and Productive Output. This is part of our work, From Specialization to Generalization: Instruction-tuned LLMs for Robust Harmful Content Mitigation, where we show that instruction-tuning an LLM for hate speech can dramatically improve its performance on hate speech-related tasks. We've collected 36 different hate speech datasets and converted them into a conversational template. We used it to train our HIPPO 4B and 32B models, which outperform several task-specific models.

Downloads last month
-
Safetensors
Model size
4B params
Tensor type
BF16
Β·
U8
Β·
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for leukas/hippo-4b

Quantized
(321)
this model

Dataset used to train leukas/hippo-4b

Collection including leukas/hippo-4b

Paper for leukas/hippo-4b