peterkirby/modernbert-large-pan2020-authorship-verification
Feature Extraction • 0.4B • Updated • 10 • 1
Our results show that contrastive learning outperforms a classification-based approach to authorship verification under the tested settings. We identify loss function, batch size, training duration, pre-trained model, input context length, and random text span data augmentation as important factors of model performance. Based on these considerations, we develop a ModernBERT Bi-Encoder model that achieves 98.4% accuracy on the PAN21 authorship verification task.
Get this paper in your agent:
hf papers read 2609.28471 curl -LsSf https://hf.co/cli/install.sh | bash No dataset linking this paper
No Space linking this paper
No Collection including this paper