YAML Metadata Warning:empty or missing yaml metadata in repo card

Check out the documentation for more information.

FinGPT-131M

FinGPT-131M is a 131M-parameter decoder-only language model trained from scratch on a GST/tax regulatory corpus.

Architecture

  • Parameters: ~131.3M
  • Layers: 13
  • Hidden size: 768
  • Attention heads: 12
  • Context length: 1024
  • Vocabulary: 50,257
  • Tokenizer: GPT-2 BPE
  • Activation: GELU
  • Attention: causal scaled dot-product attention
  • Weight tying: Yes

Training

Best checkpoint:

  • Step: 11,250
  • Validation loss: 2.7641

The model was trained from scratch; it is not fine-tuned from GPT-2.

Intended use

The model is intended for experimentation and research on language modeling over GST and tax-related text.

Limitations

This model should not be treated as a source of authoritative legal or tax advice. Tax regulations can change, and model outputs may contain hallucinations or outdated information.

Downloads last month
47
Safetensors
Model size
0.1B params
Tensor type
F32
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support