vtava's picture
Publish accepted Memory Fusion prefix [0]
3d11298 verified
|
Raw History Blame Contribute Delete
1.41 kB
---
library_name: transformers
pipeline_tag: text-generation
tags:
- tinycenn
- cenn
- language-modeling
- text-generation
- research
---
# SmolLM2-135M-MemoryFusion-Sequential-R64
Research artifact from **TinyCeNN-LM**. Architecture: `TinyCeNN-LM experiment`.
## Architecture
- Architecture/run type: `TinyCeNN-LM experiment`
- Base model: `not recorded`
- Dataset: `Not recorded`
- Source code: https://github.com/vtavakkoli/TinyCeNN-LM
## Latest saved results
No structured training report was found in this upload.
The Hugging Face repository keeps timestamped run artifacts under `runs/`. This preserves training reports, configs and run metadata independently of the temporary Colab filesystem.
## Saved experiment files
- `tokenizer_config.json`
## Reproducibility
Run the matching notebook from the TinyCeNN-LM repository. Colab notebooks use a Hugging Face write token from the `HF_TOKEN` Colab Secret; tokens should never be pasted into notebook source.
## Limitations
This is a research checkpoint. Metrics saved here are the metrics produced by the corresponding training notebook/script; unless explicitly marked as held-out evaluation, they should not be treated as publication-grade benchmark results. Generation quality can differ substantially from the base model.
## Citation
If you use this experimental checkpoint, cite the TinyCeNN-LM repository and the upstream base model.