Image-to-Video
Safetensors
English
video-generation
text-to-video
memory
MosaiChunk / README.md
evanbuzzZ's picture
Add license and citation
016a461 verified
|
Raw History Blame Contribute Delete
2.33 kB
---
license: other
license_name: mosaichunk-backbone-licenses
license_link: https://github.com/mosaichunk/MosaiChunk#license
language:
- en
tags:
- arxiv:2610.02153
- video-generation
- text-to-video
- image-to-video
- memory
- safetensors
datasets:
- mosaichunk/RememBench
---
# MosaiChunk
Learned memory-router checkpoints for **MosaiChunk: Compositing Spatio-Temporal Memory for Autoregressive Video Generation**.
[Project page](https://mosaichunk.github.io/) 路 [Paper](https://arxiv.org/abs/2610.02153) 路 [Code](https://github.com/mosaichunk/MosaiChunk) 路 [RememBench](https://huggingface.co/datasets/mosaichunk/RememBench)
| Checkpoint | Backbone |
|---|---|
| [`t2v/model.safetensors`](t2v/model.safetensors) | RAVEN-adapted MiniMax-H3 (H3-AR) |
| [`i2v/model.safetensors`](i2v/model.safetensors) | LingBot-World-Infinity |
Each folder contains router weights and `config.json`. The weights are exported without changing tensor values; optimizer and training state are excluded.
These checkpoints require their corresponding frozen video backbone. T2V also requires the pretrained RAVEN streaming adapter. Backbone and adapter weights are not bundled here.
## Download
```python
from huggingface_hub import snapshot_download
snapshot_download("mosaichunk/MosaiChunk", local_dir="checkpoints/MosaiChunk")
```
See the [code repository](https://github.com/mosaichunk/MosaiChunk) for training and inference.
## License
Each router checkpoint is subject to the license of its backbone:
- `i2v/`: LingBot-World-v2, [CC BY-NC-SA 4.0](https://creativecommons.org/licenses/by-nc-sa/4.0/)
- `t2v/`: MiniMax-H3 with the RAVEN streaming adapter, [MiniMax-H3 Community License Agreement](https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/LICENSE)
The checkpoints do not include backbone or adapter weights and grant no rights to them.
## Citation
```bibtex
@misc{zhang2026mosaichunkcompositingspatiotemporalmemory,
title={MosaiChunk: Compositing Spatio-Temporal Memory for Autoregressive Video Generation},
author={Yiwen Zhang and Haocheng Xi and Michael Tian-Yue Liu and Alexei A. Efros and Hadar Averbuch-Elor and Qianqian Wang and Haiwen Feng},
year={2026},
eprint={2610.02153},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2610.02153},
}
```