Text Generation
PEFT
Safetensors
English
super-mario-64
sm64
speedrun
tas
reasoning
qwen3
lora
unsloth
conversational
Instructions to use hugo74130/sm64 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- PEFT
How to use hugo74130/sm64 with PEFT:
from peft import PeftModel from transformers import AutoModelForCausalLM base_model = AutoModelForCausalLM.from_pretrained("unsloth/qwen3-4b-unsloth-bnb-4bit") model = PeftModel.from_pretrained(base_model, "hugo74130/sm64") - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- Unsloth Desktop
File size: 1,910 Bytes
c94b06a | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 46 47 48 49 50 51 52 53 54 55 56 57 58 59 60 61 62 | ---
license: cc-by-nc-4.0
base_model: unsloth/Qwen3-4B-unsloth-bnb-4bit
library_name: peft
tags:
- super-mario-64
- sm64
- speedrun
- tas
- reasoning
- qwen3
- lora
- unsloth
language:
- en
pipeline_tag: text-generation
---
# SM64 Speedrun / TAS Assistant — Qwen3-4B LoRA
A **QLoRA adapter** for **Qwen3-4B** fine-tuned to answer **Super Mario 64**
speedrunning and TAS questions with step-by-step reasoning inside
`<think>...</think>`, followed by a clear answer.
- **Base model:** unsloth/Qwen3-4B-unsloth-bnb-4bit (Qwen/Qwen3-4B)
- **Method:** QLoRA (4-bit), rank 16, alpha 32, 2 epochs
- **This repo:** LoRA adapter only (~127 MB) — load on top of the base model.
## Usage
```python
from peft import PeftModel
from transformers import AutoModelForCausalLM, AutoTokenizer
base = "Qwen/Qwen3-4B"
model = AutoModelForCausalLM.from_pretrained(base, device_map="auto")
model = PeftModel.from_pretrained(model, "hugo74130/sm64")
tok = AutoTokenizer.from_pretrained("hugo74130/sm64")
```
Ask **English** questions about SM64 speedrun/TAS mechanics (BLJ, PU, GLG,
RNG, categories, levels, tricks). Recommended sampling (Qwen3 thinking):
**temperature 0.6, top_p 0.95, top_k 20, min_p 0, repetition_penalty 1.0**.
Knowledge is limited to the training topics (grounded on Ukikipedia); it may
hallucinate on out-of-scope questions.
## Source & License
Training data was **derived from Ukikipedia** (https://ukikipedia.net),
licensed **CC BY-NC 4.0** (https://creativecommons.org/licenses/by-nc/4.0/).
Q&A pairs were generated/transformed from wiki content (changes made). As a
derivative work, this adapter is released under the **same license —
CC BY-NC 4.0 — for non-commercial use only.**
**Attribution:** Ukikipedia contributors — https://ukikipedia.net
Not affiliated with, or endorsed by, Ukikipedia or Nintendo.
## Disclaimer
Wiki-sourced content — factual accuracy is not guaranteed.
|