Text Generation
Safetensors
English
qwen3
conversational

TaH2-1.7B-max2

Qwen3-1.7B with adaptive latent iteration, at most two iterations per token. Uses duo attention, a learned iteration decider (threshold 0.5), and stop_prob_mix. Fine-tuned on the 1.7B subset of TaH2 AMTeam Tool.

Code: thu-nics/TaH · Paper: arXiv:2609.35748.

Install the TaH code, then load the backbone and both TaH modules:

import torch
from huggingface_hub import snapshot_download
from transformers import AutoTokenizer
from tah.model.tah_model import TaHForCausalLM

path = snapshot_download("nics-efc/TaH2-1.7B-max2")
tokenizer = AutoTokenizer.from_pretrained(path)
model = TaHForCausalLM.from_pretrained(
    path, dtype=torch.bfloat16, device_map="cuda").eval()
Downloads last month
-
Safetensors
Model size
2B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for nics-efc/TaH2-1.7B-max2

Finetuned
(447)
this model

Dataset used to train nics-efc/TaH2-1.7B-max2

Collection including nics-efc/TaH2-1.7B-max2

Paper for nics-efc/TaH2-1.7B-max2