Inkling-0.92B-MTP

Small BF16 test model using thinkingmachines/Inkling, with a toy-trained backbone and synthetic MTP weights.

Load the model

This fixture has no published quantized variant.

import torch
from transformers import InklingForConditionalGeneration

MODEL_ID = "inference-optimization/Inkling-0.92B-MTP"
model = InklingForConditionalGeneration.from_pretrained(
    MODEL_ID, dtype=torch.bfloat16, attn_implementation="eager",
)
Downloads last month
7
Safetensors
Model size
0.9B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Collections including inference-optimization/Inkling-0.92B-MTP