Liquid AI
Try LFM โ€ข Documentation โ€ข LEAP

LFM2.5-1.2B-Instruct-4bit

MLX export of LFM2.5-1.2B-Instruct for Apple Silicon inference.

Model Details

Property Value
Parameters 1.2B
Precision 4-bit
Group Size 64
Size 628 MB
Context Length 128K

Recommended Sampling Parameters

Parameter Value
temperature 0.1
top_k 50
top_p 0.1
repetition_penalty 1.05
max_tokens 512

Use with mlx

pip install mlx-lm
from mlx_lm import load, generate
from mlx_lm.sample_utils import make_sampler, make_logits_processors

model, tokenizer = load("LiquidAI/LFM2.5-1.2B-Instruct-4bit")

prompt = "What is the capital of France?"

if tokenizer.chat_template is not None:
    messages = [{"role": "user", "content": prompt}]
    prompt = tokenizer.apply_chat_template(
        messages, tokenize=False, add_generation_prompt=True
    )

sampler = make_sampler(temp=0.1, top_k=50, top_p=0.1)
logits_processors = make_logits_processors(repetition_penalty=1.05)

response = generate(
    model,
    tokenizer,
    prompt=prompt,
    max_tokens=512,
    sampler=sampler,
    logits_processors=logits_processors,
    verbose=True,
)

License

This model is released under the LFM 1.0 License.

Downloads last month
2,333
Safetensors
Model size
1B params
Tensor type
U32
ยท
BF16
ยท
MLX
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for LiquidAI/LFM2.5-1.2B-Instruct-MLX-4bit

Quantized
(114)
this model
Quantizations
1 model

Space using LiquidAI/LFM2.5-1.2B-Instruct-MLX-4bit 1