LFM2.5 2.6B

LiquidAI/LFM2.5-2.6B-GGUF:Q8_0

llama.cpp1 build benchmarkedbest build liquidai · Q8_013.8 GB84 tok/s
best build

Scores at a glance

measured avg16
Tool Calling180/27
Reasoning & Math144/28
0255075100

0–100 per task · n/m = cases passed (scores credit partial passes) · measured average covers 2/9 suite tasks

best build
config
deployment

liquidai · Q8_0

quantization
Q8_0
harness
llama.cpp
context
128K
memory
13.8 GB
throughput
84 tok/s
cost / run
$1.1119est.
provider
llama.cpp
run
2026-08-05
run this buildllama.cpp
llama-server -hf LiquidAI/LFM2.5-2.6B-GGUF:Q8_0 -c 128000

Pulls the GGUF from Hugging Face and serves an OpenAI-compatible endpoint on :8080.

record
per-test results

Full run record

Per-test scores, full transcripts, and the reproduce recipe.

open the run record →