LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K tokens) and is designed to provide higher-quality “thinking” responses in a small 1.2B model.
Modalities
Price
Free
Context
33K
Released
Jan 20, 2026
LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K tokens) and is designed to provide higher-quality “thinking” responses in a small 1.2B model.
Yes. The pricing shown on this page for LFM2.5-1.2B-Thinking is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
LFM2.5-1.2B-Thinking has a 32,768 token context window.
LFM2.5-2.6B (free) is another text model from Liquid.
LFM2.5-1.2B-Thinking was released on January 20, 2026.
Token volume and request traffic to this model over time.