LFM 2.5 1.2B Thinking API Benchmarks, Pricing & Provider Data
Compare LFM 2.5 1.2B Thinking with another model
Choose a model to open its comparison page.
Цены API LFM 2.5 1.2B Thinking у поставщиков (2): от $0.010/request до $0.010/request.
LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Возможности
Technical Details
- Вход
- Выход
- Всего параметров
- 1.2B
- Выпущена
- Jan 2026
- Токенизатор
- Other
- Архитектура
- text->text
- Модерация
- Нет
- Поддерживаемые параметры
- frequency_penaltyinclude_reasoningmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Сравнение цен
Compare LFM 2.5 1.2B Thinking API pricing across 2 providers. Prices range from $0.010/request to $0.010/request. 素墨API offers the lowest rate at $0.010/request.
Альтернативы и похожие модели
GPT-OSS
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
GLM-4.5 Air
glm-4-5-air
Zhipu AI GLM-4.5 Air is a lightweight GLM-4.5 variant optimized for low-latency Chinese and English dialogue, retrieval-augmented apps, and edge deployments.
LLFM 2.5 1.2B Instruct
lfm-2-5-1-2b-instruct
LFM 2.5 1.2B Instruct is a compact language model in the LFM series, optimized for low-latency responses and efficient inference.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Qwen3
qwen3
Alibaba Qwen3 is the Qwen family's flagship LLM series with dense and MoE variants, seamless thinking/non-thinking modes, and leading open-source performance in math, code, and agent tasks.
Qwen3.5
qwen3-5
Alibaba Qwen3.5 is a Qwen3 generation model with improved reasoning, multilingual support, and efficient inference for chat, coding, and agent applications.
Frequently Asked Questions
- What benchmark data does LFM 2.5 1.2B Thinking include?
- LMSpeed shows LFM 2.5 1.2B Thinking benchmark context, API price, output speed, first-token latency, and provider data across 2 providers when those signals are available.
- Сколько стоит API LFM 2.5 1.2B Thinking?
- Цены LFM 2.5 1.2B Thinking у поставщиков (2): от $0.010/request до $0.010/request. Самая низкая указанная цена — у 素墨API.
- What does the LFM 2.5 1.2B Thinking API pricing table include?
- The LFM 2.5 1.2B Thinking API pricing table compares 2 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- У какого провайдера самая низкая цена API LFM 2.5 1.2B Thinking?
- Сейчас самая низкая указанная цена LFM 2.5 1.2B Thinking — $0.010/request у 素墨API. Сравниваются поставщики: 2.
- Является ли API LFM 2.5 1.2B Thinking бесплатным?
- У LFM 2.5 1.2B Thinking в настоящее время нет бесплатного тарифа API на LMSpeed. Все 2 провайдеров взимают плату за token.
