Бесплатный Llama API
Сравните бесплатных Llama API провайдеров по скорости, задержке и доступности. · Бесплатных моделей 123 · Провайдеров 153
61–120 из 123
| Модель | Провайдеры | Скорость | Задержка | Тестов |
|---|---|---|---|---|
llama-3.2-11b-vision-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
llama-3.2-1b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
llama-3.2-3b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
llama-3.2-nemoretriever-1b-vlm-embed-v1MetaБесплатно | Н/Д | Н/Д | 0 | |
llama-3.2-nv-embedqa-1b-v1MetaБесплатно | Н/Д | Н/Д | 0 | |
llama-3.2-nv-embedqa-1b-v2MetaБесплатно | Н/Д | Н/Д | 0 | |
llama-3.3-70b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
llama-3.3-nemotron-super-49b-v1MetaБесплатно | Н/Д | Н/Д | 0 | |
llama-3.3-nemotron-super-49b-v1.5MetaБесплатно | Н/Д | Н/Д | 0 | |
llama-4-maverick-17b-128e-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
llama-4-maverick-17b-128e-instruct-fp8MetaБесплатно | Н/Д | Н/Д | 0 | |
llama-4-scout-17b-16e-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
llama-nemotron-embed-1b-v2MetaБесплатно | Н/Д | Н/Д | 0 | |
llama2-70bMetaБесплатно | Н/Д | Н/Д | 0 | |
llama3-chatqa-1.5-70bMetaБесплатно | Н/Д | Н/Д | 0 | |
meta-llama-3.1-405b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
meta-llama-3.1-8b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
NousResearch/Hermes-3-Llama-405BMetaБесплатно | Н/Д | Н/Д | 0 | |
nvidia/Llama-3_1-Nemotron-Ultra-253B-v1Nebius Token FactoryБесплатно ToolsOpen Weights128K | Н/Д | Н/Д | 0 | |
meta-llama/Meta-Llama-3.1-405B-Instruct-Lite-ProMetaБесплатно | Н/Д | Н/Д | 0 | |
meta-llama/Llama-Guard-3-8BNebius Token FactoryБесплатно Open Weights8.2K | Н/Д | Н/Д | 0 | |
meta-llama/Llama-3.3-70B-Instruct-fastNebius Token FactoryБесплатно ToolsOpen Weights128K | Н/Д | Н/Д | 0 | |
@hf/meta-llama/meta-llama-3-8b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
@cf/meta-llama/llama-2-7b-chat-hf-loraMetaБесплатно | Н/Д | Н/Д | 0 | |
fth/DavidAU/Llama3.3-8B-Instruct-Thinking-Claude-4.5-Opus-High-ReasoningMetaБесплатно | Н/Д | Н/Д | 0 | |
deepseek-ai/deepseek-r1-distill-llama-8bDeepSeekБесплатно | Н/Д | Н/Д | 0 | |
fth/DavidAU/Llama3.3-8B-Instruct-Thinking-Heretic-Uncensored-Claude-4.5-Opus-High-ReasoningAnthropicБесплатно | Н/Д | Н/Д | 0 | |
fth/Undi95/Meta-Llama-3.1-8B-ClaudeMetaБесплатно | Н/Д | Н/Д | 0 | |
fth/Undi95/Meta-Llama-3.1-8B-Claude-bf16MetaБесплатно | Н/Д | Н/Д | 0 | |
llama-freeMetaБесплатно | Н/Д | Н/Д | 0 | |
ollama/gemma4:31bGoogleБесплатно | Н/Д | Н/Д | 0 | |
llama-3.1-8b-instruct-turbo-freeMetaБесплатно | Н/Д | Н/Д | 0 | |
meta-llama/llama-3-70b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
meta-llama/llama-guard-3-8bMetaБесплатно | Н/Д | Н/Д | 0 | |
meta.llama3-1-70b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
meta.llama3-1-8b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
meta.llama3-3-70b-instructMetaБесплатно | Н/Д | Н/Д | 0 | |
nousresearch/hermes-2-pro-llama-3-8bMetaБесплатно | Н/Д | Н/Д | 0 | |
alfredpros/codellama-7b-instruct-solidityMetaБесплатно | Н/Д | Н/Д | 0 | |
文本Meta Llama 3.1 extends the Llama 3 family with stronger reasoning, tool use, and long-context support across 8B to 405B scales. | Н/Д | Н/Д | 0 | |
ToolsOpen Weights128KMeta Llama 3.3 is an updated Llama 3 open model with improved instruction following, multilingual support, and efficient inference. | Н/Д | Н/Д | 0 | |
Dracarys Llama 3.1 Instruct is an instruction-tuned variant, optimized for following instructions and conversational tasks. | Н/Д | Н/Д | 0 | |
Usdcode Llama 3.1 Instruct is an instruction-tuned variant, optimized for following instructions and conversational tasks. | Н/Д | Н/Д | 0 | |
Tools33KOpen Weights128KMeta Llama 3.1 Instruct is an instruction-tuned variant in the Llama series, optimized for following instructions and conversational tasks. | Н/Д | Н/Д | 0 | |
Meta Llama 3.2 Nemoretriever 300m Embed v1 is an embedding model, designed for generating vector representations of text for retrieval and semantic search. | Н/Д | Н/Д | 0 | |
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads. | Н/Д | Н/Д | 0 | |
Google Gemini 3 Flash is a next-generation fast multimodal model for responsive assistants, document understanding, and high-throughput API traffic. | Н/Д | Н/Д | 0 | |
Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context. It’s post-trained for agentic workflows (RAG, tool calling) via SFT across math, code, science, and... | Н/Д | Н/Д | 0 | |
ToolsVision131.1KLlama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward... | Н/Д | Н/Д | 0 | |
ReasoningToolsOpen WeightsVisionMeta Llama 4 Scout is a compact open-weight model in the Llama 4 family, designed for efficient inference, on-device deployment, and low-latency agent workloads. | Н/Д | Н/Д | 0 | |
Aion RP Llama 3.1 is a roleplay-tuned variant in the Aion series, optimized for character-driven dialogue and creative writing. | Н/Д | Н/Д | 0 | |
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across... | Н/Д | Н/Д | 0 | |
This is [Sao10K](/sao10k)'s experiment over [Euryale v2.2](/sao10k/l3.1-euryale-70b). | Н/Д | Н/Д | 0 | |
Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.2](/models/sao10k/l3-euryale-70b). | Н/Д | Н/Д | 0 | |
ReasoningToolsOpen Weights128KThe Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model... | Н/Д | Н/Д | 0 | |
文本Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate... | Н/Д | Н/Д | 0 | |
文本Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it... | Н/Д | Н/Д | 0 | |
文本Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and... | Н/Д | Н/Д | 0 | |
Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b). | Н/Д | Н/Д | 0 | |
Nous Research Hermes 3 is a generalist instruct model fine-tuned on Meta Llama 3.1, with strong reasoning, roleplay, multi-turn chat, tool calling, and structured JSON output. | Н/Д | Н/Д | 0 |
LMSpeed отслеживает 123 бесплатных моделей Llama API у 153 провайдеров. Все данные скорости и задержки получены из реальных API тестов и регулярно обновляются.
