Бесплатный Llama API

Сравните бесплатных Llama API провайдеров по скорости, задержке и доступности. · Бесплатных моделей 123 · Провайдеров 153

61–120 из 123
МодельПровайдерыСкоростьЗадержкаТестов
llama-3.2-11b-vision-instructMetaБесплатно
Н/ДН/Д0
llama-3.2-1b-instructMetaБесплатно
Н/ДН/Д0
llama-3.2-3b-instructMetaБесплатно
Н/ДН/Д0
llama-3.2-nemoretriever-1b-vlm-embed-v1MetaБесплатно
Н/ДН/Д0
llama-3.2-nv-embedqa-1b-v1MetaБесплатно
Н/ДН/Д0
llama-3.2-nv-embedqa-1b-v2MetaБесплатно
Н/ДН/Д0
llama-3.3-70b-instructMetaБесплатно
Н/ДН/Д0
llama-3.3-nemotron-super-49b-v1MetaБесплатно
Н/ДН/Д0
llama-3.3-nemotron-super-49b-v1.5MetaБесплатно
Н/ДН/Д0
llama-4-maverick-17b-128e-instructMetaБесплатно
Н/ДН/Д0
llama-4-maverick-17b-128e-instruct-fp8MetaБесплатно
Н/ДН/Д0
llama-4-scout-17b-16e-instructMetaБесплатно
Н/ДН/Д0
llama-nemotron-embed-1b-v2MetaБесплатно
Н/ДН/Д0
llama2-70bMetaБесплатно
Н/ДН/Д0
llama3-chatqa-1.5-70bMetaБесплатно
Н/ДН/Д0
meta-llama-3.1-405b-instructMetaБесплатно
Н/ДН/Д0
meta-llama-3.1-8b-instructMetaБесплатно
Н/ДН/Д0
NousResearch/Hermes-3-Llama-405BMetaБесплатно
Н/ДН/Д0
nvidia/Llama-3_1-Nemotron-Ultra-253B-v1Nebius Token FactoryБесплатно
ToolsOpen Weights128K
Н/ДН/Д0
meta-llama/Meta-Llama-3.1-405B-Instruct-Lite-ProMetaБесплатно
Н/ДН/Д0
meta-llama/Llama-Guard-3-8BNebius Token FactoryБесплатно
Open Weights8.2K
Н/ДН/Д0
meta-llama/Llama-3.3-70B-Instruct-fastNebius Token FactoryБесплатно
ToolsOpen Weights128K
Н/ДН/Д0
@hf/meta-llama/meta-llama-3-8b-instructMetaБесплатно
Н/ДН/Д0
@cf/meta-llama/llama-2-7b-chat-hf-loraMetaБесплатно
Н/ДН/Д0
fth/DavidAU/Llama3.3-8B-Instruct-Thinking-Claude-4.5-Opus-High-ReasoningMetaБесплатно
Н/ДН/Д0
deepseek-ai/deepseek-r1-distill-llama-8bDeepSeekБесплатно
Н/ДН/Д0
fth/DavidAU/Llama3.3-8B-Instruct-Thinking-Heretic-Uncensored-Claude-4.5-Opus-High-ReasoningAnthropicБесплатно
Н/ДН/Д0
fth/Undi95/Meta-Llama-3.1-8B-ClaudeMetaБесплатно
Н/ДН/Д0
fth/Undi95/Meta-Llama-3.1-8B-Claude-bf16MetaБесплатно
Н/ДН/Д0
llama-freeMetaБесплатно
Н/ДН/Д0
ollama/gemma4:31bGoogleБесплатно
Н/ДН/Д0
llama-3.1-8b-instruct-turbo-freeMetaБесплатно
Н/ДН/Д0
meta-llama/llama-3-70b-instructMetaБесплатно
Н/ДН/Д0
meta-llama/llama-guard-3-8bMetaБесплатно
Н/ДН/Д0
meta.llama3-1-70b-instructMetaБесплатно
Н/ДН/Д0
meta.llama3-1-8b-instructMetaБесплатно
Н/ДН/Д0
meta.llama3-3-70b-instructMetaБесплатно
Н/ДН/Д0
nousresearch/hermes-2-pro-llama-3-8bMetaБесплатно
Н/ДН/Д0
alfredpros/codellama-7b-instruct-solidityMetaБесплатно
Н/ДН/Д0
Llama 3.1MetaБесплатно
文本Meta Llama 3.1 extends the Llama 3 family with stronger reasoning, tool use, and long-context support across 8B to 405B scales.
Н/ДН/Д0
Llama 3.3MetaБесплатно
ToolsOpen Weights128KMeta Llama 3.3 is an updated Llama 3 open model with improved instruction following, multilingual support, and efficient inference.
Н/ДН/Д0
Dracarys Llama 3.1 InstructMetaБесплатно
Dracarys Llama 3.1 Instruct is an instruction-tuned variant, optimized for following instructions and conversational tasks.
Н/ДН/Д0
Usdcode Llama 3.1 InstructMetaБесплатно
Usdcode Llama 3.1 Instruct is an instruction-tuned variant, optimized for following instructions and conversational tasks.
Н/ДН/Д0
Meta Llama 3.1 InstructMetaБесплатно
Tools33KOpen Weights128KMeta Llama 3.1 Instruct is an instruction-tuned variant in the Llama series, optimized for following instructions and conversational tasks.
Н/ДН/Д0
Llama 3.2 Nemoretriever 300m Embed v1MetaБесплатно
Meta Llama 3.2 Nemoretriever 300m Embed v1 is an embedding model, designed for generating vector representations of text for retrieval and semantic search.
Н/ДН/Д0
MiniMax M2.7MiniMaxБесплатно
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
Н/ДН/Д0
Gemini 3 FlashGoogleБесплатно
Google Gemini 3 Flash is a next-generation fast multimodal model for responsive assistants, document understanding, and high-throughput API traffic.
Н/ДН/Д0
Llama 3.3 Nemotron Super 49B V1.5NVIDIAБесплатно
Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context. It’s post-trained for agentic workflows (RAG, tool calling) via SFT across math, code, science, and...
Н/ДН/Д0
Llama 4 MaverickMetaБесплатно
ToolsVision131.1KLlama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Н/ДН/Д0
Llama 4 ScoutMetaБесплатно
ReasoningToolsOpen WeightsVisionMeta Llama 4 Scout is a compact open-weight model in the Llama 4 family, designed for efficient inference, on-device deployment, and low-latency agent workloads.
Koyeb AI GatewaySWT-APICHB API
+1 ещёAI API
Н/ДН/Д0
Aion RP Llama 3.1MetaБесплатно
Aion RP Llama 3.1 is a roleplay-tuned variant in the Aion series, optimized for character-driven dialogue and creative writing.
Н/ДН/Д0
R1 Distill Llama 70BDeepSeekБесплатно
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...
Н/ДН/Д0
Llama 3.1 70B Hanami x1Sao10KБесплатно
This is [Sao10K](/sao10k)'s experiment over [Euryale v2.2](/sao10k/l3.1-euryale-70b).
Н/ДН/Д0
Llama 3.3 Euryale 70BSao10KБесплатно
Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.2](/models/sao10k/l3-euryale-70b).
Н/ДН/Д0
Llama 3.3 70B InstructMetaБесплатно
ReasoningToolsOpen Weights128KThe Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
Н/ДН/Д0
Llama 3.2 1B InstructMetaБесплатно
文本Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...
Н/ДН/Д0
Llama 3.2 3B InstructMetaБесплатно
文本Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
Н/ДН/Д0
Llama 3.2 11B Vision InstructMetaБесплатно
文本Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...
Н/ДН/Д0
Llama 3.1 Euryale 70B v2.2Sao10KБесплатно
Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).
Н/ДН/Д0
Hermes 3 Llama 3.1MetaБесплатно
Nous Research Hermes 3 is a generalist instruct model fine-tuned on Meta Llama 3.1, with strong reasoning, roleplay, multi-turn chat, tool calling, and structured JSON output.
Н/ДН/Д0

LMSpeed отслеживает 123 бесплатных моделей Llama API у 153 провайдеров. Все данные скорости и задержки получены из реальных API тестов и регулярно обновляются.