发现各大提供商提供的免费大模型 API,对比速度与延迟数据。免费模型 2773服务商 157免费服务 8727
| 模型 | 服务商 | 速度 | 延迟 | 测试 |
|---|---|---|---|---|
ERNIE-3.5-128K百度免费 | N/A | N/A | 0 | |
gemini-3.1-flash-lite-image-ssvipGoogle免费 | N/A | N/A | 0 | |
gemini-3.1-flash-lite-imageGoogle免费 | N/A | N/A | 0 | |
LMSpeed 追踪了来自 157 家 API 提供商的 2773 个免费大语言模型。各提供商的免费额度不同——有的限制每日请求数,有的为新用户提供免费额度。所有速度数据均来自真实 API 测试。
先找到合适的免费模型,对比它背后的服务商,再进入模型或服务商详情页确认限制与表现。
使用搜索、模型系列筛选和能力标签,把目录缩小到真正想测试的免费模型。
查看服务商数量、测试量、吞吐速度和首字延迟,不只按价格选择。
在服务商详情页确认可用性、健康检测、价格备注和免费额度限制,再用于测试或集成。
关于免费额度与如何使用这份目录的常见问题。
| N/A |
| N/A |
| 0 |
Baichuan-M2免费 | N/A | N/A | 0 |
hunyuan-vision腾讯免费 | N/A | N/A | 0 |
gemini-3.1-flash-image-preview-ssvipGoogle免费 | N/A | N/A | 0 |
gpt-5.5-1mOpenAI免费 | N/A | N/A | 0 |
chirp-v5免费 | N/A | N/A | 0 |
gpt-image-1-dmx00OpenAI免费 | N/A | N/A | 0 |
gpt-image-2-03OpenAI免费 | N/A | N/A | 0 |
Doubao-pro-32k字节跳动免费 | N/A | N/A | 0 |
gemini-3-pro-image-ssvipGoogle免费 | N/A | N/A | 0 |
gemini-3-pro-image-preview-ssvipGoogle免费 | N/A | N/A | 0 |
gemini-3-pro-image-preview-dfsx-0.3Google免费 | N/A | N/A | 0 |
gemini-3-pro-image-preview-0.3Google免费 | N/A | N/A | 0 |
gpt-5.4-cdxOpenAI免费 | N/A | N/A | 0 |
Doubao-pro-128k字节跳动免费 | N/A | N/A | 0 |
minimax_files_retrieve免费 | N/A | N/A | 0 |
qwen-mt-turboAlibaba (China)免费 16.4K | N/A | N/A | 0 |
gemini-3-flash-preview-ssvipGoogle免费 | N/A | N/A | 0 |
Doubao-lite-4k字节跳动免费 | N/A | N/A | 0 |
gpt-5.4-nano-ssvipOpenAI免费 | N/A | N/A | 0 |
mj_turbo_shorten免费 | N/A | N/A | 0 |
gpt-5.4-mini-ssvipOpenAI免费 | N/A | N/A | 0 |
mj_turbo_variation免费 | N/A | N/A | 0 |
qwen3.7-plus-80off阿里巴巴免费 | N/A | N/A | 0 |
MBZUAI-IFM/K2-Think-v2免费 | N/A | N/A | 0 |
Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM. | N/A | N/A | 0 |
gpt-5.5-pro20xOpenAI免费 codexcodingopenai | N/A | N/A | 0 |
gpt-5.4-pro20xOpenAI免费 | N/A | N/A | 0 |
ReasoningToolsFilesVisionAnthropic 的 Claude Opus 4.8是旗舰大语言模型,专为高复杂度通用任务与多步推理优化。 | 43.48 t/s | 5.57 s | 5 |
codex-gpt-image-2OpenAI免费 | N/A | N/A | 0 |
gpt-5.5-openai-compactOpenAI免费 | N/A | N/A | 0 |
gpt-image-2-4kOpenAI免费 imagegpt4kultra | N/A | N/A | 0 |
LongCat-2.0-PreviewLongCat免费 | 22.83 t/s | 11.07 s | 5 |
deepseek-v4-flash-searchDeepSeek免费 | 38.08 t/s | 6.06 s | 5 |
qwen/qwen3-next-80b-a3b-thinking阿里巴巴免费 | N/A | N/A | 0 |
baichuan-inc/baichuan2-13b-chat免费 | N/A | N/A | 0 |
microsoft/phi-3.5-vision-instruct免费 | N/A | N/A | 0 |
microsoft/phi-4-mini-flash-reasoning免费 | N/A | N/A | 0 |
meta/llama3-70b-instructMeta免费 | N/A | N/A | 0 |
meta/llama-3.1-405b-instructMeta免费 | N/A | N/A | 0 |
nvidia/ai-synthetic-video-detector免费 | N/A | N/A | 0 |
grok-4.20-0309-superxAI免费 | N/A | N/A | 0 |
grok-4.20-0309-reasoning-superxAI免费 | N/A | N/A | 0 |
stockmark/stockmark-2-100b-instruct免费 | N/A | N/A | 0 |
LongCat-Flash-Thinking-2601LongCat免费 | N/A | N/A | 0 |
LongCat-Flash-LiteLongCat免费 翻译模型,并发无上限 | N/A | N/A | 0 |
LongCat-Flash-ChatLongCat免费 | N/A | N/A | 0 |
yentinglin/llama-3-taiwan-70b-instructMeta免费 | N/A | N/A | 0 |
deepseek-v4-vision-search-nothinkingDeepSeek免费 | N/A | N/A | 0 |
deepseek-v4-vision-nothinkingDeepSeek免费 | N/A | N/A | 0 |
mistralai/mixtral-8x22b-instruct-v0.1Mistral免费 | N/A | N/A | 0 |
deepseek-v4-pro-search-nothinkingDeepSeek免费 | N/A | N/A | 0 |
deepseek-v4-pro-searchDeepSeek免费 | N/A | N/A | 0 |
mistralai/mistral-medium-3-instructMistral免费 | N/A | N/A | 0 |
mistralai/mistral-medium-3.5-128bMistral免费 | N/A | N/A | 0 |
deepseek-v4-flash-search-nothinkingDeepSeek免费 | N/A | N/A | 0 |
moonshotai/kimi-k2.6Moonshot免费 | N/A | N/A | 0 |
指的是某家提供商上输入和输出价格都为 $0 / token 的模型组合。部分提供商提供长期免费额度,部分仅向新账号一次性赠送。只有在公开价格为零时才会被标记为免费,价格页会定期复核。
大多数有速率限制——每分钟请求数、每日总量、上下文长度上限——很多要求实名注册并绑定手机号或支付方式。部分是阶段性活动。在依赖某个免费端点之前,请仔细阅读提供商的条款和配额。
部分条目是社区运营的「公益站」或聚合中转——站长用自掏腰包的上游 key 打包再免费分发。它们通常额度更大、可用模型更全,但稳定性远低于官方 API:站长随时可能停服或「跑路」,额度和价格也可能毫无预警地变动;很多站点是邀请制,需要 GitHub 邀请码、社区推荐或特定圈子才能注册,部分长期关闭注册。建议把它们当作可遇不可求的备用通道,重要工作流仍走官方付费端点。
速度因模型和提供商而异。按「最多测试」排序找数据最稳的,或在上方按模型系列筛选。每行展示真实 API 测试得到的 tokens/秒 中位值与首字延迟。
对每家提供商使用相同 prompt 做五轮压力测试,使用 tiktoken 统计输出 token 数,测量吞吐量(tokens/秒)和首字到达时间。所有指标以中位数聚合以抵抗异常值,并按固定节奏刷新。
用于原型、个人项目和低流量工具是可以的。生产流量很快就会撞到速率限制。建议把免费额度当成「试用通道」:先验证模型和服务商表现,再切到同模型的付费端点上线。
可能目前没有提供商提供免费版本、免费活动已结束,或者尚未被基准测试覆盖。可以在该模型的详情页查看付费方案,或通过页脚反馈链接告诉我们漏了哪家。
查看 LMSpeed 最新收录并发布的大语言模型,进入模型页对比服务商、价格、性能与免费可用性。
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...