Cerebras提供基于云的AI API,专注于大语言模型和其他AI工作负载的高性能推理和训练。该平台利用Cerebras的定制硬件架构,特别是晶圆级引擎(WSE),旨在加速计算密集型的AI任务。主要功能包括模型服务、批量处理和可扩展的训练管道。典型用例涉及企业和研究人员运行需要专用硬件加速的AI应用。
api.cerebras.ai完成站点认证,解锁以下权益:
Cerebras提供基于云的AI API,专注于大语言模型和其他AI工作负载的高性能推理和训练。该平台利用Cerebras的定制硬件架构,特别是晶圆级引擎(WSE),旨在加速计算密集型的AI任务。主要功能包括模型服务、批量处理和可扩展的训练管道。典型用例涉及企业和研究人员运行需要专用硬件加速的AI应用。
按 1 个模型行对比最新 audit、最新测速、吞吐、延迟与按 token 计费价格。
| 模型 | 输入 ($/M) | 输出 ($/M) | 速度 | 延迟 |
|---|---|---|---|---|
| - | - | 400.0 t/s | 3.57 s |
当前显示 1 / 1 个模型行
| 时间 | 模型 | 速度 | 延迟 |
|---|---|---|---|
| Jan 13, 04:32 PM | zai-glm-4.7 | 400.04 tok/s | 3.57s |
排名基于社区提交的测试数据与定期健康探测,仅供参考,非官方数据。
用 59 个 LMSpeed 信号,把 Cerebras 的 6 个相近 API 替代服务商放在一起比较:共享模型覆盖、价格、实测速度、可用性和免费模型。
| 服务商 | 对比理由 | 模型数 | 免费项 | 均价 | 速度 | 30 天可用性 |
|---|---|---|---|---|---|---|
| Cerebras api-cerebras-ai Provides AI inference and training APIs leveraging Cerebras hardware for large-scale model deployment. | 当前服务商基线 | 1 | 0 | N/A | 400 tok/s | 99.8% |
| OpenRouter openrouter A unified API interface providing access to over 300 models from 60+ providers, including OpenAI, Anthropic, and Google. |
| 265 | 0 | N/A | 85 tok/s | 99.9% |
huawei-modelarts Huawei ModelArts MaaS (Model as a Service) platform offering Pangu series models and third-party LLMs via OpenAI-compatible API. |
| 11 | 0 | N/A | 31 tok/s | 99.9% |
api-kriora-com Provides OpenAI-compatible APIs and managed GPU instances for deploying and scaling open-source AI models. |
| 9 | 0 | N/A | 565 tok/s | 99.7% |
api-suanli-cn Gongji Compute (共绩算力) provides pay-as-you-go, standard API access to mainstream LLMs and elastic GPU computing for AI applications. |
| 6 | 5 | N/A | 36 tok/s | 63.8% |
chutes Chutes provides LLM inference API service, offering access to various open-source AI models through OpenAI-compatible endpoints. |
| 2 | 0 | N/A | 27 tok/s | 99.9% |
siliconflow Provides cost-effective generative AI cloud services based on open-source models for text, image, video, and audio generation. |
| 62 | 0 | N/A | 56 tok/s | 56.4% |