GuaiHub is a unified AI model aggregation and distribution gateway. It supports exposing upstream models through OpenAI-compatible, Claude-compatible, and Gemini-compatible interfaces.

Models
229 models
From
$0.0041/M
Speed
70 tok/s
Updated
6/8/2026
Latency
0.00 s
Created At
4/16/2026
Recharge Rate
¥1.00 per $1 quota

Features

DrawingTaskData ExportCheck-in

Login Methods

GitHubPasskeyGoogle

API Endpoints

  • Main Site Endpoint
    https://guaihub.com

    Stable and reliable, suitable for production environments

  • Backup Endpoint
    https://api.guaihub.cc

    Overseas optimized route for latency-sensitive scenarios

  • Endpoint 3Historical / Unverified
    https://api.guaihao.cn
  • Endpoint 4Historical / Unverified
    https://api2.guaihao.cn
Claim this provider

Verify ownership to unlock provider management features:

  • Edit provider name, content, and links
  • Get featured with priority traffic and visibility boost
  • Display a verified badge to build user trust

Leaderboard Rankings

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.

About GuaiHub

GuaiHub positions itself as a centralized AI model hub for aggregation and distribution. The site describes support for converting multiple large-model formats into OpenAI-compatible, Claude-compatible, and Gemini-compatible APIs for personal and enterprise use.

Health Check

100%Recent availability
History (72 pts)
PastNow

API Benchmarks & Pricing

Compare 185 model rows across audit recency, latest speed tests, throughput, latency, and per-token pricing.

Model
Input ($/M)
Output ($/M)
claude-opus-4-5-20251101-thinkingcc-sale
$0.548/M$2.74/M
claude-opus-4-5-20251101-thinkingcc
$1.37/M$6.85/M
codex-auto-reviewCodex
$0.171/M$1.03/M
doubao-seedance-2-0-260128Seedance-officially
$6.71/M$6.71/M
doubao-seedance-2-0-fast-260128Seedance-officially
$6.71/M$6.71/M
gemini-3-flash-agentgemini
$0.137/M$0.822/M
gemini-pro-agentgemini
$20.55/M$82.19/M
grok-4.20-0309-non-reasoningfree
$0.017/M$0.034/M
grok-4.20-0309-non-reasoninggrok
$0.146/M$0.291/M
grok-4.20-0309-reasoningfree
$0.017/M$0.034/M
grok-4.20-0309-reasoninggrok
$0.146/M$0.291/M
grok-composer-2.5-fastfree
$1.03/M$1.03/M
grok-composer-2.5-fastgrok
$8.73/M$8.73/M
grok-imagine-imagegrok
$0.0023/request-
grok-imagine-imagefree
$0.0003/request-
video-ds-2.0Seedance-en
$0.945/request-
video-ds-2.0-fastSeedance-en
$0.808/request-
$0.548/M$2.74/M
$1.37/M$6.85/M
claude-opus-5Claude-officially
$3.42/M$17.12/M
$0.411/M$2.47/M
$0.200/M$0.200/M
$1.70/M$1.70/M
kimi-k3Kimi-officially
$2.19/M$10.96/M
$5.14/M$30.82/M

Showing 25 of 185 model rows

Recent Test Records

TimeModelSpeedLatency
Apr 16, 07:01 AM
gpt-5.4-fast
69.82 tok/s
2.96s

Similar API Provider Alternatives to Compare

Compare GuaiHub alternatives against 6 nearby API providers using 711 LMSpeed signals across shared model coverage, pricing, benchmark speed, uptime, and free-model availability.

ProviderWhy compareModelsFreeAvg priceSpeed30d uptime
GuaiHub

guaihub

GuaiHub is a unified AI model aggregation and distribution gateway. It supports exposing upstream models through OpenAI-compatible, Claude-compatible, and Gemini-compatible interfaces.

Current provider baseline2290$0.0055/M70 tok/s9980%
钱多多 API

api2-aigcbest-top

Provides AI-generated content APIs for various applications, including text and image generation.

  • Lower average pricing
  • Faster measured speed
  • More free-model options
  • Broader model coverage
1,4182$0.0025/M101 tok/s9940%
向量引擎

api-vectorengine-ai

Vector Engine provides an API platform aggregating access to over 500 AI large models with OpenAI API compatibility and global deployment.

  • Lower average pricing
  • Faster measured speed
  • More free-model options
  • Broader model coverage
8052$0.0005/M89 tok/s9970%
艾可API

aicanapi-com

AicanAPI provides an enterprise-grade API management platform with high-performance architecture, security compliance, and cost optimization for business digital transformation.

  • Faster measured speed
  • More free-model options
  • Broader model coverage
545103N/A70 tok/s5890%
Koyeb AI Gateway

new-api-koyeb-app

An OpenAI-compatible API gateway deployed on Koyeb, providing access to multiple AI models.

  • More free-model options
  • Broader model coverage
578420N/A33 tok/s9960%
Fengsili API

api-fengsili-online

An OpenAI-compatible API relay service providing access to multiple AI models.

  • Faster measured speed
  • Broader model coverage
3800$3.00/M73 tok/s4770%
APIMart

apimart

APIMart is a pay-as-you-go, OpenAI-compatible AI API gateway operated by Hangzhou Huanzhi Network Technology Co., Ltd. It provides unified access to more than 500 third-party text, image, video, and audio models, with consolidated billing, multi-provider routing, and automatic failover.

  • More free-model options
  • Broader model coverage
25716$0.015/MN/A9980%

Announcements

default8/22/2026

[DeepSeek Peak/Off-Peak Billing Adjustment] DeepSeek-V4 now follows the latest official billing policy. Still in the deepseek-officially group, with a limited-time activity multiplier of 0.8*. Weekdays (Monday to Friday): Continue with the original peak/off-peak segmented billing. Weekends (Saturday and Sunday): No longer distinguish between peak and off-peak periods; all-day calls are charged at the off-peak rate. Thank you for your understanding and support!

default8/22/2026

[New Model] DeepSeek-V4-Flash-Vision-Exp model promotion launched A new model, DeepSeek-V4-Flash-Vision-Exp, has been added to the DeepSeek official group. It is now open for calls, with a limited-time 20% discount. For specific billing rates and promotion period, please refer to the platform billing page.

default8/16/2026

【DeepSeek Price Adjustment】DeepSeek-V4 has aligned with the official billing strategy. This adjustment does not affect existing API calls. After the adjustment, API prices have increased and peak/off-peak pricing is now in effect, with off-peak prices at half of peak-hour prices. We recommend planning your business calls accordingly to optimize usage.

default8/13/2026

[New Model] Grok 4.6 is now available.

default7/31/2026

[GPT 5.6 Price Reduction] OpenAI GPT 5.6 has been price-synced with the official price cut. New prices take effect immediately. This adjustment does not affect account balance or billing method.

default7/26/2026

[New Model] claude-opus-5 groups are now available.

default7/23/2026

[Important Notice] codex-sale group will be removed. Due to cost impact from O's control policies and quota reduction, service availability has declined. To ensure stability and continuity, we will temporarily remove the codex-sale group starting July 25, 2026 at 12:00. We will monitor policy changes—if O tightens controls further, rates may increase; if the situation eases, we will restore it promptly. As a suggestion, you can use the codex group for now.

default7/21/2026

[New Model] Kimi-K3 is now available - Moonshot AI flagship model.

default7/11/2026

[Promotion] Valid until July 31, 2026. All GuaiHub users get a 20% discount on evening model calls from 21:00 to 8:00 (Asia/Shanghai). Covered models: gpt-5.6-luna, gpt-5.6-terra, gpt-5.6-sol. When conditions are met, the total order price is multiplied by 0.8. Check usage logs for details. Thank you for your support!

default7/9/2026

[New Model] Grok-4.5 is now available on the model plaza. Grok 4.5 is SpaceXAI's smartest model for coding, intelligent tasks, and knowledge work.

default7/8/2026

[Gemini Group Change] The original Gemini group (regular channel) has been replaced with Gemini (stable version). The multiplier has been updated from 0.6x to 2x. Please note the group change. The Gemini-SLB group has been removed.

default7/7/2026

🚀 [Server Upgrade Complete] To handle recent high concurrency growth and improve stability and response speed, we have completed a server upgrade today. This upgrade optimizes server performance, network bandwidth, and concurrent processing capabilities. After the upgrade, request handling and response speed during peak times will be significantly improved, providing a more stable and smooth experience.

default7/2/2026

[bailian] Group Removal Notice. The bailian group has completed its mission and will be officially removed on July 3. Please migrate to other groups as soon as possible.

New groups available: [Minimax-officially] Official channel group - multiplier 0.5x [Kimi-officially] Official channel group - multiplier 0.8x

default6/23/2026

[New Groups] Zai-sale and Zai-officially groups are now open. To enrich model selection, billing multipliers are x0.6 and x0.9 respectively. Included models: GLM-4.7, GLM-5, GLM-5.2. GLM-5.2 is the latest flagship long-cycle task model. GuaiHub has completed adaptation; both groups are ready for use. Thank you for your understanding and support!

default6/13/2026

[Model Removal] Due to US government export control directives, Anthropic has banned all users from using Claude Fable 5 and Claude Mythos 5. We have removed them in sync with the official decision.

default6/7/2026

[API Request Chain Optimized] We have recently completed an API request chain optimization, including Cloudflare access configuration, interface caching strategy, HTTPS security chain, static resource caching, and basic anti-scraping rate limiting adjustments. This optimization will further improve interface access stability, reduce streaming response lag, connection interruptions, and abnormal request interference. We will continue to optimize based on actual usage.

FAQ

Why does GuaiHub offer discounts compared to buying directly?

We serve users globally and leverage the combined API call volume of all users on our platform to negotiate exclusive discounts with model providers. Through our collective purchasing power, we secure the lowest prices that individual or enterprise users cannot obtain directly, and pass these savings directly to every user.

Is there any difference in quality and speed compared to calling the official API directly?

There is no difference. We essentially request the official APIs, and through the deployment of a global acceleration network, we can significantly improve the response speed and stability of API requests, ensuring it is more efficient and faster than when you call independently.

What is quota? How is it calculated?

## Quota Calculation Rules Quota consumption is calculated separately for input, output, and cache reads. `Quota consumption = Group multiplier × ((Input Tokens × Input unit price + Cache Tokens × Cache read unit price + Output Tokens × Output unit price) / 1M)` **Explanation:** - `Input Tokens`: Tokens consumed by user requests - `Output Tokens`: Tokens consumed by model responses - `Cache Tokens`: Tokens that hit context cache and are billed at the cache read price - `Group multiplier`: The price multiplier for the current group **Note:** - Different models may have different unit prices for input, output, and cache reads - For non-streaming requests, the upstream may only return total tokens; if input/output breakdown is missing, the actual deduction is subject to the platform's final settlement - The displayed billing result is for reference only; the actual deduction prevails

A message from the site owner

<p>From the very beginning, we never intended to follow the path of "ultra-low prices and cutthroat competition."</p> <p>GuaiHub places more importance on long-term stability of service, consistent availability of routes, and transparent and stable pricing. Low prices are certainly attractive, but if the cost is frequent fluctuations and unreliable availability, that is not what we want to do.</p> <p>Regarding data security, we have always been cautious. The platform mainly handles request forwarding, authentication, and billing. When connecting directly to official providers, we focus more on link stability and service quality rather than the specific content users send.</p> <p>We always prioritize high availability and stable service. Within reasonable profit margins, we strive to keep both service and pricing as stable as possible. GuaiHub may not be the cheapest, but we will do our best to be serious, reliable, and dedicated, and to live up to everyone's trust.</p>

Notes

  • Health checks: Scope: the 72-hour chart and recent availability measure API connectivity only. Each bar summarizes one hour of checks. Targets: LMSpeed tries the configured health check URL and provider status URL first, then API endpoints derived from known API hosts and recent speed-test base URLs. A website host is considered only when it looks like an API endpoint. Probe steps: each candidate goes through DNS lookup, TCP connection, TLS handshake for HTTPS, and an HTTP HEAD request with redirects followed. Probing stops after the first reachable candidate. Reachable criteria: every required network step must succeed. An HTTP response below 500 is treated as reachable, including 401 because it confirms that an authenticated API endpoint responded, except for statuses classified as blocked. Blocked results: HTTP 403, 429, 521, 525, and 530, plus detected WAF or Cloudflare challenges, are shown as blocked and excluded from availability calculations because LMSpeed cannot determine whether the API itself is down. Model availability: when a dedicated test key is configured, LMSpeed sends an authenticated GET request to a derived /models endpoint and compares returned model IDs with this provider's listed models. These per-model results appear in Models & Pricing and are not included in the provider connectivity percentage. Timeouts: TCP connection, TLS handshake, HTTP connectivity, and model requests each use a 20-second timeout. A full run can take longer when several candidates are tried. Frequency: a background worker checks all providers every 5 minutes by default. The 72-hour chart combines those samples into hourly bars, and the schedule may be changed by the service operator. Limit: automated samples are not an SLA and do not guarantee account quota, every model, every region, or successful completion requests. Check the provider's own status page before making operational decisions.
  • Domain Rating data is sourced from Ahrefs. It is a 0–100 backlink-based domain strength signal and does not measure API speed or reliability.
  • Announcements and FAQ are read from this provider's NewAPI status snapshot when available. LMSpeed stores the original content and optional English translations from the provider status source, then shows the localized fields on this page.