An enterprise-grade API platform enabling unified access to over 50 AI models, including OpenAI, Claude, and Gemini.

Models
70 models
From
$0.0082/M
Speed
47 tok/s
Updated
8/11/2026
Latency
0.00 s
Created At
4/25/2026
Recharge Rate
¥1.00 per $1 quota

Features

DrawingTaskData ExportCheck-in

Login Methods

GitHubDiscordLinuxDO

Billing

Payment methods

AlipayWeChat PayStripeApple PayCredit cardGoogle Pay

Billing types

Usage-based
Refunds
Supported

API Endpoints

  • Endpoint 1Historical / Unverified
    https://www.uocode.com
  • Endpoint 2Historical / Unverified
    https://us.uocode.com
  • Endpoint 3Historical / Unverified
    https://uo.mentoe.com
Claim this provider
HVerified by hongdong tian

Leaderboard Rankings

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.

About UoCode

UoCode is an enterprise-level API aggregation and management platform. It provides a single interface for developers to call on over 50 leading AI models, including OpenAI's GPT series, Anthropic's Claude series, Google's Gemini, Qwen, Llama, and Midjourney. The service is fully compatible with the OpenAI API format, requiring developers to only change the base_url and api_key in their existing OpenAI SDK code.

Key features include a unified routing system for easy model switching, a management console for user, token, and quota administration, and detailed logging with token-based billing. The platform advertises a 99.9% service availability SLA and supports integration with popular applications like NextChat, LobeChat, and Dify.

Health Check

100%Recent availability
History (72 pts)
PastNow

API Benchmarks & Pricing

Compare 93 model rows across audit recency, latest speed tests, throughput, latency, and per-token pricing.

Model
Input ($/M)
Output ($/M)
Model availability
Speed
Latency
claude-opus-4-6-thinkingClaude_K1通道
$0.171/M$0.856/M
100%
--
claude-opus-4-7-thinkingClaude_K1通道
$0.171/M$0.856/M
100%
--
gemini-2.5-flash-thinkingGoogle Gemini
$0.0082/M$0.192/M
100%
--
gpt-5.6-lunaGTP Codex 全家桶
$0.027/M$0.164/M
100%
--
gpt-5.6-lunaGPT Codex
$0.029/M$0.173/M
100%
--
$0.034/M$0.205/M
100%
--
gpt-5.6-lunaGPT全家桶-日卡专用
$0.137/M$0.822/M
100%
--
gpt-5.6-terraGTP Codex 全家桶
$0.068/M$0.411/M
100%
--
gpt-5.6-terraGPT Codex
$0.072/M$0.432/M
100%
--
$0.086/M$0.514/M
100%
--
gpt-5.6-terraGPT全家桶-日卡专用
$0.342/M$2.05/M
100%
--
gpt-5.6-solGTP Codex 全家桶
$0.137/M$0.822/M
100%
--
gpt-5.6-solGPT Codex
$0.144/M$0.863/M
100%
--
$0.171/M$1.03/M
100%
--
gpt-5.6-solGPT全家桶-日卡专用
$0.685/M$4.11/M
100%
--
gpt-5.5-openai-compactGTP Codex 全家桶
$0.137/M$0.822/M
100%
--
gpt-5.5-openai-compactGPT Codex
$0.144/M$0.863/M
100%
--
gpt-5.5-openai-compactGPT全家桶-日卡专用
$0.685/M$4.11/M
100%
--
gpt-5.5GTP Codex 全家桶
$0.137/M$0.822/M
100%
39.1 t/s3.40 s
gpt-5.5GPT Codex
$0.144/M$0.863/M
100%
39.1 t/s3.40 s
gpt-5.5gptpro
$0.171/M$1.03/M
100%
39.1 t/s3.40 s
gpt-5.5GPT全家桶-日卡专用
$0.685/M$4.11/M
100%
39.1 t/s3.40 s
gpt-5.4-miniGTP Codex 全家桶
$0.021/M$0.123/M
100%
--
gpt-5.4-miniGPT Codex
$0.022/M$0.129/M
100%
--
$0.026/M$0.154/M
100%
--

Showing 25 of 93 model rows

Recent Test Records

TimeModelSpeedLatency
May 13, 11:48 AM
claude-opus-4-7
54.59 tok/s
2.78s
May 13, 11:46 AM
claude-opus-4-6
38.35 tok/s
2.61s
May 13, 11:46 AM
claude-haiku-4-5-20251001
96.62 tok/s
1.34s
May 13, 11:44 AM
gpt-5.4
31.43 tok/s
2.36s
May 13, 11:42 AM
gpt-5.5
39.11 tok/s
3.40s
Apr 26, 06:28 AM
claude-sonnet-4-6
39.99 tok/s
5.50s
Apr 26, 06:24 AM
claude-sonnet-4-6
18.72 tok/s
6.52s
Apr 26, 06:19 AM
claude-sonnet-4-6
39.07 tok/s
6.38s
Apr 26, 06:14 AM
claude-sonnet-4-6
30.16 tok/s
4.64s
Apr 26, 06:13 AM
claude-sonnet-4-6
31.47 tok/s
7.32s

Similar API Provider Alternatives to Compare

Compare UoCode alternatives against 6 nearby API providers using 384 LMSpeed signals across shared model coverage, pricing, benchmark speed, uptime, and free-model availability.

ProviderWhy compareModelsFreeAvg priceSpeed30d uptime
UoCode

uocode

An enterprise-grade API platform enabling unified access to over 50 AI models, including OpenAI, Claude, and Gemini.

Current provider baseline700$0.021/M47 tok/s9960%
钠 API

naapi-cc

Na API (naapi.cc) is an OpenAI-compatible LLM API gateway with competitive pricing and stable access to 100+ models from OpenAI, Anthropic, Google, and more.

  • Lower average pricing
  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
  • Same provider category
5660$0.0031/M649 tok/s9970%
Cuz AI

ai-cuz-lab-space

Cuz AI runs an OpenAI-compatible relay at ai.cuz-lab.space with broad model coverage, public pricing, and stable throughput for chat and coding workloads.

  • Lower average pricing
  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
  • Same provider category
3930$0.020/M170 tok/s9980%
CatClaw API

www-catclawai-top

CatClaw API provides an AI model relay with access to GPT, Claude, and other large language models.

  • Faster measured speed
  • Higher 30-day availability
  • More free-model options
  • Broader model coverage
  • Same provider category
1292N/A83 tok/s9990%
初叶🍂Furry API

ai-chuyel-top

A free, community-supported API service providing access to various AI models for developers and users.

  • Faster measured speed
  • More free-model options
  • Broader model coverage
  • Same provider category
1,021152$18.00/M171 tok/s2670%
天絮 API

tianxu-api

Tianxu API provides an AI model relay service with multiple access points and stable connectivity.

  • Lower average pricing
  • Faster measured speed
  • Broader model coverage
  • Same provider category
7590$0.014/M93 tok/s9910%
Dext API

ai-dext-top

Dext API is an OpenAI-compatible API relay offering access to 500+ LLM models at competitive prices, with multi-channel routing and broad model coverage.

  • Lower average pricing
  • More free-model options
  • Broader model coverage
  • Same provider category
667219$0.017/M19 tok/s3010%

Announcements

default8/30/2026

The multiplier for the GTP Codex family bucket has been lowered because we recently adjusted pricing. Please check the pricing page for the full price. We have also launched two new models: grok-4.5 and grok-4.6.

default8/29/2026

The group monitoring page is currently under maintenance, expected to take 1–2 days. Please be patient during this time, normal usage will not be affected.

default8/27/2026

GPT Codex family bundle is a special discount channel, with average stability. The price has been reduced. GPT Codex regular channel has better stability, and the price remains unchanged.

default8/22/2026

Performance degradation: We are currently investigating stability issues with the Claude_K1 and AG channels. Please bear with us while we work to restore full functionality.

default8/10/2026

The server became inaccessible about 1 hour ago, but has now been restored to normal and can be used normally.

default8/8/2026

Currently, because the API address of this site cannot be used normally, we are now officially switching the API call address. There are currently two available addresses:

1: https://ncapi.eu.cc 2: https://apinew.eu.cc

Please replace the API address in time and continue using.

default8/6/2026

Due to poor stability of this model, we have decided to take gpt-5.6-luna offline. Other models are unaffected and can be used normally.

default7/24/2026

CC Max and CC Max_1 channel prices have increased by 0.2. Please go to the pricing plaza for detailed rates.

default7/23/2026

Dear users, due to rising upstream costs and market fluctuations, GPT service prices will be increased from today. Please visit the pricing plaza for detailed rates.

default7/21/2026

Due to the current unstable state of GPT services, we will adjust prices at 21:00 GMT+8 on July 21. After the adjustment, please visit the pricing page to see the latest prices. No further notice will be given. If the service stabilizes, we will lower prices according to the actual situation.

default7/20/2026

Models under the Google Gemini group are now relisted, using a pay-as-you-go billing model. For specific pricing details, please go to the pricing plaza.

default7/12/2026

The multiplier prices for the CC Channel 2 and Claude_C3 groups have been reduced. Please go to the pricing plaza for specific prices.

default7/11/2026

It is forbidden to use gpt-5.4-mini for immersive translation, distillation tasks, or non-streaming calls.
If discovered, the account will be banned directly.

Such usage consumes a large amount of quota and affects normal use by other users, so it is not allowed.

default7/10/2026

Final notice: When using WeChat Pay, please do not select Apple Pay, Google Pay, or bank card (Credit/Debit Card) payment methods. If any of the above methods are used for payment, the corresponding handling fee will be deducted manually.

Please use only WeChat Pay to complete payment. Apple Pay, Google Pay, and bank card (Credit/Debit Card) payments are not supported.

default7/2/2026

We have launched waffo Pancake payment, with a minimum payment amount of $30. To facilitate normal use for Hong Kong users, we have also launched WeChat Pay payment with no handling fee. Alipay+ currently requires a service provider fee of 2.31 HKD + 3%; while waffo Pancake internal payment has no handling fee.

default7/2/2026

CC_MAX price increased by $0.1. Please check the pricing plaza for details.

default7/1/2026

CC_MAX price has been reduced by $0.1. Please check the pricing plaza for details; some groups have launched the Claude Sonnet 5 model.

default6/26/2026

GTP Codex family bundle price has been reduced for the second time today.

Notes

  • Health checks: Scope: the 72-hour chart and recent availability measure API connectivity only. Each bar summarizes one hour of checks. Targets: LMSpeed tries the configured health check URL and provider status URL first, then API endpoints derived from known API hosts and recent speed-test base URLs. A website host is considered only when it looks like an API endpoint. Probe steps: each candidate goes through DNS lookup, TCP connection, TLS handshake for HTTPS, and an HTTP HEAD request with redirects followed. Probing stops after the first reachable candidate. Reachable criteria: every required network step must succeed. An HTTP response below 500 is treated as reachable, including 401 because it confirms that an authenticated API endpoint responded, except for statuses classified as blocked. Blocked results: HTTP 403, 429, 521, 525, and 530, plus detected WAF or Cloudflare challenges, are shown as blocked and excluded from availability calculations because LMSpeed cannot determine whether the API itself is down. Model availability: when a dedicated test key is configured, LMSpeed sends an authenticated GET request to a derived /models endpoint and compares returned model IDs with this provider's listed models. These per-model results appear in Models & Pricing and are not included in the provider connectivity percentage. Timeouts: TCP connection, TLS handshake, HTTP connectivity, and model requests each use a 20-second timeout. A full run can take longer when several candidates are tried. Frequency: a background worker checks all providers every 5 minutes by default. The 72-hour chart combines those samples into hourly bars, and the schedule may be changed by the service operator. Limit: automated samples are not an SLA and do not guarantee account quota, every model, every region, or successful completion requests. Check the provider's own status page before making operational decisions.
  • Domain Rating data is sourced from Ahrefs. It is a 0–100 backlink-based domain strength signal and does not measure API speed or reliability.
  • Announcements and FAQ are read from this provider's NewAPI status snapshot when available. LMSpeed stores the original content and optional English translations from the provider status source, then shows the localized fields on this page.