0CHAT provides an AI API gateway for accessing multiple chat models through unified endpoints.

Models
44 models
From
$0.025/M
Speed
51 tok/s
Updated
6/8/2026
Latency
0.00 s
Created At
3/23/2026
Recharge Rate
¥0.35 per $1 quota

Features

DrawingTaskData ExportCheck-in

API Endpoints

  • Optimized Route
    https://api.0chat.vip

    CDN load balancing

  • Endpoint 2
    https://api1.0chat.vip

    CN2 optimized route

  • Endpoint 3
    https://api2.0chat.vip

    Cloudflare international route

Claim this provider

Verify ownership to unlock provider management features:

  • Edit provider name, content, and links
  • Get featured with priority traffic and visibility boost
  • Display a verified badge to build user trust

Leaderboard Rankings

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.

About 0CHAT

0CHAT is a unified API gateway that aggregates access to various large language models (LLMs) and AI services from multiple providers. It offers standardized endpoints compatible with OpenAI's API format, allowing developers to switch between different models by changing the base URL.

Key capabilities:

  • Chat completions via /v1/chat/completions, /v1/responses, and /v1/messages endpoints
  • Text embeddings through /v1/embeddings
  • Image generation and editing with /v1/images/generations, /v1/images/edits, and /v1/images/variations
  • Audio processing including speech synthesis (/v1/audio/speech), transcription (/v1/audio/transcriptions), and translation (/v1/audio/translations)
  • Model listing via /v1beta/models
  • Text reranking with /v1/rerank

Supported providers include: MoonshotAI, OpenAI, Grok, Zhipu, Volcengine, Cohere, Claude, Gemini, Suno, Minimax, Wenxin, Spark, Qingyan, DeepSeek, Qwen, Midjourney, AzureAI, Hunyuan, Xinference, and others.

Notable features: No subscription required, pay-as-you-go pricing model with competitive rates mentioned as 'better price', and emphasis on stability. The service positions itself as a drop-in replacement for existing OpenAI API implementations.

Health Check

99%Recent availability
History (72 pts)
PastNow

API Benchmarks & Pricing

Compare 32 model rows across audit recency, latest speed tests, throughput, latency, and per-token pricing.

Model
Input ($/M)
Output ($/M)
Speed
Latency
codex-auto-reviewdefault
$2.52/M$2.52/M--
codex-auto-reviewOpenAI
$2.52/M$2.52/M--
$0.0034/request---
$0.0034/request---
$0.029/request---
$0.029/request---
$0.046/request---
$0.046/request---
$0.192/M$0.575/M--
gpt-5.5OpenAI
$0.168/M$1.01/M--
gpt-5.5default
$0.168/M$1.01/M--
$1.53/M$9.21/M--
$0.025/M$0.151/M--
$0.025/M$0.151/M--
gpt-5.4default
$0.084/M$0.503/M51.0 t/s2.53 s
gpt-5.4OpenAI
$0.084/M$0.503/M51.0 t/s2.53 s
$0.059/M$0.470/M--
$0.059/M$0.470/M--
$0.059/M$0.470/M--
$0.059/M$0.470/M--
$1.53/M$9.21/M--
$0.384/M$2.30/M--
gpt-5.2OpenAI
$0.059/M$0.470/M--
gpt-5.2default
$0.059/M$0.470/M--
$0.0077/request---

Showing 25 of 32 model rows

Recent Test Records

TimeModelSpeedLatency
Mar 23, 04:10 PM
gpt-5.4
47.17 tok/s
2.77s
Mar 23, 04:07 PM
gpt-5.4
54.81 tok/s
2.29s

Similar API Provider Alternatives to Compare

Compare 0CHAT alternatives against 6 nearby API providers using 292 LMSpeed signals across shared model coverage, pricing, benchmark speed, uptime, and free-model availability.

ProviderWhy compareModelsFreeAvg priceSpeed30d uptime
0CHAT

api-0chat-vip

0CHAT provides an AI API gateway for accessing multiple chat models through unified endpoints.

Current provider baseline440$0.088/M51 tok/s9970%
钱多多 API

api2-aigcbest-top

Provides AI-generated content APIs for various applications, including text and image generation.

  • Lower average pricing
  • Faster measured speed
  • More free-model options
  • Broader model coverage
1,4182$0.0025/M101 tok/s9940%
TommyLam API

new-api-tommylam-me

TommyLam API offers an OpenAI-compatible API gateway for accessing multiple AI models at competitive rates.

  • Faster measured speed
  • Higher 30-day availability
  • More free-model options
  • Broader model coverage
231100N/A61 tok/s10000%
PackyAPI

codex-api-packycode-com

PackyAPI (codex-api.packycode.com) is an OpenAI-compatible API relay for Codex and other models via a single interface.

  • Lower average pricing
  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
780$0.041/M69 tok/s9980%
F2API

api-f2api-com

F2API is an enterprise AI gateway providing a unified OpenAI-compatible interface for accessing multiple large language models.

  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
1,2620N/A272 tok/s9990%
哈基米API站

api-gemai-cc

哈基米API站 offers a unified API gateway for large language models with competitive pricing and stable service.

  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
3230$0.930/M119 tok/s9980%
ChooseC API

ipv4-beta-kxcym-top-3001

ChooseC API is an OpenAI-compatible relay service hosted at ipv4-beta.kxcym.top, providing access to multiple AI models through standardized endpoints.

  • Faster measured speed
  • More free-model options
  • Broader model coverage
1229$0.095/M51 tok/s0%

Announcements

default5/11/2026

Important Notice: GPT Group Multiplier Adjustment

Dear Users,

We are adjusting the GPT grouping policy:

Effective immediately, the GPT group multiplier is uniformly adjusted to 0.7 (0.7x).

This adjustment has taken effect. Any future changes will be announced via on-site notices. Please refer to the platform's real-time display.

Thank you for your understanding and support.

warning4/26/2026

📢 Recharge Suspension Notice

Dear Users:

Due to business adjustments, the recharge function is temporarily suspended starting immediately, with the restoration time to be determined.

Please rest assured that your existing balance and API calls are not affected and can be used normally.

We apologize for any inconvenience. If you have any questions, please contact the purchase point.

Recharge Suspension Notice

default4/16/2026

[Important Notice] Explanation of GPT Group Rate Adjustment

Dear Users:

The GPT group strategy is being adjusted:

Effective immediately, the GPT group rate is uniformly adjusted to 0.7 (0.7x).

This adjustment has taken effect. Any future changes will be announced via on-site notices. Please refer to the platform's real-time display for accuracy.

Thank you for your understanding and support.

default4/14/2026

[Important Notice] Explanation of GPT Group Rate Adjustment Dear Users:

After platform resource coordination and optimization, to ensure stable service operation, the GPT group strategy is being adjusted:

Effective immediately, the GPT group rate is restored to 1.

This adjustment has taken effect. Any future changes will be announced via on-site notices. Please refer to the platform's real-time display for accuracy.

Thank you for your understanding and support.

ongoing4/6/2026

[Notice] GPT Rate Increase and Refund Explanation

Due to stricter OpenAI risk controls, to maintain service stability, effective immediately, the GPT group rate is uniformly increased to 1.25. We apologize for any inconvenience.

If dissatisfied with the adjustment, you can apply for a refund:

  1. Original channel orders: Please apply at the original purchase point.
  2. Site payment orders: Please send an email to [email protected] to process.

Thank you for your understanding and support!

success4/1/2026

grok official API temporarily available, free to use, only $50, until exhausted.

error3/28/2026

[Important Notice] Explanation of GPT Group Rate Adjustment Dear Users:

After platform resource coordination and optimization, to ensure stable service operation, the GPT group strategy is being adjusted:

Effective immediately, the GPT group rate is restored to 0.5.

This adjustment has taken effect. Any future changes will be announced via on-site notices. Please refer to the platform's real-time display for accuracy.

Thank you for your understanding and support.

success3/27/2026

[Important Notice] Explanation of GPT Group Rate Adjustment

Dear Users:

The GPT group strategy is being adjusted:

Effective immediately, the GPT group rate is uniformly adjusted to 1 (1x).

This adjustment has taken effect. Any future changes will be announced via on-site notices. Please refer to the platform's real-time display for accuracy.

Thank you for your understanding and support.

warning3/27/2026

[Important Notice] Preview of GPT Group Rate Adjustment

Dear Users:

Due to OpenAI risk controls and ban situations, the platform's available resources have fluctuated recently, and the access and maintenance costs for GPT series models continue to rise. To ensure the stability and sustainable operation of API services, the platform may dynamically adjust the GPT group rate based on resource conditions in the future.

Thank you for your understanding and support.

success3/20/2026

The API site has completed GPT5.4 series optimization. This time, not only is Fast mode enabled by default in Codex, but also the API call chains for all channels have been optimized. Currently, the speed of calling GPT5.4 series APIs from all channels has significantly improved.

Furthermore, this optimization does not incur any additional fees, and users can enjoy a faster calling experience without extra payment.

ongoing3/17/2026

A temporary adjustment is being made to OpenAI related model rates:

  • Adjusted rate: 0.5
  • Which is 50% of the original rate (half price)

This adjustment will cover all OpenAI related models on the platform. During the OpenAI event, model call costs will be settled according to the new rate standard.

default3/16/2026

Currently, the GPT series model channels are sourced from Codex, billed at 1x rate; Gemini series model channels are sourced from GCP, billed at 4x rate due to higher access costs. Users can select the corresponding group when generating tokens to use different model services as needed.

Notes

  • Health checks: Scope: the 72-hour chart and recent availability measure API connectivity only. Each bar summarizes one hour of checks. Targets: LMSpeed tries the configured health check URL and provider status URL first, then API endpoints derived from known API hosts and recent speed-test base URLs. A website host is considered only when it looks like an API endpoint. Probe steps: each candidate goes through DNS lookup, TCP connection, TLS handshake for HTTPS, and an HTTP HEAD request with redirects followed. Probing stops after the first reachable candidate. Reachable criteria: every required network step must succeed. An HTTP response below 500 is treated as reachable, including 401 because it confirms that an authenticated API endpoint responded, except for statuses classified as blocked. Blocked results: HTTP 403, 429, 521, 525, and 530, plus detected WAF or Cloudflare challenges, are shown as blocked and excluded from availability calculations because LMSpeed cannot determine whether the API itself is down. Model availability: when a dedicated test key is configured, LMSpeed sends an authenticated GET request to a derived /models endpoint and compares returned model IDs with this provider's listed models. These per-model results appear in Models & Pricing and are not included in the provider connectivity percentage. Timeouts: TCP connection, TLS handshake, HTTP connectivity, and model requests each use a 20-second timeout. A full run can take longer when several candidates are tried. Frequency: a background worker checks all providers every 5 minutes by default. The 72-hour chart combines those samples into hourly bars, and the schedule may be changed by the service operator. Limit: automated samples are not an SLA and do not guarantee account quota, every model, every region, or successful completion requests. Check the provider's own status page before making operational decisions.
  • Domain Rating data is sourced from Ahrefs. It is a 0–100 backlink-based domain strength signal and does not measure API speed or reliability.
  • Announcements and FAQ are read from this provider's NewAPI status snapshot when available. LMSpeed stores the original content and optional English translations from the provider status source, then shows the localized fields on this page.