Sponsored byFusecodeEnterprise coding API for Claude Code, Codex, and model workflows.
LogoLMSpeed
  • Free
  • Models
  • Providers
  • Leaderboard
  • Docs
LogoLMSpeed
  1. Home
  2. Providers
  3. Yunwu AI
LogoLMSpeed

The best API speed test tool

GitHubGitHubTwitterX (Twitter)Email
Product
  • Features
  • Pricing
  • FAQ
Leaderboard
  • Overview
  • Speed Ranking
  • Latency Ranking
  • Health Ranking
  • Model Pricing
  • Model Speed
  • Reasoning
  • Coding
Models
  • All Models
  • GPT
  • Claude
  • Gemini
  • DeepSeek
  • Llama
  • Qwen
Free Models
  • All Free Models
  • Free GPT
  • Free Claude
  • Free Gemini
  • Free DeepSeek
  • Free Llama
  • Free Qwen
Tools
  • Speed Test
  • Provider Audit
Company
  • About
Resources
  • Provider Directory
  • Documentation
  • Public API
  • Botab
  • VidBee
Legal
  • Cookie Policy
  • Privacy Policy
  • Terms of Service
© 2026 LMSpeed All Rights Reserved.Made by Nexmoe with ❤️

YUNWU API

X (Twitter)Share on X
yunwu.ai
Models
1997 models
From
$0.0002/request
Speed
115 tok/s
Updated
6/8/2026

A unified API gateway providing access to multiple large language models with direct connectivity in China.

OverviewHealthBenchmarks1997Tests135AlternativesEmbed
Latency4.41 s
Created At8/13/2025
Recharge Rate¥0.50 per $1 quota

Features

DrawingTaskData ExportCheck-in

Login Methods

GitHubLinuxDO
Website

API Endpoints

  • Main Site
    https://yunwu.ai

    US high-defense load balancing site cluster

  • CF Site
Claim this provider

Verify ownership to unlock provider management features:

  • Edit provider name, content, and links
  • Get featured with priority traffic and visibility boost
  • Display a verified badge to build user trust

Notes

  • Health checks: Scope: the 72-hour chart and recent availability measure API connectivity only. Each bar summarizes one hour of checks. Targets: LMSpeed tries the configured health check URL and provider status URL first, then API endpoints derived from known API hosts and recent speed-test base URLs. A website host is considered only when it looks like an API endpoint. Probe steps: each candidate goes through DNS lookup, TCP connection, TLS handshake for HTTPS, and an HTTP HEAD request with redirects followed. Probing stops after the first reachable candidate. Reachable criteria: every required network step must succeed. An HTTP response below 500 is treated as reachable, including 401 because it confirms that an authenticated API endpoint responded, except for statuses classified as blocked. Blocked results: HTTP 403, 429, 521, 525, and 530, plus detected WAF or Cloudflare challenges, are shown as blocked and excluded from availability calculations because LMSpeed cannot determine whether the API itself is down. Model availability: when a dedicated test key is configured, LMSpeed sends an authenticated GET request to a derived /models endpoint and compares returned model IDs with this provider's listed models. These per-model results appear in Models & Pricing and are not included in the provider connectivity percentage. Timeouts: TCP connection, TLS handshake, HTTP connectivity, and model requests each use a 20-second timeout. A full run can take longer when several candidates are tried. Frequency: a background worker checks all providers every 5 minutes by default. The 72-hour chart combines those samples into hourly bars, and the schedule may be changed by the service operator. Limit: automated samples are not an SLA and do not guarantee account quota, every model, every region, or successful completion requests. Check the provider's own status page before making operational decisions.
  • Domain Rating data is sourced from Ahrefs. It is a 0–100 backlink-based domain strength signal and does not measure API speed or reliability.
  • Announcements and FAQ are read from this provider's NewAPI status snapshot when available. LMSpeed stores the original content and optional English translations from the provider status source, then shows the localized fields on this page.
https://api.apiplus.org

Global CDN, average domestic access speed

  • No-Login Page API Interface
    https://api.zhongzhuan.chat

    Convenient for reselling KEYs

  • Domestic Server
    https://api3.wlai.vip

    Domestic high-defense server line

  • Endpoint 5Historical / Unverified
    https://yunwu.zeabur.app
  • Data as of Jun 8, 2026, 05:19 PM·Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.

    About YUNWU API

    YUNWU API is a unified LLM API gateway that offers stable and reliable API relay services. It supports a wide range of AI models, including ChatGPT, Claude, Gemini, Grok, MoonshotAI, Zhipu, and over 30 others from providers like OpenAI, Cohere, and DeepSeek. Key features include direct connectivity within China without requiring a proxy, low latency, high concurrency, and pay-as-you-go pricing. The service is designed for applications needing access to multiple LLMs through a single interface, with common integrations for tools like LobeChat and OpenCat.

    Health Check

    100%Recent availability
    History (72 pts)
    PastNow

    API Benchmarks & Pricing

    Compare 28 model rows across audit recency, latest speed tests, throughput, latency, and per-token pricing.

    View all 1997 models
    ModelInput ($/M)Output ($/M)AuditSpeedLatency
    OpenAI
    gpt-5.6-lunaCodex专属
    $0.055/M$0.329/M———
    OpenAI
    gpt-5.6-lunadefault
    $0.068/M$0.411/M———
    OpenAI
    gpt-5.6-luna限时体验
    $0.096/M$0.575/M———
    OpenAI
    gpt-5.6-luna纯AZ
    $0.103/M$0.616/M———
    OpenAI
    gpt-5.6-luna官转
    $0.205/M$1.23/M———
    OpenAI
    gpt-5.6-luna官转OpenAI
    $0.411/M$2.47/M———
    OpenAI
    gpt-5.6-luna优质官转OpenAI
    $0.548/M$3.29/M———
    OpenAI
    gpt-5.6-luna-2026-07-09纯AZ
    $0.103/M$0.616/M———
    OpenAI
    gpt-5.6-luna-2026-07-09官转
    $0.205/M$1.23/M———
    OpenAI
    gpt-5.6-terraCodex专属
    $0.137/M$0.822/M———
    Claude
    claude-opus-4-7
    --—162.4 t/s1.91 s
    DeepSeek
    deepseek-v4-flash
    --84688610073.8 t/s2.82 s
    Grok
    grok-4-1-fast-non-reasoning
    --—111.2 t/s3.38 s
    mimo-v2-flash
    --—60.5 t/s2.09 s
    OpenAI
    gpt-5.4-nano
    --—147.0 t/s4.37 s
    OpenAI
    gpt-5.4-mini
    --1008486100200.7 t/s1.44 s
    Gemini
    gemini-2.5-pro-thinking
    --—105.7 t/s9.62 s
    Grok
    grok-4-fast
    --—84.1 t/s4.01 s
    Mistral
    mistral-small-latest
    --—474.5 t/s1.51 s
    Qwen
    qwen-max-latest
    --—23.5 t/s1.55 s
    Qwen
    qwen3-coder-480b-a35b-instruct
    --—58.4 t/s1.43 s
    qwq-plus-latest
    --—43.9 t/s19.49 s
    Claude
    claude-sonnet-4-5-20250929
    --—41.4 t/s3.27 s
    Gemini
    gemini-2.5-flash-lite-nothinking
    --—318.1 t/s0.99 s
    OpenAI
    gpt-5-mini
    --—82.7 t/s6.57 s
    OpenAI
    gpt-5-nano
    --—91.0 t/s14.07 s
    OpenAI
    gpt-5-nano-2025-08-07
    --—112.0 t/s10.26 s
    OpenAI
    gpt-4o
    --—57.6 t/s2.58 s

    Showing 28 of 28 model rows

    Recent Test Records

    TimeModelSpeedLatency
    Jul 4, 04:15 PM
    Claudeclaude-opus-4-7
    162.36 tok/s
    1.91s
    May 15, 04:00 PM
    OpenAIgpt-5-mini
    82.67 tok/s
    6.57s
    Apr 24, 11:02 AM
    OpenAIgpt-5.4-nano
    19.82 tok/s
    7.05s
    Apr 24, 11:00 AM
    DeepSeekdeepseek-v4-flash
    67.17 tok/s
    2.68s
    Apr 24, 05:00 AM
    DeepSeekdeepseek-v4-flash
    80.40 tok/s
    2.96s
    Apr 1, 11:04 AM
    OpenAIgpt-5-nano-2025-08-07
    111.99 tok/s
    10.26s
    Apr 1, 11:00 AM
    OpenAIgpt-5-nano
    91.01 tok/s
    14.07s
    Mar 26, 05:04 AM
    OpenAIgpt-5.4-mini
    200.70 tok/s
    1.44s
    Mar 26, 05:02 AM
    OpenAIgpt-5.4-nano
    274.17 tok/s
    1.68s
    Mar 26, 05:00 AM
    mimo-v2-flash
    49.16 tok/s
    2.81s

    FAQ

    What is the billing model for the relay station?

    1. Offers multiple billing models: by request count, by token quantity, etc. 2. Real-time display of user's API usage and costs.

    What should I do if an API request fails?

    Common reasons for API request failures and solutions: 1. Authentication error: Check if the API key is correct 2. Insufficient balance: Please recharge your account in time 3. Parameter error: Refer to the documentation to check request parameters 4. Model unavailable: Try switching to another available model 5. Request timeout: May be due to network issues or high service load, please retry later If unresolved, contact online customer service.

    How do I view my API call records and usage?

    After logging in, you can view detailed API call records on the "Usage Log" page, including time, model, consumed token quantity, and cost information.

    How do you ensure data security?

    1. We do not store your request content and response data 2. All API requests use TLS encrypted transmission 3. Strict access control and permission management 4. Regular security audits and vulnerability scans.

    What guarantees are there for service availability?

    We promise 99.9% service availability, ensured through global distributed deployment and load balancing for stable service. For enterprise users, we provide Service Level Agreement (SLA) guarantees.

    How do I get help if I encounter problems?

    1. Consult the detailed development documentation 2. Contact online customer service support.

    Is there example code available for development?

    We provide example code and SDKs for multiple programming languages, including Python, Node.js, Java, etc., see the "Documentation" section at the top for details.

    Similar API Provider Alternatives to Compare

    Compare YUNWU API alternatives against 6 nearby API providers using 1020 LMSpeed signals across shared model coverage, pricing, benchmark speed, uptime, and free-model availability.

    ProviderWhy compareModelsFreeAvg priceSpeed30d uptime
    YUNWU API

    yunwu-ai

    A unified API gateway providing access to multiple large language models with direct connectivity in China.

    Current provider baseline2166$4.39/M115 tok/s99.4%

    newapi-higobs-com

    An OpenAI-compatible API gateway providing access to multiple large language models and AI services.

    • Lower average pricing
    • Faster measured speed
    1370$2.21/M118 tok/s0.5%

    apitoken-online

    ApiToken Online offers an OpenAI-compatible API gateway at apitoken.online with transparent per-model pricing and multi-provider routing.

    • Lower average pricing
    • More free-model options
    3316$0.736/M81 tok/s72.9%

    new-api-bxhm-onrender-com

    Kingo API Sharing Station is a LinuxDO API relay by user Kingo deployed on Render. Proxies Claude haiku / sonnet 4.5. Small daily sign-in credits; some 1.21-era models may return 502.

    • Lower average pricing
    • Higher 30-day availability
    30$1.71/MN/A99.6%

    new-waadri-top

    WAADRI runs a unified AI model gateway that exposes aggregated model access through OpenAI-, Claude-, and Gemini-compatible interfaces.

    • More free-model options
    192710N/AN/A5.5%

    api-n1n-ai

    N1N provides API access to a wide range of AI models including GPT-4, Claude 3, Gemini, and others for text, image, and video generation.

    • Higher 30-day availability
    1776$8.77/M90 tok/s99.7%

    180txt-cn

    180txt API provides an OpenAI-compatible API relay for multiple AI models.

    • More free-model options
    378$31.07/M42 tok/s4.8%
    Higobs API
    ApiToken Online
    Kingo API分享站
    WAADRI
    N1N
    180txt API