An AI model API aggregation platform providing access to multiple providers including OpenAI, Claude, Gemini, and others for production environments.

Models
611 models
From
$0.0022/M
Speed
--
Updated
9/10/2026
Latency
--
Created At
3/23/2026
Recharge Rate
¥1.00 per $1 quota

Features

DrawingTaskData Export

Login Methods

LinuxDO

API Endpoints

  • Main Site
    https://poloai.top

    US high-defense load balancing site cluster

  • Backup
    https://cf.poloai.top

    Cloudflare acceleration, text models only

  • Endpoint 3Historical / Unverified
    https://fast.poloai.top
  • Endpoint 4Historical / Unverified
    https://direct.poloai.top
  • Endpoint 5Historical / Unverified
    https://polocdn.580ai.net
  • Endpoint 6Historical / Unverified
    https://api.newapi.life
  • Endpoint 7Historical / Unverified
    https://ai.newapi.life
  • Endpoint 8Historical / Unverified
    https://apia.jianying580.com
Claim this provider

Verify ownership to unlock provider management features:

  • Edit provider name, content, and links
  • Get featured with priority traffic and visibility boost
  • Display a verified badge to build user trust

Leaderboard Rankings

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.

About PoloAPI

PoloAPI is an AI model API aggregation platform designed for production environments. It offers access to multiple AI providers through a unified interface.

Key capabilities:

  • Aggregates APIs from providers like OpenAI, Claude, Gemini, Grok, DeepSeek, Qwen, Doubao, Mistral, Llama, Cohere, Perplexity, Midjourney, Kimi, Stability AI, Hugging Face, Yi, Baichuan, GLM, Spark, and Wenxin.
  • Provides both direct official connections and relay/transit API options.
  • Supports high-concurrency usage.

Notable aspects:

  • Pricing is grouped into tiers (S-level premium resources, A-level standard resources, limited-time special offers) with rates listed in Chinese yuan per token or per call.
  • Targets enterprise clients and AI出海 (AI going global) services.
  • Offers additional services like model fine-tuning, quantitative investment, and AI application development.

Typical use cases: Production AI applications, enterprise AI integration, developers needing multi-provider access.

Health Check

98%Recent availability
History (72 pts)
PastNow

API Benchmarks & Pricing

Compare 2400 model rows across audit recency, latest speed tests, throughput, latency, and per-token pricing.

Model
Input ($/M)
Output ($/M)
claude-haiku-3-5Claude-code优质
$0.074/M$0.074/M
claude-haiku-3-5中转低价|一key调用全模型
$0.082/M$0.082/M
claude-opus-5-thinkingClaude-code稳定
$0.685/M$3.42/M
claude-opus-5-thinkingClaude-code稳定-A
$0.685/M$3.42/M
claude-opus-5-thinkingClaude-code稳定-S
$0.822/M$4.11/M
claude-opus-5-thinkingClaude-code优质
$1.23/M$6.16/M
claude-opus-5-thinking中转低价|一key调用全模型
$1.37/M$6.85/M
claude-sonnet-5-thinkingClaude-code稳定-A
$0.274/M$1.37/M
claude-sonnet-5-thinkingClaude-code稳定
$0.274/M$1.37/M
claude-sonnet-5-thinkingClaude-code稳定-S
$0.329/M$1.64/M
claude-sonnet-5-thinking中转低价|一key调用全模型
$0.548/M$2.74/M
codex-auto-reviewCodex-gpt
$0.548/M$3.29/M
codex-auto-reviewClaude-code稳定
$0.685/M$4.11/M
codex-auto-reviewCodex-gpt-蒸馏
$0.822/M$4.93/M
codex-auto-review中转低价|一key调用全模型
$1.37/M$8.22/M
dall-e-2gpt-vip
$0.024/request-
dall-e-2gpt-openai
$0.034/request-
dall-e-2A-Claude/GPT-svip 极速通道
$0.047/request-
dall-e-2企业极速|一key调用全模型
$0.038/request-
deepseek-flashSvip国内模型
$12.33/M$12.33/M
deepseek-flash中转低价|一key调用全模型
$20.55/M$20.55/M
doubao-seed-2-0-code-preview-260215豆包
$1.32/M$6.58/M
doubao-seed-2-0-lite-260215豆包
$0.247/M$1.48/M
doubao-seed-2-0-mini-260215豆包
$0.110/M$1.10/M
doubao-seed-2-0-pro-260215豆包
$1.32/M$6.58/M

Showing 25 of 2400 model rows

Recent Test Records

No test records available

Similar API Provider Alternatives to Compare

Compare PoloAPI alternatives against 6 nearby API providers using 1332 LMSpeed signals across shared model coverage, pricing, benchmark speed, uptime, and free-model availability.

ProviderWhy compareModelsFreeAvg priceSpeed30d uptime
PoloAPI

poloai-top

An AI model API aggregation platform providing access to multiple providers including OpenAI, Claude, Gemini, and others for production environments.

Current provider baseline6110$0.0022/MN/A9940%
鲨鱼魔法

openai-sharkmagic-top

鲨鱼魔法 (sharkmagic.top) provides access to official paid AI models like Gemini, Claude, and GPT through a unified proxy with OpenAI-compatible API.

  • Higher 30-day availability
  • More free-model options
  • Broader model coverage
1,57715$7.20/M81 tok/s9980%
OpenRouter

openrouter

A unified API interface providing access to over 300 models from 60+ providers, including OpenAI, Anthropic, and Google.

  • Higher 30-day availability
  • Broader model coverage
6140N/A93 tok/s10000%
HotaruAPI

api-hotaruapi-top

HotaruAPI provides API access to AI models for developers, including a model marketplace and self-service options.

  • More free-model options
3784N/A59 tok/s0%
NVIDIA NIM

nvidia-nim

NVIDIA NIM provides optimized AI model inference APIs for LLMs, vision, and embedding models through NVIDIA cloud infrastructure.

  • Higher 30-day availability
1680N/A82 tok/s9960%
WONG公益站

wzw-pp-ua

WONG公益站 is a free, community-driven OpenAI-compatible API relay. It aggregates 70+ AI models from providers like OpenAI, Anthropic, DeepSeek, xAI, Moonshot, and Zhipu AI, offering free API quotas for developers and researchers.

  • Higher 30-day availability
1550$0.0055/M46 tok/s9990%
SkyAI

api-071572-xyz

SkyAI provides an OpenAI-compatible API relay with access to multiple AI models and stable performance.

  • More free-model options
6414N/A98 tok/s0%

Announcements

default7/31/2026

The official price for GPT-5.6 has been reduced. This site has been updated to match the latest official pricing.

default7/15/2026

Due to resource issues, the veo model is now under maintenance. Please switch to other video models.

default7/8/2026

Due to official restrictions from Banana, resources are congested. If you encounter frequent errors, please try again later.

default7/3/2026

Payment system maintenance has been completed, and payment functions are now back to normal. Recharging, purchasing, renewing, and other related operations are all available. If you encounter any issues during use, please contact customer service promptly. Thank you for your understanding and support!

default7/3/2026

Dear users: Due to ongoing maintenance and upgrades to the payment system, the platform will temporarily disable payment functions. During maintenance, payment-related operations such as recharging, purchasing, and renewing will be unavailable. Existing services will not be affected. Please follow subsequent announcements for the restoration time. We apologize for any inconvenience caused. Thank you for your understanding and support.

default6/12/2026

Notice: Starting from June 15 at 24:00, this site will adjust its access policy. Access from mainland China will be closed, and only overseas access will be supported. Thank you for your understanding and support.

default6/4/2026

Due to official reasons, veo is currently unavailable. We are working urgently on repairs and maintenance.

default5/11/2026

Statement on closing individual new user registration (2026/5/11) Registration restrictions: Temporarily closing individual new user registration is to control site traffic and ensure stability for enterprise customers. Rights commitment: The platform solemnly promises that even if there are future business adjustments or shutdown plans, all individual user account balances will be fully refunded, and user interests will never be harmed. Current status: All existing user functions are operating normally, and data assets are secure and under control. All information is subject to platform notifications.

default1/26/2026

Qwen all models, GLM up to 4.7, have been connected to official resources. All series recruiting agents, with volume, consult for price adjustment.

default1/26/2026

Daily consumption exceeding 5000 RMB, contact WeChat customer service for price adjustment.

FAQ

What is the billing model for the relay station?

1. Provides multiple billing models: by request count, by token quantity, etc. 2. Real-time display of user's API usage and costs.

What to do if an API request fails?

Common reasons and solutions for API request failures: 1. Authentication error: Check if the API key is correct. 2. Insufficient balance: Please top up your account promptly. 3. Parameter error: Refer to the documentation to check request parameters. 4. Model unavailable: Try switching to another available model. 5. Request timeout: May be due to network issues or high service pressure, please try again later. If unresolved, contact online customer service.

How to view API call records and usage?

After logging in, you can view detailed API call records on the "Usage Logs" page, including time, model, consumed token quantity, and cost information.

How is data security ensured?

1. We do not store your request content and response data. 2. All API requests use TLS encrypted transmission. 3. Strict access control and permission management. 4. Regular security audits and vulnerability scans.

What is the service availability guarantee?

We promise 99.9% service availability, ensured through global distributed deployment and load balancing. For enterprise users, we provide Service Level Agreement (SLA) guarantees.

How to get help when encountering issues?

1. Consult the detailed development documentation. 2. Contact online customer service support.

Notes

  • Health checks: Scope: the 72-hour chart and recent availability measure API connectivity only. Each bar summarizes one hour of checks. Targets: LMSpeed tries the configured health check URL and provider status URL first, then API endpoints derived from known API hosts and recent speed-test base URLs. A website host is considered only when it looks like an API endpoint. Probe steps: each candidate goes through DNS lookup, TCP connection, TLS handshake for HTTPS, and an HTTP HEAD request with redirects followed. Probing stops after the first reachable candidate. Reachable criteria: every required network step must succeed. An HTTP response below 500 is treated as reachable, including 401 because it confirms that an authenticated API endpoint responded, except for statuses classified as blocked. Blocked results: HTTP 403, 429, 521, 525, and 530, plus detected WAF or Cloudflare challenges, are shown as blocked and excluded from availability calculations because LMSpeed cannot determine whether the API itself is down. Model availability: when a dedicated test key is configured, LMSpeed sends an authenticated GET request to a derived /models endpoint and compares returned model IDs with this provider's listed models. These per-model results appear in Models & Pricing and are not included in the provider connectivity percentage. Timeouts: TCP connection, TLS handshake, HTTP connectivity, and model requests each use a 20-second timeout. A full run can take longer when several candidates are tried. Frequency: a background worker checks all providers every 5 minutes by default. The 72-hour chart combines those samples into hourly bars, and the schedule may be changed by the service operator. Limit: automated samples are not an SLA and do not guarantee account quota, every model, every region, or successful completion requests. Check the provider's own status page before making operational decisions.
  • Domain Rating data is sourced from Ahrefs. It is a 0–100 backlink-based domain strength signal and does not measure API speed or reliability.
  • Announcements and FAQ are read from this provider's NewAPI status snapshot when available. LMSpeed stores the original content and optional English translations from the provider status source, then shows the localized fields on this page.