www.fluapi.com

An OpenAI-compatible API service providing access to multiple AI models.

Models
26 models
From
$5.00/M
Speed
42 tok/s
Updated
7/25/2026
Latency
0.00 s
Created At
4/22/2026
Recharge Rate
¥7.30 per $1 quota

Features

DrawingTaskData Export

API Endpoints

  • Premium High-Speed
    https://vip.fluapi.com/v1

    Recommended route 1

  • Endpoint 2Historical / Unverified
    https://vip.fluapi.com
  • Official Website
    https://www.fluapi.com/v1

    Primary endpoint

  • Global Load Balancing
    https://new.fluapi.com/v1

    Dedicated institutional line with high concurrency support

  • Premium High-Speed
    https://svip.fluapi.com/v1

    Recommended route 2

  • Endpoint 6Historical / Unverified
    https://fluapi.com
  • Endpoint 7Historical / Unverified
    https://www.fluapi.com
  • Endpoint 8Historical / Unverified
    https://img.fluapi.com
  • Endpoint 9Historical / Unverified
    https://new.fluapi.com
  • Endpoint 10Historical / Unverified
    https://image.fluapi.com
  • Endpoint 11Historical / Unverified
    https://svip.fluapi.com
Claim this provider

Verify ownership to unlock provider management features:

  • Edit provider name, content, and links
  • Get featured with priority traffic and visibility boost
  • Display a verified badge to build user trust

Leaderboard Rankings

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.

About FluAPI

FluAPI is an OpenAI-compatible API service providing access to multiple AI models. The platform offers standardized API endpoints for integrating with various language models, supporting text generation and chat completion use cases.

Health Check

100%Recent availability
History (72 pts)
PastNow

API Benchmarks & Pricing

Compare 38 model rows across audit recency, latest speed tests, throughput, latency, and per-token pricing.

Model
Input ($/M)
Output ($/M)
video-v1default
$100.00/request-
video-v1vip
$100.00/request-
$75.00/M$375.00/M
$75.00/M$375.00/M
$5.00/M$40.00/M
$5.00/M$40.00/M
$12.50/M$100.00/M
$12.50/M$100.00/M
$25.00/M$200.00/M
$25.00/M$200.00/M
$15.00/M$75.00/M
$15.00/M$75.00/M
$2.00/request-
$2.00/request-
$150.00/M$750.00/M
$150.00/M$750.00/M
$30.00/M$150.00/M
$30.00/M$150.00/M
$30.00/M$150.00/M
$30.00/M$150.00/M
gpt-5.5default
$25.00/M$200.00/M
$25.00/M$200.00/M
gpt-5.2
--
$100.00/request-
$100.00/request-

Showing 25 of 38 model rows

Recent Test Records

TimeModelSpeedLatency
Apr 22, 07:05 AM
gpt-5.3-codex
41.54 tok/s
5.55s

Similar API Provider Alternatives to Compare

Compare FluAPI alternatives against 6 nearby API providers using 191 LMSpeed signals across shared model coverage, pricing, benchmark speed, uptime, and free-model availability.

ProviderWhy compareModelsFreeAvg priceSpeed30d uptime
FluAPI

www-fluapi-com

An OpenAI-compatible API service providing access to multiple AI models.

Current provider baseline260$15.00/M42 tok/s9970%
ChooseC API

ipv4-beta-lm-studio

ChooseC API is a unified AI model aggregation gateway supporting 260+ mainstream models including Claude, GPT, Qwen, DeepSeek, Kimi, and GLM with OpenAI, Claude, and Gemini compatibility.

  • Lower average pricing
  • Faster measured speed
  • More free-model options
  • Broader model coverage
3359$0.095/M138 tok/s9950%
ChooseC API

ipv4-beta-kxcym-top-3001

ChooseC API is an OpenAI-compatible relay service hosted at ipv4-beta.kxcym.top, providing access to multiple AI models through standardized endpoints.

  • Lower average pricing
  • Faster measured speed
  • More free-model options
  • Broader model coverage
1229$0.095/M51 tok/s0%
PackyAPI

codex-api-packycode-com

PackyAPI (codex-api.packycode.com) is an OpenAI-compatible API relay for Codex and other models via a single interface.

  • Lower average pricing
  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
780$0.041/M69 tok/s9980%
F2API

api-f2api-com

F2API is an enterprise AI gateway providing a unified OpenAI-compatible interface for accessing multiple large language models.

  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
1,2620N/A272 tok/s9990%
SMLC666 API

api-smlc666-top

SMLC666 API provides an OpenAI-compatible API relay service with unified access to various AI models.

  • Lower average pricing
  • Faster measured speed
  • Broader model coverage
1830$0.240/M43 tok/s9970%
WONG公益站

wzw-pp-ua

WONG公益站 is a free, community-driven OpenAI-compatible API relay. It aggregates 70+ AI models from providers like OpenAI, Anthropic, DeepSeek, xAI, Moonshot, and Zhipu AI, offering free API quotas for developers and researchers.

  • Lower average pricing
  • Faster measured speed
  • Broader model coverage
1400$0.013/M46 tok/s9970%

Announcements

default8/20/2026

Notice: Grok4.5 and Grok4.6 are now available. Welcome to try them!

default8/19/2026

Regarding the appearance of: ⚠ Selected model is at capacity. Please try a different model. This prompt is caused by official rate limiting and insufficient computing power. Switching models can temporarily solve this problem. We are also currently researching how to solve it completely~

default8/14/2026

Notice: seedance-2.0 model is temporarily unavailable for calls.

default8/7/2026

The server and website have been fully updated. If you find any bugs or issues, please let us know via the feedback feature, and we will respond as soon as we see it. Enjoy your use!

default6/16/2026

🚀 OpenAI Full Model Series Backend Upgrade Completed All OpenAI series models on the site have switched to the official strongest stable API backend, with core capabilities significantly enhanced:

Ultra-fast response: Token throughput is 2-3 times that of regular APIs, with low latency for the first token Enhanced reasoning: Deeper code reasoning, more accurate output in complex scenarios Stable high concurrency: Smoother service operation, significantly improved concurrency capacity

default5/14/2026

📢 Service Upgrade Announcement Relay service: Smart routing is now live, resolving latency and lag issues. Concurrency increased from 50 to 500. Small and medium API sites are welcome to connect as upstream. Image generation service: The website has been fixed. Overall cost reduced from 0.1 CNY/image to 0.025 CNY/image, supporting high-concurrency generation.

default4/25/2026

Major new release! The brand new gpt-image-2 AI image generation model is officially launched 🎉

For usage instructions, refer to the documentation in the top navigation bar under Image Generation API Call. Pricing is less than 0.1 yuan, welcome to try it out.

default4/24/2026

Service Restoration Notice This afternoon, the platform was affected by a large-scale DDoS attack, resulting in server crashes and service interruptions. The issue has been thoroughly investigated and fixed, and all platform services are now back to normal. We have comprehensively strengthened server security defenses and improved system stability. We deeply apologize for any inconvenience caused and thank everyone for your patience and support!

default4/23/2026

The website is currently adding global CDN acceleration, which may cause connection interruptions. Adjustments are in progress, please wait a moment.

default4/20/2026

Website Upgrade Maintenance Successfully Completed This time, we completed a complete underlying architecture reconstruction, upgraded to high-performance computer cluster deployment, and successfully achieved millisecond-level ultra-low latency, significantly improving overall system stability and smoothness. Thank you for your patience and strong support, welcome to continue using it with confidence!

default4/14/2026

[Maintenance Completion Notice] The API relay service on this site has completed maintenance and upgrade, and all interfaces are now back to normal use. Thank you for your patience and support, and wish you a pleasant experience!

default4/1/2026

The website maintenance is complete, and it can now be used normally!

default4/1/2026

Server upgrade in progress for better experience and stability. Please wait, we'll notify you once done.

FAQ

How to configure Codex?

Check the tutorial in the navigation bar. It only takes two quick steps.

How to redeem a redemption code?

Go to Console > Wallet, scroll to the bottom. Below "Confirm Top-up", there is a redemption field. Enter your code there.

What should I do if I can't connect?

The correct URL is https://new.fluapi.com/v1. Different software may require or not require the '/v1' suffix. Try both options. If your client (CC) adds an extra '/v1', remove it to avoid duplication.

How to get an API key?

In Console > API Keys > Create API Key > Copy the key. (Currently, the default group and VIP group have no difference.)

Notes

  • Health checks: Scope: the 72-hour chart and recent availability measure API connectivity only. Each bar summarizes one hour of checks. Targets: LMSpeed tries the configured health check URL and provider status URL first, then API endpoints derived from known API hosts and recent speed-test base URLs. A website host is considered only when it looks like an API endpoint. Probe steps: each candidate goes through DNS lookup, TCP connection, TLS handshake for HTTPS, and an HTTP HEAD request with redirects followed. Probing stops after the first reachable candidate. Reachable criteria: every required network step must succeed. An HTTP response below 500 is treated as reachable, including 401 because it confirms that an authenticated API endpoint responded, except for statuses classified as blocked. Blocked results: HTTP 403, 429, 521, 525, and 530, plus detected WAF or Cloudflare challenges, are shown as blocked and excluded from availability calculations because LMSpeed cannot determine whether the API itself is down. Model availability: when a dedicated test key is configured, LMSpeed sends an authenticated GET request to a derived /models endpoint and compares returned model IDs with this provider's listed models. These per-model results appear in Models & Pricing and are not included in the provider connectivity percentage. Timeouts: TCP connection, TLS handshake, HTTP connectivity, and model requests each use a 20-second timeout. A full run can take longer when several candidates are tried. Frequency: a background worker checks all providers every 5 minutes by default. The 72-hour chart combines those samples into hourly bars, and the schedule may be changed by the service operator. Limit: automated samples are not an SLA and do not guarantee account quota, every model, every region, or successful completion requests. Check the provider's own status page before making operational decisions.
  • Domain Rating data is sourced from Ahrefs. It is a 0–100 backlink-based domain strength signal and does not measure API speed or reliability.
  • Announcements and FAQ are read from this provider's NewAPI status snapshot when available. LMSpeed stores the original content and optional English translations from the provider status source, then shows the localized fields on this page.