Profundo AI offers flat-rate access to a growing catalog of models through one OpenAI-compatible API, with Chat Completions, Responses, streaming, structured outputs, and tool calling.

Models
40 models
From
--
Speed
244 tok/s
Updated
8/13/2026
Latency
0.00 s
Created At
8/11/2026

Billing

Payment methods

StripeApple Pay

Billing types

Subscription
Refunds
Not supported
Invoicing
Not supported

API Endpoints

  • Endpoint 1
    profundoai.com
  • Endpoint 2
    api.profundoai.com
Claim this provider
HVerified by Henry Hermes

Leaderboard Rankings

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.

About Profundo AI

Profundo AI is a small-team, subscription-based AI API service launched in 2026. It offers a growing catalog of models through a single OpenAI-compatible endpoint and API key. Billing is plan-based rather than per token, and every plan includes the full catalog.

The API supports Chat Completions and the Responses API, plus streaming, structured outputs, native and parallel tool calling, and configurable reasoning effort where the selected model supports them. The documented base URL is https://api.profundoai.com/v1.

Plans use daily request allowances instead of token metering: Plus includes 250 requests per day, Priority includes 2,000, while Founder and Performance are described as unlimited subject to fair-use and service-protection controls. The catalog and model availability may change, and Profundo AI does not publish an uptime, throughput, concurrency, or availability SLA.

Health Check

99%Recent availability
History (72 pts)
PastNow

API Benchmarks & Pricing

Compare 40 model rows across audit recency, latest speed tests, throughput, latency, and per-token pricing.

Model
Model availability
Speed
Latency
claude-fable-5-1
94%
--
deepseek-v4-pro
94%
--
glm-5.3
94%
--
glm-5.3-flash
94%
--
gpt-6-astra
94%
--
qwen3.8-max
94%
--
claude-sonnet-5
94%
--
deepseek-v4-flash
94%
--
glm-5.2
94%
--
94%
67.4 t/s8.33 s
94%
22.2 t/s24.59 s
94%
62.6 t/s3.97 s
94%
48.4 t/s9.03 s
94%
112.8 t/s20.65 s
94%
97.0 t/s11.90 s
claude-haiku-4-5
61%
--
claude-opus-4.6
61%
--
claude-opus-4.8
61%
--
codestral
61%
--
gemma-4
61%
--
gemma-4-26b-a4b-it
61%
100.4 t/s29.66 s
gemma-4-31b-it
61%
145.7 t/s56.75 s
glm-4.7
61%
--
gpt-5.3-codex-spark
61%
--
gpt-5.4
61%
--

Showing 25 of 40 model rows

Recent Test Records

TimeModelSpeedLatency
Aug 30, 03:20 AM
claude-fable-5
112.83 tok/s
20.65s
Aug 30, 03:20 AM
claude-opus-4-8
96.99 tok/s
11.90s
Aug 30, 03:20 AM
claude-opus-5
67.36 tok/s
8.33s
Aug 30, 03:20 AM
gemini-3.1-pro
1277.15 tok/s
39.86s
Aug 30, 03:20 AM
gemini-3.5-flash
343.23 tok/s
74.53s
Aug 30, 03:20 AM
gemma-4-26b-a4b-it
100.41 tok/s
29.66s
Aug 30, 03:20 AM
gemma-4-31b-it
145.68 tok/s
56.75s
Aug 30, 03:20 AM
glm-4.7-flash
91.17 tok/s
92.89s
Aug 30, 03:20 AM
gpt-5.6-luna
62.61 tok/s
3.97s
Aug 30, 03:20 AM
gpt-5.6-sol
48.43 tok/s
9.03s

Similar API Provider Alternatives to Compare

Compare Profundo AI alternatives against 6 nearby API providers using 277 LMSpeed signals across shared model coverage, pricing, benchmark speed, uptime, and free-model availability.

ProviderWhy compareModelsFreeAvg priceSpeed30d uptime
Profundo AI

profundo-ai

Profundo AI offers flat-rate access to a growing catalog of models through one OpenAI-compatible API, with Chat Completions, Responses, streaming, structured outputs, and tool calling.

Current provider baseline400N/A244 tok/s8590%
钠 API

naapi-cc

Na API (naapi.cc) is an OpenAI-compatible LLM API gateway with competitive pricing and stable access to 100+ models from OpenAI, Anthropic, Google, and more.

  • Faster measured speed
  • Higher 30-day availability
  • More free-model options
  • Broader model coverage
  • Same provider category
1,0862$0.082/M649 tok/s9950%
星见雅 API

api-xinjianya-top

Xinjianya API is an OpenAI-compatible API relay at api.xinjianya.top. Service availability may be limited.

  • Higher 30-day availability
  • More free-model options
  • Broader model coverage
  • Same provider category
498211N/A67 tok/s9770%
RenRen API

llm-whitedream-top

RenRen API runs a New API-powered gateway on llm.whitedream.top for aggregated access to multiple AI models.

  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
  • Same provider category
4770$0.0075/M271 tok/s9930%
6345ywz API

api-6345ywz-cn

6345ywz API is an OpenAI-compatible API relay providing access to multiple AI models with competitive pricing.

  • Faster measured speed
  • Higher 30-day availability
  • Broader model coverage
  • Same provider category
4730$0.0001/M272 tok/s9900%
MapleLeaf API

ai-071129-xyz

MapleLeaf API runs a New API-powered gateway on ai.071129.xyz for aggregated access to multiple AI models.

  • Higher 30-day availability
  • More free-model options
  • Broader model coverage
  • Same provider category
470292N/A38 tok/s9950%
91VIP

91vip-futureppo-top

91VIP is a non-profit API service providing access to various AI models including Codex, Claude Code, and Open Code, with specific unlimited-use groups.

  • Faster measured speed
  • More free-model options
  • Broader model coverage
  • Same provider category
40210$0.149/M301 tok/s0%

Notes

  • Health checks: Scope: the 72-hour chart and recent availability measure API connectivity only. Each bar summarizes one hour of checks. Targets: LMSpeed tries the configured health check URL and provider status URL first, then API endpoints derived from known API hosts and recent speed-test base URLs. A website host is considered only when it looks like an API endpoint. Probe steps: each candidate goes through DNS lookup, TCP connection, TLS handshake for HTTPS, and an HTTP HEAD request with redirects followed. Probing stops after the first reachable candidate. Reachable criteria: every required network step must succeed. An HTTP response below 500 is treated as reachable, including 401 because it confirms that an authenticated API endpoint responded, except for statuses classified as blocked. Blocked results: HTTP 403, 429, 521, 525, and 530, plus detected WAF or Cloudflare challenges, are shown as blocked and excluded from availability calculations because LMSpeed cannot determine whether the API itself is down. Model availability: when a dedicated test key is configured, LMSpeed sends an authenticated GET request to a derived /models endpoint and compares returned model IDs with this provider's listed models. These per-model results appear in Models & Pricing and are not included in the provider connectivity percentage. Timeouts: TCP connection, TLS handshake, HTTP connectivity, and model requests each use a 20-second timeout. A full run can take longer when several candidates are tried. Frequency: a background worker checks all providers every 5 minutes by default. The 72-hour chart combines those samples into hourly bars, and the schedule may be changed by the service operator. Limit: automated samples are not an SLA and do not guarantee account quota, every model, every region, or successful completion requests. Check the provider's own status page before making operational decisions.
  • Domain Rating data is sourced from Ahrefs. It is a 0–100 backlink-based domain strength signal and does not measure API speed or reliability.
  • Announcements and FAQ are read from this provider's NewAPI status snapshot when available. LMSpeed stores the original content and optional English translations from the provider status source, then shows the localized fields on this page.