$ uprouter — directory
Connect-ready AI API routers
The 211 routers below expose OpenAI- or Claude-compatible endpoints, so they can be wired into Uprouter Connect and called through one API key — the largest free tiers belong to Nscale, BluesMinds, Mistral AI.
Compatibility is data-derived from each provider's documented API surface, not a marketing claim. Connect-ready entries are the ones you can route behind failover aliases and a single compute balance; the rest of the directory still has value for direct integrations.
Ordering here reflects free-credit value, not sponsorship — no listing position can be bought. Always review each provider's terms and data-handling policy before you point traffic at it.
- Nscaleinference.api.nscale.comOfficial free tier
Nscale is a European AI hyperscaler providing Serverless Inference for large-scale AI deployment. It offers an OpenAI-compatible API to leading open models like Llama and DeepSeek. New users receive a $5 signup credit to start prototyping without initial cost.
models: Llama family, DeepSeek family, gpt-oss-120b
up · 279ms$1,089.78estconnectLow risk - BluesMindsapi.bluesminds.comFree relay
BluesMinds is a primarily-paid AI relay reselling access to Claude and GPT at significant discounts. It offers a large free credit bonus ($500) for accounts linked to aged GitHub profiles, and provides free daily calls for specialized models like GPT-4.1.
up · 385ms$500.00estconnectMedium risk - AgentRouteragentrouter.orgFree relay
Agent Router is an open-source, public-welfare AI API aggregation platform designed to unify access to agents and MCP tools. It is 'forever free' for both users and developers, providing a discovery registry for skills (ClawHub) and a live A2A directory.
models: Claude (full Claude Code family), DeepSeek, Zhipu GLM +1
up · 997ms$200.00connectMedium risk - Mistral AIconsole.mistral.aiOfficial free tier
Mistral AI's official platform (La Plateforme) provides access to their open and proprietary models via an OpenAI-compatible API. It features a free 'Experiment' tier for developers and pay-as-you-go pricing for production workloads.
models: Mistral Small 4, Mistral Medium 3.5, Mixtral 8x22B Instruct +3
down$200.00estconnectLow risk - SenseNovatoken.sensenova.cnOfficial free tier
token.sensenova.cn is the official Token Plan for SenseTime's SenseNova multimodal AI platform. During its public beta, it offers generous free quotas for models like SenseNova 6.8 Flash Lite and U1 Fast, suitable for complex office workflows.
models: SenseNova 6.7 Flash-Lite, SenseNova U1 Fast, DeepSeek V4 Flash +1
up · 718ms$200.00estconnectLow risk - LongCatlongcat.chatFree product
LongCat is Meituan's large model API platform (longcat.chat/platform), offering OpenAI/Anthropic-compatible endpoints. During its public beta phase in 2026, it provides generous daily free token allowances for models including LongCat-Flash and LongCat-2.0, with a context window of 131K tokens.
models: LongCat-Flash-Chat, LongCat-Flash-Thinking, LongCat-Flash-Lite
up · 1223ms$150.00estconnectLow risk - xAI Consoleconsole.x.aiOfficial free tier
The xAI Console is the official developer interface for Elon Musk's Grok models. It offers a usage-based API with prepaid credits, prioritizing performance and direct access to their flagship large language models.
models: grok-beta, grok-2, grok-3 +1
up · 19ms$150.00estconnectLow risk - Gemaiapi.gemai.ccFree relay
Hajimi API (api.gemai.cc) is a Chinese third-party relay aggregation platform. It specializes in routing requests to Claude, Gemini, and GPT models through a unified OpenAI-compatible endpoint, primarily catering to users requiring simplified access to international models.
models: Gemini series, GPT series, claude-sonnet-4-6 +3
up · 204ms$74.39estconnectHigh risk - ModelScopemodelscope.cnOfficial free tier
ModelScope (modelscope.cn), backed by Alibaba Cloud, is an open-source model community and inference platform. It provides a free API tier allowing users to make up to 2,000 OpenAI-compatible calls per day across a wide range of open-source models (Qwen, DeepSeek, GLM).
models: Qwen family (most stable), DeepSeek-V3.1, Kimi +2
up · 1972ms$60.00estconnectLow risk - AnyRouteranyrouter.topFree relay
AnyRouter (anyrouter.top) is a public-welfare AI API relay primarily serving the mainland Chinese developer community. It specializes in routing Claude Code and other frontier models for free, supported by community sign-in bonuses and referral credits.
models: Claude 4 Opus, Claude 4 Sonnet
up · 1320ms$50.00estconnectMedium risk - Groqconsole.groq.comOfficial free tier
Groq Cloud provides ultra-fast LLM inference using its proprietary LPU (Language Processing Unit) hardware. It offers an OpenAI-compatible API with a substantial free tier for developers to test and build applications at scale.
models: Qwen family, Whisper Large v3, Llama-3.1-8B-Instruct +1
up · 200ms$50.00estconnectLow risk - AeroLinkaerolink.latCommercial aggregator
AeroLink (aerolink.lat) is a commercial AI API gateway and relay service that routes multiple coding assistants and automation tools through a single key. It offers tiered subscription plans with rolling usage windows and claims access to high-end models like Claude Opus 4.8.
models: Claude Code-related models, claude-sonnet-4-6, claude-opus-4-8
up · 602ms$35.00connectHigh risk - Basetenapp.baseten.coOfficial free tier
Baseten is a robust model inference and deployment platform that enables enterprises to run custom and open-weights models in production. It offers high scalability, dedicated GPU infrastructure, and a developer-friendly API for deploying models from a comprehensive library.
models: Any supported model in their library (billed by compute; no fixed free-model list)
up$30.00connectLow risk - Cerebrascerebras.aiOfficial free tier
Cerebras Inference is powered by wafer-scale WSE chips, offering world-leading inference speeds. It provides an OpenAI-compatible API with $5 in free credits and generous rate limits for developers. Details were checked against the provider's own page at www.cerebras.ai.
models: gpt-oss-120b, Qwen3 32B, Llama 4 Scout +1
up$30.00estconnectLow risk - Cerebras APIapi.cerebras.aiOfficial free tier
Cerebras Inference offers ultra-fast AI inference powered by its Wafer-Scale Engine (WSE-3) technology, achieving record-breaking speeds up to 3000 tokens/s. The official API is OpenAI-compatible and features a generous free tier alongside professional pay-as-you-go access.
models: Llama 3.1/3.3 family, Qwen 3 family (per official model list), Llama 4 Scout +1
up · 43ms$30.00estconnectLow risk - Z.aiapi.z.aiOfficial free tier
Z.ai is Zhipu AI's international developer platform, offering access to the GLM (General Language Model) family. It provides a high-performance alternative to Western models, with the GLM-Flash series offered permanently for free to developers to encourage global adoption of its flagship reasoning models.
models: GLM-4.5-Flash, GLM-4.6V-Flash, GLM-4.7-Flash
up · 1650ms$30.00estconnectLow risk - Tokenesstokeness.ioCommercial aggregator
Tokeness (tokeness.io) is an AI model aggregator providing access to 50+ frontier models via a standard, unified API protocol. It focuses on stability and security, relaying requests without retaining conversation data. It includes a referral program and free routing options.
models: (mostly paid) Claude Opus 4.6/4.7, (mostly paid) OpenAI family
up · 3906ms$29.76estconnectMedium risk - Inference.netinference.netOfficial free tier
Inference.net (formerly Kuzco) is a distributed GPU inference network offering a unified OpenAI-compatible API for open-source models. It specializes in low-cost deployment of Llama 3.1 models and offers substantial initial credits ($1 + $25 for surveys) for new developers.
models: Meta-Llama-3.1-70B-Instruct
up$26.00estconnectMedium risk - xAI APIapi.x.aiOfficial free tier
xAI (founded by Elon Musk) offers the official API for the Grok series of LLMs. Known for leading benchmarks in coding and reasoning, the Grok-4.6 flagship model is available via an OpenAI-compatible API, featuring low-latency inference and high agentic tool-calling capabilities.
models: Grok 4.20
up · 116ms$25.00estconnectMedium risk - Feifeimiaoapi.feifeimiao.topFree relay
Feifeimiao (api.feifeimiao.top) is a Chinese third-party relay service offering an OpenAI-compatible API gateway. Built on the New-API framework, it aggregates various models and focuses on providing access to tools like Codex and Claude Code for developers in the region.
models: gpt-image-2, claude-opus-4-7, gpt-5.5
up · 284ms$22.32estconnectMedium risk - Cun.aicun.aiCommercial aggregator
Cun.ai is a commercial AI relay aggregator based on the New-API framework, providing access to GPT, Claude, and DeepSeek. It targets Chinese developers with localized payment methods and a substantial signup credit promotion.
up · 84ms$20.83estconnectHigh risk - Alibaba Cloud Bailianbailian.console.alibabacloud.comOfficial free tier
Alibaba Cloud Model Studio (Bailian) is the official international platform for Qwen and other foundation models. It offers 1 million free tokens per model for new users and a 50% discount on batch inference.
up$20.00estconnectLow risk - Camel-Hubapi.camel-hub.comCommercial aggregator
Camel-Hub (api.camel-hub.com) is a commercial AI API relay aggregating 330+ models. It emphasizes prompt caching to reduce user costs and requires real-name verification (KYC) for full access. Free access is limited to tool-class models like Suno and Midjourney.
models: Suno (tool), Midjourney (tool), embeddings (tool)
up · 415ms$20.00connectMedium risk - DeepInfradeepinfra.comPaid API
DeepInfra is a high-speed inference cloud for open-source AI models. It provides an OpenAI-compatible API for text generation, embeddings, and image generation. Authentication uses a standard API key from its dashboard.
models: Meta: Llama 3.1 405B Instruct, Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct +3
unknown$20.00estconnectLow risk - Google AI Studioai.google.devOfficial free tier
Google AI Studio (ai.google.dev) is the official developer platform for Google's Gemini model family. It offers both a generous free tier for prototyping and a paid production tier with higher rate limits, context caching, and data privacy guarantees.
models: Gemma 4 26B A4B , Gemma 4 31B, Gemini 2.5 Flash Lite +3
up · 215ms$15.00estconnectLow risk - Lightning AIlightning.aiOfficial free tier
Lightning AI (LitAI) provides a unified gateway to frontier models through an OpenAI-compatible API. New users receive 40 million free tokens (approx. $15 in credits) to test models like Claude, GPT, and Gemini. Usage beyond the initial credit is billed as pay-as-you-go.
models: gpt-5 / gpt-5.5 family, gemini-3.5-flash / gemini-2.5-pro, nemotron-3-ultra-550b +3
up · 87ms$15.00estconnectLow risk - NLP Cloudnlpcloud.comOfficial free tier
NLP Cloud is a commercial provider hosting various open-model APIs (Llama, Mixtral, Dolphin) with both pay-as-you-go and subscription options. It grants new users $15 in free credit upon phone verification.
models: Llama 3.1 405B, gpt-oss-120b, Llama 3.3 70B Instruct
up$15.00estconnectLow risk - StepFunplatform.stepfun.comOfficial free tier
StepFun Open Platform is the official developer platform of a Chinese AI unicorn, serving its Step-series multimodal models via an OpenAI-compatible API. It supports text, vision, and end-to-end speech models with tiered rate limits and a credits-based free allowance for new accounts.
models: Step-3, Step-series text LLMs, Step-series multimodal/vision models +1
up · 1157ms$15.00estconnectLow risk - AI21 Labsapi.ai21.comOfficial free tier
AI21 Labs is an established LLM vendor known for its Jamba model family. The developer platform (api.ai21.com) offers OpenAI-style endpoints for its high-performance models, with a trial credit for new accounts to explore its capabilities.
models: Jamba Large (1.7), Jamba Mini (1.7), Jamba 1.6 Large
up · 89ms$10.00connectLow risk - Eden AIedenai.coCommercial aggregator
Eden AI is a commercial AI aggregation gateway providing a single API to 500+ models. It features cost tracking, automatic fallback, and a unified format for LLMs, vision, and translation. Details were checked against the provider's own page at www.edenai.co.
up · 188ms$10.00connectLow risk - FreeModelfreemodel.devCommercial aggregator
FreeModel is a multi-model gateway that intelligently routes requests to the best-performing open models. It supports both OpenAI and Anthropic API formats and provides automatic model fallback and smart provider selection to optimize for quality.
models: (promo also claims Claude Sonnet 4.6 / Opus 4.7), gpt-5.5, gpt-5.4 +2
up · 64ms$10.00estconnectMedium risk - Upstageconsole.upstage.aiOfficial free tier
Upstage Console provides an OpenAI-compatible API for their high-performance Solar models. It features a developer-friendly $10 signup credit and commitment-based tiers for businesses requiring higher rate limits and support.
models: Solar Pro, Solar Mini, Solar Pro 3
up$10.00estconnectLow risk - FreeTheAIfreetheai.xyzFree relay
FreeTheAi is a free OpenAI-compatible API gateway aggregating 80+ models behind a single base URL. It operates via Discord-based key signup and daily check-ins, emphasizing persistent free access for developers and roleplay users without credit card requirements.
up · 106ms$7.50estconnectHigh risk - Inception Labsapi.inceptionlabs.aiOfficial free tier
Inception Labs offers a cloud inference platform for its in-house diffusion-based LLMs, such as Mercury 2 and Mercury Coder. Designed for extreme efficiency, it provides an OpenAI-compatible API that significantly undercuts GPU-based inference costs while maintaining high reasoning quality.
models: mercury (general chat diffusion model), mercury-coder (code, with fim/completions), mercury-2
up · 98ms$7.00estconnectLow risk - Ollamaapi.ollama.comOfficial free tier
Ollama Cloud is the official managed inference platform for the Ollama ecosystem. It allows developers to deploy and call popular open-weights models (Llama, DeepSeek, Qwen) via a unified OpenAI-compatible API, providing a bridge between local development and cloud scale.
models: deepseek-v3.1:671b-cloud, gpt-oss-120b, gpt-oss:20b +1
up · 70ms$7.00estconnectLow risk - Requestyrequesty.aiCommercial aggregator
Requesty is an AI model gateway and router that aggregates over 600 models (Claude, GPT, Gemini, DeepSeek) through a single OpenAI-compatible API. It offers cost optimization, unified billing, and advanced routing features for enterprise and individual developers.
up · 198ms$6.00estconnectLow risk - EasyMaxeasymax.aiCommercial aggregator
EasyMax (Easy API) is a multi-model aggregation gateway compatible with OpenAI and Claude API protocols. It focuses on stability and pricing, supporting GPT, Gemini, and image generation models. Details were checked against the provider's own page at easymax.ai.
models: gpt-image-2, gemini-3-pro-image-preview, gemini-3.1-flash-image +3
up · 95ms$5.00estconnectMedium risk - Flatkeyconsole.flatkey.aiCommercial aggregator
Flatkey.ai is a commercial AI API gateway aggregating over 200 official models under a single API key. It uses a prepaid pay-as-you-go model, claiming to be significantly cheaper than official pricing by leveraging high-volume enterprise agreements.
models: veo-3.1-generate-preview, grok-imagine-image-pro, MiniMax-H3 +3
up · 378ms$5.00connectMedium risk - FreeModel APIapi.freemodel.devCommercial aggregator
FreeModel is a frontier model API router that provides a single, unified endpoint for top-tier models like GPT-5 and Claude Code. It features intelligent routing to optimize for output quality and cost, supporting OpenAI-compatible SDKs for seamless developer integration.
up · 146ms$5.00connectMedium risk - IAMHCapi.iamhc.cnFree relay
IAMHC (Xinjiang Huancheng Wangan) is a public-welfare AI API gateway providing unified access to over 199 mainstream models. It features a rate-limited free tier and emphasizes compliance with built-in moderation for developers in the China region.
models: Free tier = small-parameter test models (specific names not individually disclosed), Paid tiers cover 199+ mainstream models including GPT/Claude/Gemini
up · 2333ms$5.00estconnectMedium risk - InternAIchat.intern-ai.org.cnOfficial free tier
InternLM (Shanghai AI Laboratory) provides an open platform with an OpenAI-compatible API. It offers a monthly free token allowance for its InternVL and long-thinking reasoning models. Details were checked against the provider's own page at internlm.intern-ai.org.cn.
models: intern-latest, intern-s1, intern-s1-mini +2
up · 1077ms$5.00estconnectLow risk - Pollinationspollinations.aiFree product
Pollinations.ai is a Berlin-based open-source platform providing a signup-free, OpenAI-compatible API for text, image, and audio generation. It aggregates multiple open-weights models (DeepSeek, Qwen, Mistral) and offers them for free to foster creative AI development.
models: OpenAI GPT-5 (Mini/Nano/5.2), Google Gemini 3 Flash / 2.5 Flash Lite, Mistral Small 3.2 +3
up · 60ms$5.00estconnectLow risk - SambaNovacloud.sambanova.aiOfficial free tier
SambaNova Cloud provides fast LLM inference on specialized RDU hardware. It offers an OpenAI-compatible API with a persistent free tier (Developer Tier) and pay-as-you-go credits for higher rate limits.
models: Llama 3.1 405B, DeepSeek-V3.1/V3.2, gpt-oss-120b +2
up · 293ms$5.00connectLow risk - Vercel AI Gatewayvercel.comOfficial free tier
Vercel AI Gateway is a unified API gateway that allows routing to hundreds of models from various providers through a single OpenAI-compatible endpoint. It provides observability, budgets, and zero-data-retention (ZDR) options with no token markup over upstream list prices.
models: Access to OpenAI/Anthropic/Google/Meta/DeepSeek and more via the gateway (at upstream list prices)
up · 40ms$5.00connectLow risk - Docodedocode.ccCommercial aggregator
DoCode is a relay specifically optimized for AI coding workflows (Claude Code, Cursor, Codex). It uses heavy caching and model routing to lower costs, with VIP tiers for further discounts. Authentication via sk- key.
up · 420ms$4.00estconnectHigh risk - AnyAPIanyapi.aiCommercial aggregator
AnyAPI is a commercial multi-model API aggregation gateway providing access to 400+ models via a single key. It offers a generous free tier for developers and scales to high-volume enterprise plans, supporting models from OpenAI, Anthropic, and open-weights families.
models: QwQ 32B, Gemma 4, Qwen3 Coder +3
up · 55ms$3.00estconnectLow risk - Dahlinference.dahl.globalOfficial free tier
Dahl Inference is a high-performance platform offering decentralized access to open models like DeepSeek and MiniMax. It features an OpenAI-compatible API and a generous signup promotion of 100 million free tokens.
models: Llama-3.1-8B-Instruct, Qwen2.5 72B Instruct
unknown$3.00connectMedium risk - Difydify.aiPaid API
Dify is an open-source LLM app development platform. Its cloud service (Dify Cloud) provides hosted model access via an API, with credits consumable across OpenAI, Anthropic, Gemini, and others. Details were checked against the provider's own page at dify.ai.
unknown$3.00estconnectLow risk - Qinius.qiniu.comOfficial free tier
Qiniu Cloud AI LLM Inference is a managed MaaS platform that provides a unified, compliant API for over 50 mainstream LLMs. It is compatible with both OpenAI and Anthropic protocols, offering new users substantial free tokens and flexible resource packages for production scaling.
models: DeepSeek-V4 series, Qwen, GLM (Zhipu) +3
up · 1823ms$3.00estconnectLow risk - VSLLMvsllm.comFree relay
VSLLM (Weiyun Model Open Platform) is a commercial AI API relay gateway that aggregates multiple large models. It is OpenAI-compatible and features a daily auto-refreshing free quota for new sign-ups. Claims to offer highly competitive, transparent pricing based on actual token usage.
models: gemini-3.1-pro-request-antigravity, mimo-v2.5-tts-voicedesign, glm-5.3-anthropic +3
up · 224ms$3.00estconnectMedium risk - SiliconFlowapi.siliconflow.cnOfficial free tier
SiliconFlow (SiliconCloud) is a premier Chinese AI model platform offering high-performance inference for a wide range of open-source models (Qwen, Llama, DeepSeek). It is notable for providing free, permanent API access to models with parameters under 9B, alongside competitive pay-as-you-go rates for larger flagship
models: Qwen2.5-7B-Instruct, Llama-3.1-8B-Instruct
up · 1005ms$2.38estconnectLow risk - Kimiplatform.kimi.comOfficial free tier
The official Moonshot AI developer platform for Kimi models. Offers pay-as-you-go access to Kimi K3, K2.7-Code, and K2.6 over an OpenAI-compatible API. One-off trial credits available upon verification.
up · 745ms$2.23estconnectLow risk - Moonshot AIplatform.moonshot.cnOfficial free tier
Official Moonshot AI developer platform for Kimi models. Offers pay-as-you-go access to Kimi K3, K2.7-Code, and K2.6 over an OpenAI-compatible API. One-off trial credits available upon verification. Details were checked against the provider's own page at platform.moonshot.cn.
models: moonshot-v1-8k, moonshot-v1-32k, moonshot-v1-128k +1
up · 1947ms$2.23estconnectLow risk - ChatAnywherechatanywhere.techFree relay
ChatAnywhere is a widely used free LLM relay. Users bind a GitHub account to get a free API key and access GPT-4o, Claude, and Gemini models through an OpenAI-compatible interface, optimized for access from mainland China.
models: gemini, grok, Qwen +3
up · 349ms$2.00estconnectMedium risk - ChatAnywhere APIapi.chatanywhere.orgFree relay
ChatAnywhere (GPT_API_free) is a widely used AI API relay providing unified access to global models including OpenAI, Claude, Gemini, and DeepSeek. It offers a GitHub-authorized free tier and low-cost paid access with endpoints optimized for users in mainland China and overseas.
models: GPT-5 series (5/day), DeepSeek (30/day), several small models +2
up · 359ms$2.00estconnectMedium risk - DuiAPIduiapi.comCommercial aggregator
DuiAPI is a Chinese commercial-aggregator relay emphasizing compliant, ICP-registered access to official AI platforms. It supports Qwen, DeepSeek, and Kimi models through OpenAI and Anthropic compatible endpoints.
up · 468ms$2.00connectHigh risk - Super-NBapi.super-nb.meCommercial aggregator
Super NB (api.super-nb.me) is a Chinese AI API gateway providing OpenAI-compatible access to global LLM providers. It aggregates GPT and Claude models, offering a simplified top-up and usage experience for developers in the region through a New-API based interface.
up · 204ms$2.00estconnectHigh risk - Unity2unity2.aiCommercial aggregator
Unity2.Ai is a commercial AI infrastructure provider offering seamless access to top Chinese and global AI models (Claude, GPT, Gemini, DeepSeek) via a single OpenAI-compatible API. Features self-developed cluster architecture for high concurrency and smart upstream scheduling. Supports Alipay and WeChat via Stripe.
degraded$2.00estconnectMedium risk - GitHub Modelsmodels.github.aiOfficial free tier
GitHub Models (models.github.ai) was an official model-inference platform by GitHub for prototyping. As of July 30, 2026, the service has been fully retired and the catalog/API are no longer available. Users are directed to Azure AI for production workloads.
models: Llama family, GPT-4o, GPT-4o-mini
up · 28msno free tierconnectLow risk - OpenRouteropenrouter.aiCommercial aggregator
OpenRouter is a commercial LLM gateway aggregating 500+ models from every major provider. It offers a single OpenAI-compatible API with a platform fee (5.5%) and a selection of free models for development.
models: Auto Router, Free Models Router, qwen/qwen3-next-80b-a3b-instruct:free +3
up · 49ms$1.50estconnectLow risk - APIPlusapi.apiplus.cloudFree relay
APIPlus is a third-party AI API relay built on the New-API framework, reselling access to Claude, GPT, and Gemini. It offers a small registration credit and an intermittently available 'free channel' with deep discounts, primarily targeting the low-cost reseller market.
models: GPT family (free channel at 0.01x, not free)
down$1.00estconnectMedium risk - Agnes AIagnes-ai.comFree product
Agnes AI, developed by Singapore-based Sapiens AI, is an omni-modal AI platform offering indefinitely free access to core text, image, and video models. It supports OpenAI-compatible API keys and tiered 'Token Plans' for production-level throughput.
models: Agnes-2.0-Flash (text), Agnes-Image-2.0-Flash (image), Agnes-Video-V2.0 (video)
up · 271ms$1.00estconnectMedium risk - Bytezbytez.comOfficial free tier
Bytez is an AI model hosting platform that offers an OpenAI-compatible API for a wide range of models. It provides $1 in free credits that refresh every 4 weeks and uses a low-cost pay-as-you-go pricing model.
models: Llama-3.1-8B-Instruct, Anthropic: Claude 3.5 Sonnet, GPT-4o
unknown$1.00connectMedium risk - Electron Hubwww.electronhub.aiCommercial aggregator
Electron Hub provides a unified API to 600+ models, including premium frontier models (GPT, Claude, Gemini). It offers weekly credit refills on both free and paid subscription plans, plus permanent top-ups.
models: Meta: Llama 3.1 405B Instruct, claude-sonnet-5, Google Gemini Pro Latest +1
unknown$1.00connectLow risk - FastRouterfastrouter.aiCommercial aggregator
FastRouter is a high-performance gateway for AI model routing, comparison, and operations. It supports 200+ models via an OpenAI-compatible API with intelligent cost optimization and failover. Details were checked against the provider's own page at fastrouter.ai.
models: gemini-3.5-flash-lite, gpt-5.6-luna, DeepSeek V4 Flash +3
unknown$1.00estconnectLow risk - Fireworks AIfireworks.aiOfficial free tier
Fireworks AI is a high-speed serverless inference platform for open-weights models (Llama, Qwen, DeepSeek). It provides an OpenAI-compatible API optimized for low latency and high throughput. Details were checked against the provider's own page at docs.fireworks.ai.
models: Llama family, DeepSeek family, Qwen family
up · 171ms$1.00connectLow risk - Hyperbolichyperbolic.aiOfficial free tier
Hyperbolic (hyperbolic.ai) is an open-access AI cloud providing high-performance serverless inference and GPU rentals. It offers an OpenAI-compatible API to 25+ open-source models including DeepSeek, Qwen, and Llama 3.1 405B. New users receive roughly $1 in trial credits upon phone verification.
models: DeepSeek V3, Llama 3.1 405B (BF16, currently the only public 405B Base), Qwen family +2
up · 173ms$1.00estconnectLow risk - Hyperbolic XYZhyperbolic.xyzOfficial free tier
Hyperbolic (hyperbolic.xyz) is the API-specific endpoint for Hyperbolic's AI cloud. It provides serverless inference for over 25 open-source models through an OpenAI-compatible /v1 endpoint. The platform supports text, image, and audio generation with pay-as-you-go pricing.
models: DeepSeek V3, Llama 3.1 405B, Qwen family +1
up · 55ms$1.00estconnectLow risk - Nebiusnebius.comOfficial free tier
Nebius AI Studio is a European AI cloud offering an OpenAI-compatible API to 60+ open-source LLMs; it provides high-performance inference-as-a-service with European data residency and transparent pay-as-you-go pricing.
models: DeepSeek family, Llama/Qwen and 60+ open-source models, Qwen3-235B-A22B +1
up · 1321ms$1.00estconnectLow risk - Nebius Studioapi.studio.nebius.comOfficial free tier
Nebius AI Studio is a European-based GPU cloud platform providing managed inference for top open-source LLMs like Llama 3.1 and DeepSeek R1. It leverages its own EU data centers to offer high-performance, privacy-compliant inference with transparent pay-as-you-go pricing tailored for both developers and enterprises.
models: DeepSeek family, Qwen family, Llama-3.1-8B-Instruct
up · 387ms$1.00estconnectLow risk - OVHcloudovhcloud.comOfficial free tier
OVHcloud AI Endpoints is an inference API from the established European cloud provider OVHcloud. It offers OpenAI-compatible access to open-weight models with anonymous and registered free tiers for developers.
models: Qwen, Mistral, Llama +1
up · 86ms$1.00estconnectLow risk - iFlyTek Sparkxfyun.cnOfficial free tier
iFlytek Spark (Xunfei Xinghuo) is a large model platform providing access to the Spark LLM family. It offers a permanently free Spark Lite version and significant one-time token allowances for newer users. The API is OpenAI-compatible and supports various modalities including text, image, and voice.
models: Spark Lite (permanently free), Spark Max (limited/campaign free), Spark Pro/other versions (limited free tokens)
up · 2034ms$1.00estconnectLow risk - JuCodexjucodex.comCommercial aggregator
JuCodex is a Chinese-market AI relay specializing in coding models, reselling access to OpenAI's Codex and GPT-4 variants. It features a custom ratio system for billing and provides a small initial credit for new accounts to test compatibility with IDE extensions.
models: gpt-image-2
up · 716ms$0.89estconnectHigh risk - DeepSeekplatform.deepseek.comOfficial free tier
DeepSeek is a leading AI research lab providing frontier-class models (V4-Flash/Pro, R1) via an OpenAI-compatible API. It is known for industry-leading cost efficiency and specialized thinking/reasoning modes.
models: DeepSeek Chat, DeepSeek-R1, DeepSeek V4 Flash Vision Exp +2
unknown$0.70estconnectMedium risk - AionLabsapi.aionlabs.aiFree product
Aion Labs provides specialized fine-tuned models for immersive roleplay and storytelling. Its API is OpenAI-compatible and features a long-term daily free allowance for developers, with higher tiers available upon account top-up.
models: aion-2.5, aion-1.0, aion-1.0-mini +2
up · 190ms$0.50estconnectLow risk - DXNT / DX Tokenwww.dxnt.comCommercial aggregator
DXNT (DX Token) is an AI model API gateway providing unified access to GLM, Kimi, MiniMax, and DeepSeek. It offers high compatibility with IDEs and agents, featuring smart routing and MCP support. Details were checked against the provider's own page at dxnt.com.
models: claude-sonnet-5, GPT-4o
unknown$0.50estconnectMedium risk - Novita AInovita.aiOfficial free tier
Novita AI is a developer-focused inference cloud offering an OpenAI-compatible API to models like DeepSeek, Qwen, and Llama. New users receive a $0.50 trial credit; service is otherwise pay-as-you-go with very low latency.
models: DeepSeek family, Qwen family, Llama family
up · 341ms$0.50estconnectLow risk - CloudCodecloudcode.oneCommercial aggregator
CloudCode.ONE is a commercial AI proxy optimized for coding workflows (e.g., Claude Code), offering pay-as-you-go access to high-performance models like DeepSeek V4 Pro as a backend replacement with a small trial credit.
models: DeepSeek V4 Pro (pay-as-you-go; free limited to a tiny trial credit)
up · 357ms$0.48estconnectMedium risk - Scalewayconsole.scaleway.comOfficial free tier
Scaleway Generative APIs offer a serverless, OpenAI-compatible endpoint for deploying open LLMs. Hosted in Europe, it emphasizes data sovereignty and grants 1,000,000 free tokens to all customers to jumpstart development.
models: Mistral family, Qwen family, Pixtral 12B (2409) +3
up$0.22estconnectLow risk - Bazaarlinkbazaarlink.aiCommercial aggregator
BazaarLink is a Taiwan-based AI aggregator providing access to 100+ models via an OpenAI-compatible API. It specializes in TWD billing and unified enterprise invoices. It offers a small signup credit and a persistent free model tier.
models: deepseek/deepseek-v4-flash-free, auto:free
up · 423ms$0.20estconnectMedium risk - Jeniyajeniya.cnCommercial aggregator
Jeniya (jeniya.cn) is an anonymously operated Chinese AI API relay reselling access to Claude, GPT, and Grok models via a unified OpenAI-compatible endpoint. It uses the One-API panel for management and supports domestic payment methods like Alipay and WeChat.
models: gemini-3-pro-image-preview, gemini-3.1-flash-image-preview, gemini-3.1-flash-image +3
up · 1132ms$0.20estconnectHigh risk - Poepoe.comSubscription router
Poe by Quora is a major AI aggregator and bot-building platform. It provides an OpenAI-compatible API that allows developers to access hundreds of models (GPT, Claude, Gemini, Llama) using a unified credits-based billing system. Poe also supports custom server bots via the Poe Protocol.
up · 81ms$0.18estconnectLow risk - Levolinkai.levolink.comCommercial aggregator
Levolink is an anonymously operated Chinese AI API relay built on the New-API framework. It aggregates a vast array of models (500+) including GPT, Claude, Gemini, and Doubao, primarily targeting the RMB-based domestic market with cheap top-ups.
models: doubao-seedream-4-5-251128, doubao-seedream-5-0-260128, grok-imagine-image +3
up · 320ms$0.12connectHigh risk - Cloudflare Workers AIdevelopers.cloudflare.comOfficial free tier
Cloudflare Workers AI allows running AI models on Cloudflare's global network. It provides an OpenAI-compatible API for inference, embeddings, and image generation, billed in compute 'Neurons'. Details were checked against the provider's own page at developers.cloudflare.com.
models: Gemma family, Llama family, Qwen family +3
up$0.11connectLow risk - Hugging Facehuggingface.coOfficial free tier
Hugging Face is the leading hub for open-source AI, offering model hosting, datasets, and serverless inference. It provides a permanent free tier for basic CPU spaces and community GPU grants, with professional-grade Inference Endpoints available on a pay-as-you-go basis.
models: Llama family, Qwen family, Gemma family +2
up · 17ms$0.10connectLow risk - Nomicnomic.aiOfficial free tier
Nomic AI provides the Atlas Embedding API and open-weight models like nomic-embed. It offers a monthly free tier for its embedding API and is transitioning to focus on document AI for specialized industries.
models: nomic-embed-text-v1.5, nomic-embed-text, nomic-embed-vision
up · 305ms$0.10estconnectLow risk - 9Router9router.xunleiyk.comFree relay
9Router is a specialized AI API router designed for smart fallback and cost optimization, allowing users to pool their own subscription keys. It emphasizes 'unlimited free AI coding' by routing through free models or maximizing existing subscriptions via a local proxy or VPS.
up · 1505msfree tier · not quantifiedconnectUnrated - AI Routerai-router.devSubscription router
AI-ROUTER is an enterprise-focused AI API gateway providing unified access to OpenAI, Claude, and Gemini models. It uses a unique 'fuel package' and subscription model, offering managed quotas with daily, weekly, and monthly limits for intensive development and team workflows.
up · 173msno free tierconnectMedium risk - AI/ML APIaimlapi.comPaid API
AI/ML API is a major AI model aggregator providing unified access to 1000+ models from OpenAI, Anthropic, Google, DeepSeek, and more. It offers a single API key for LLM, image, voice, and video models with a focus on production-ready pay-as-you-go access.
models: Phi 4, Claude 3 Haiku, Gemini 2.5 Flash +3
unknownfree tier unknownconnectMedium risk - Aiaiai001api.aiaiai001.comCommercial aggregator
Aiaiai001 (Longcheng AI Toolbox) is a Chinese commercial AI API relay aggregating frontier models via a New-API based backend. It provides an OpenAI-compatible endpoint and requires quota purchases through an external shop (pay.ldxp.cn).
up · 91msno free tierconnectHigh risk - Ant Ling / Ring (inclusionAI)developer.ant-ling.comPaid API
Ant Ling (by inclusionAI) provides an OpenAI-compatible API for its proprietary Ling model family. It supports high-speed inference and is designed to integrate seamlessly with standard SDKs by switching the base URL, serving both developer and enterprise needs.
models: Ling 3.0 Flash, Ling 3.0 Flash Fin, inclusionAI: Ling 3.0 Flash Sante (free)
unknownfree tier unknownconnectMedium risk - Api.airforceapi.airforceOfficial free tier
Api.airforce is a unified AI model aggregator reselling access to 100+ models from Anthropic, OpenAI, and xAI. It offers a large selection of free models (including Grok-3 and Claude 3.7) alongside paid subscription tiers for higher throughput.
models: Gemini 2.5 Flash, DeepSeek Chat, Grok 4.20 +3
unknownfree tier · not quantifiedconnectMedium risk - Apishopapishop.orgCommercial aggregator
Apishop is a professional AI API gateway designed for developers in mainland China. It provides direct, low-latency access to Anthropic, OpenAI, and Gemini models with zero code changes required. Billing is transparently mapped 1:1 to official USD rates but paid in local currency.
up · 751msfree tier · not quantifiedconnectLow risk - AssemblyAIassemblyai.comSearch / Audio API
AssemblyAI is a leading speech-to-text API provider offering high-accuracy transcription (Universal-3.5 Pro) and a multi-model LLM Gateway. It supports OpenAI-compatible routing for its language model features. Access is pay-as-you-go after a free trial.
unknownfree tier · not quantifiedconnectLow risk - Aurikowww.auriko.aiPaid API
Auriko is an AI API router and aggregator that provides a unified interface to multiple LLM providers. It offers both platform-managed and Bring Your Own Key (BYOK) modes. It publishes high rate limits for both modes.
unknownfree tier unknownconnectMedium risk - Azure AI Foundrylearn.microsoft.comPaid API
Azure AI Foundry (formerly Azure AI Studio) provides access to a vast catalog of models including OpenAI, Anthropic, and Llama. It features a unified API surface and enterprise-grade management. Pricing is unified under Azure's pay-as-you-go model.
models: Claude 3 Haiku, DeepSeek Chat, Phi 4 +1
unknownfree tier unknownconnectLow risk - Azure OpenAIazure.microsoft.comPaid API
Azure OpenAI Service provides REST API access to OpenAI's powerful language models including the GPT-4 and GPT-5 series with Azure's enterprise capabilities. It uses a strictly pay-as-you-go or provisioned throughput model.
models: gpt-5.6-luna, GPT-4o-mini, GPT-4o +2
unknownfree tier unknownconnectLow risk - Baichuanwww.baichuan-ai.comPaid API
Baichuan is a leading Chinese AI company offering large-scale language models. Its platform provides model inference via native and OpenAI-compatible APIs, focusing on high-quality Chinese language performance.
unknownfree tier · not quantifiedconnectMedium risk - Baidu (ERNIE)ernie.baidu.comPaid API
Baidu's Qianfan platform provides access to the ERNIE Bot model family. It is a full-stack AI development platform offering model training, fine-tuning, and inference. New users typically receive a small trial credit.
models: Baidu: ERNIE Lite, Baidu: ERNIE Speed, Baidu: ERNIE 3.5 +2
unknownfree tier · not quantifiedconnectMedium risk - Baidu Qianfancloud.baidu.comPaid API
Baidu Qianfan is the official Model-as-a-Service (MaaS) platform of Baidu Intelligent Cloud. It serves the ERNIE (Wenxin Yiyan) series models and other third-party models via an OpenAI-compatible v2 API, offering high-performance inference for enterprise and individual developers.
models: Baidu: ERNIE 4.0 Turbo, ERNIE 4.5 VL 424B A47B , Baidu: ERNIE 4.0 +1
unknownfree tier · not quantifiedconnectMedium risk - Blackbox AIblackbox.aiPaid API
Blackbox AI provides a unified endpoint for 300+ models with end-to-end encryption. The service has migrated to an enterprise-focused gated model (enterprise.blackbox.ai) after deprecating its public inference surface (api.blackbox.ai).
models: Anthropic: Claude 3.5 Sonnet, Google: Gemini Pro 1.5, GPT-4o +1
unknownno free tierconnectHigh risk - Celebrasapi.celebras.aiUnclassified
Provider slug 'api-celebras-ai' appears to be a typographical error of 'api-cerebras-ai'. No distinct AI service or official documentation was found for this specific domain (celebras.ai). Research suggests 'Celebros' exists as an e-commerce search tool, but it is unrelated to LLM inference.
downfree tier · not quantifiedconnectUnrated - Charm Hyperhyper.charm.landPaid API
Hyper (by Charm) is an AI inference solution purpose-built for coding. It manages its own infrastructure to optimize open-source coding models. It provides 100 free Hypercredits monthly on signup. Details were checked against the provider's own page at hyper.charm.land.
models: DeepSeek: DeepSeek Coder V2 Instruct, Qwen2.5 72B Instruct, Meta-Llama-3.1-70B-Instruct +1
unknownfree tier · not quantifiedconnectMedium risk - Chat Oripeapi.oriper.comPaid API
Chat Oripe is an AI API aggregator offering access to multiple LLM families via a unified interface. It targets low-cost inference with standardized pricing across different upstream models. Details were checked against the provider's own page at api.oriper.com.
models: Llama-3.1-8B-Instruct, Mistral Nemo, DeepSeek Chat
unknownfree tier unknownconnectMedium risk - Cheaper Inferencecheaperinference.comCommercial aggregator
Cheaper Inference is a low-cost AI model aggregator offering frontier models like GPT-5.4 at significant discounts. It uses an OpenAI-compatible API interface. Details were checked against the provider's own page at cheaperinference.com.
models: gpt-5.4, Anthropic: Claude 3.5 Sonnet
unknownfree tier unknownconnectMedium risk - Chenzk APIchenzk.topCommercial aggregator
Chenzk API is a personal/community AI API relay service providing access to mainstream LLMs at normalized market rates. It uses an OpenAI-compatible base URL. Details were checked against the provider's own page at chenzk.top.
models: Anthropic: Claude 3.5 Sonnet, GPT-4o
unknownfree tier unknownconnectMedium risk - Chuteschutes.aiCommercial aggregator
Chutes is a decentralized serverless AI inference marketplace on Bittensor (Subnet 64). It provides an OpenAI-compatible API for deploying and scaling open-source models using high-performance VLLM templates, with per-token pricing and optional monthly subscription plans.
models: Meta-Llama 3.1, Mistral, DeepSeek V3.2
up · 184msfree tier · not quantifiedconnectMedium risk - Clarifaidocs.clarifai.comPaid API
Clarifai provides a production-ready AI API for developers, offering an OpenAI-compatible endpoint to run inferences on Clarifai-hosted models from labs like Meta, Google, and Mistral, alongside their own specialized vision and NLP models.
models: Mistral Large, Anthropic: Claude 3.5 Sonnet, Meta: Llama 3.1 405B Instruct
unknownfree tier · not quantifiedconnectMedium risk - Claude-ZHclaude-zh.cnCommercial aggregator
Claude-ZH (claude-zh.cn) is a third-party relay aggregator optimized for Chinese developers, providing access to Claude models via a homegrown CLI ('lucky') and OpenAI-compatible endpoints with local recharge options.
up · 913msno free tierconnectMedium risk - ClinePasscline.botSubscription router
ClinePass is a $9.99/month subscription service for the Cline IDE and CLI, bundling access to frontier open-weight models from labs like DeepSeek, Z.ai, and Qwen with generous quotas and no separate API key management.
models: glm-5.3-flash, [次]kimi-k2.5, DeepSeek Chat +3
unknownfree tier unknownconnectMedium risk - CoderPlancoderplan.aiCommercial aggregator
CoderPlan is a commercial AI API relay focused on developer tools like Claude Code and Gemini CLI. It provides pay-as-you-go top-up tiers for Chinese developers, claiming to match official API pricing for models like Claude and GPT.
up · 1025msfree tier · not quantifiedconnectMedium risk - Codex Cloudopenai.comOfficial free tier
OpenAI Codex is the specialized model family for coding tasks, now integrated into the flagship GPT models. It provides a specialized developer tier for high-performance code generation, planning, and task automation.
models: gpt-5.6-sol
unknownfree tier · not quantifiedconnectLow risk - Command Codecommandcode.aiSubscription router
Command Code is an AI coding API aggregator offering subscription-based access to a curated selection of open and premium models. It simplifies key management for developers by providing a unified gateway with generous monthly usage limits.
models: gpt-5.6-luna, Claude Haiku 4.5, claude-sonnet-5 +3
unknownfree tier unknownconnectMedium risk - CrofAIcrof.aiCommercial aggregator
CrofAI is a commercial AI API provider offering unified access to popular open models like Llama 3.1 and Qwen 2.5. It focuses on simplicity and ease of integration for developers requiring reliable per-token billing.
models: Llama-3.1-8B-Instruct, Mistral Nemo, Qwen2.5 72B Instruct
unknownfree tier unknownconnectMedium risk - Cursor APIcursor.comPaid API
Cursor API provides the backend intelligence for the Cursor AI editor, offering access to frontier models like Claude 3.5 Sonnet and GPT-4o. It features a tiered subscription model with on-demand usage billed for overages.
models: gpt-3.5-turbo, Anthropic: Claude 3.5 Sonnet, GPT-4o
unknownfree tier · not quantifiedconnectMedium risk - DGriddgrid.aiPaid API
DGrid AI is a decentralized AI network that routes requests across a distributed mesh of nodes. It features an OpenAI-compatible gateway and is integrated into tools like Chatbox and Claude Code. Operates with on-chain transparency.
models: Meta: Llama 3.1 405B Instruct, DeepSeek Chat, claude-sonnet-5 +2
unknownfree tier unknownconnectMedium risk - DIT.aidit.aiCommercial aggregator
DIT (AI Token Exchange) is a marketplace for AI model access, offering significant discounts (30-70%) compared to official provider rates. It exposes a single OpenAI-compatible API to 50+ models. Details were checked against the provider's own page at dit.ai.
models: DeepSeek Chat, claude-sonnet-5, Google Gemini Pro Latest +1
unknownfree tier unknownconnectMedium risk - Databrickswww.databricks.comPaid API
Databricks Foundation Model Serving provides enterprise-grade access to open foundation models like Llama 3.3 and DBRX. It is billed per token using Databricks Units (DBUs), integrated into the broader Databricks Data Intelligence Platform.
models: Meta: Llama 3.1 405B Instruct, Meta-Llama-3.1-70B-Instruct, Mistral: Mixtral 8x7B Instruct +1
unknownfree tier unknownconnectLow risk - DigitalOceandocs.digitalocean.comPaid API
DigitalOcean Inference provides a unified control plane for AI model inference. It offers serverless access to foundation models (Anthropic, OpenAI, DeepSeek, Kimi) at provider-aligned rates, plus dedicated GPU deployments.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Mistral: Mixtral 8x7B Instruct
unknownfree tier unknownconnectLow risk - Empowerdocs.empower.devPaid API
Empower provides an AI API router with an OpenAI-compatible endpoint. It focuses on bridging multiple model providers (Llama, etc.) through a single gateway for standard clients and SDKs. Details were checked against the provider's own page at docs.empower.dev.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct
unknownfree tier unknownconnectMedium risk - Factoryfactory.aiEnterprise
Factory provides the 'Autonomy Stack' for enterprise engineering teams, featuring autonomous 'Droids' that code, review PRs, and manage wikis. It offers model routing across 200+ models. Details were checked against the provider's own page at factory.ai.
models: minimax-m2.7, [次]kimi-k2.5, GLM-5.1
unknownfree tier unknownconnectLow risk - Featherless AIfeatherless.aiPaid API
Featherless AI provides serverless access to LLMs and agent runtimes. It offers subscription-based 'Chat' plans with unlimited tokens and credit-based 'Developer' plans for API usage. Details were checked against the provider's own page at featherless.ai.
models: Ling 3.0 Flash, Laguna S 2.1, Qwen3.8 27B +3
unknownfree tier · not quantifiedconnectLow risk - FenayAIfenayai.comPaid API
FenayAI is an AI API router and aggregator providing unified access to major models like GPT-6 Astra and Claude Fable. It features an OpenAI-compatible gateway with BYOK support. Details were checked against the provider's own page at fenayai.com.
models: DeepSeek V4 Flash, Gemini 3.8 Flash, GPT-6 Astra +1
unknownfree tier unknownconnectMedium risk - Free.aifree.aiFree product
Free.ai is an AI tool aggregator offering 477+ tools across chat, image, and video. It provides 30,000 free tokens daily for registered users and supports premium models like GPT-6 and Claude via paid plans. The platform features an in-browser IDE for deploying AI applications.
models: Claude 3 Haiku, Gemini 2.5 Flash, Llama-3.1-8B-Instruct +1
unknownfree tier · not quantifiedconnectMedium risk - FreeAIAPIKeyfreeaiapikey.comCommercial aggregator
FreeAIAPIKey is a commercial aggregator that provides unified access to frontier models including GPT-5.1, Claude 4.5, and Gemini 3 at claimed discounts of 80% off official rates. It offers a single OpenAI-compatible endpoint for easy drop-in replacement.
models: DeepSeek V4 Flash, Gemini 3.8 Flash, GPT-6 Astra +1
unknownfree tier unknownconnectHigh risk - FreeInferencefreeinference.orgOfficial free tier
FreeInference.org is a research-focused AI API gateway built at Harvard SEAS MadSys Lab. It provides free OpenAI-compatible access to frontier open models for the research and education community. No credit card is required for access.
models: Qwen3 Coder, MiniMax-M3, DeepSeek V4 Flash +1
unknownfree tier · not quantifiedconnectLow risk - FriendliAIfriendli.aiOfficial free tier
FriendliAI is a high-performance AI inference platform specializing in serving frontier open-weight models like GLM, DeepSeek, and Llama with industry-leading speed. It offers serverless Model APIs, dedicated GPU endpoints, and on-premise containers.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Mistral Nemo +1
unknownfree tier · not quantifiedconnectLow risk - Fuka APIapi.fuka.winCommercial aggregator
Fuka API is a comprehensive AI API infrastructure provider offering a unified gateway to a vast selection of models via the New-API protocol. It supports high concurrency, transparent pay-as-you-go billing, and multi-protocol compatibility for developers and teams.
up · 367msno free tierconnectMedium risk - GLM Coding (China)open.bigmodel.cnOfficial free tier
BigModel.cn (智谱AI) is the official platform for GLM series models, offering high-performance text and multimodal (image/video) inference. It features flagship models like GLM-5.3 and provides a permanent free tier for Flash models along with significant token grants for new users.
models: glm-5.3-flash, glm-5.3, GLM 5 Turbo
unknownfree tier · not quantifiedconnectLow risk - Gitlawb Opengateway (MiMo)opengateway.gitlawb.comCommercial aggregator
Opengateway (by Gitlawb) is a pay-as-you-go inference gateway providing unified OpenAI- and Anthropic-compatible access to models from Xiaomi MiMo, Google, MiniMax, and Qwen. It features smart routing via 'model: auto' and a permanent free tier for Nemotron 3 Ultra.
unknownfree tier · not quantifiedconnectMedium risk - GoAPIapi.getgoapi.comCommercial aggregator
GetGoAPI is an AI API aggregator that provides unified access to over 500 models, including GPT-6 and Claude 4.5. It offers a flat 20% discount compared to official provider pricing with no platform fees or monthly minimums.
models: DeepSeek V4 Flash, Gemini 3.8 Flash, GPT-6 Astra +1
unknownfree tier · not quantifiedconnectMedium risk - HelixMindhelixmind.onlineCommercial aggregator
HelixMind is an AI API router designed to manage and optimize costs across hundreds of LLMs through a single endpoint. It native speaks the OpenAI API format and features automatic model selection by price and task requirement, ensuring enterprise-grade reliability.
models: gpt-5.6-luna, Llama 4 Maverick, DeepSeek V4 Pro 0813 +3
unknownfree tier · not quantifiedconnectMedium risk - Helyx AIhelyxai.spaceFree relay
Helyx AI is a multi-model gateway providing access to 50+ frontier models like Claude 5 and GPT-4o. It features a unique 'three-stage waterfall' quota system including a large daily free claim, a permanent pool earned via referrals/reviews, and a paid credit balance.
models: Claude Haiku 4.5, deepseek-v4-flash-0731, Qwen3 32B +3
unknownfree tier · not quantifiedconnectHigh risk - Huancheng Public APIapi.hcnsec.cnCommercial aggregator
Huancheng Public API is a high-concurrency AI model router based on the New-API framework. It aggregates multiple upstream providers including OpenAI, Anthropic, and Google, offering a stable and low-latency endpoint for Chinese and global developers with real-time usage monitoring.
models: DeepSeek Chat, Qwen2.5 72B Instruct, Meta: Llama 3.1 405B Instruct +3
unknownfree tier unknownconnectHigh risk - IBM watsonx.ai Gatewaywww.ibm.comPaid API
IBM watsonx.ai Gateway provides managed access to foundation models from IBM and third parties. It offers an OpenAI-compatible endpoint and enterprise-grade tools for model governance and scaling. Access is available via pay-as-you-go billing per million tokens.
models: Meta-Llama-3.1-70B-Instruct, Mistral Large
unknownfree tier · not quantifiedconnectLow risk - JuziAIapi.juziai.ccCommercial aggregator
Juzi AI (api.juziai.cc) is a Chinese commercial-aggregator relay built on the New-API framework. It aggregates high-end models like GPT-5 and Claude Opus by leveraging various upstream account pools, offering pay-as-you-go access to users in mainland China.
models: [新稳]gemini-2.5-pro-maxthinking, nai-diffusion-5-curated, [kiro]claude-opus-4-7-thinking +3
up · 874msfree tier · not quantifiedconnectHigh risk - KIE.AIkie.aiPaid API
KIE.AI is an AI model API relay providing access to frontier models like Claude 3.5 Sonnet and GPT-4o. It operates on a pay-as-you-go basis and provides an integrated dashboard for API key management and usage tracking.
models: gemini-3.1-pro-preview, claude-opus-4-7, gpt-5.5
unknownfree tier unknownconnectMedium risk - Kenarikenari.idPaid API
Kenari is an Indonesian AI gateway providing a unified OpenAI-compatible API to 38+ models from Claude, GPT, and Gemini. It is uniquely tailored for the Indonesian market, accepting local payment methods like QRIS and billing in Rupiah (IDR).
models: Muse Spark 1.3 Contributor, glm-5.3-flash, deepseek-v4-flash-0731 +3
unknownfree tier · not quantifiedconnectMedium risk - Kilo Gatewaykilo.aiPaid API
Kilo Gateway is a developer-focused AI model router integrated with Kilo Code. It provides a unified gateway to multiple providers (OpenRouter, etc.) at their original rates with no added markup. It supports free models for initial development and pay-as-you-go for production usage.
models: GPT-4o, GPT-6 Astra
unknownfree tier · not quantifiedconnectMedium risk - LLM Gatewayllmgateway.ioPaid API
LLM Gateway is a unified AI API router that allows access to over 200 models from multiple providers. It features a 'Hosted Free' plan where free models are accessible with a 5 requests per 10 minutes limit when no credits are available in the account.
models: MiniMax-M3, glm-5.3, Gemini 3.8 Flash +3
unknownfree tier · not quantifiedconnectMedium risk - LLM.Kiwillm.kiwiPaid API
LLM.Kiwi is an AI model router offering access to various LLMs via an OpenAI-compatible endpoint. It provides a free tier with a rate limit of 40 requests per hour for specific high-resource models and higher limits for pro subscribers.
models: Auto Router
unknownfree tier · not quantifiedconnectMedium risk - LLM7api.llm7.ioFree relay
LLM7.io is an open-source AI API gateway providing unified access to leading models like GPT and DeepSeek. It stands out for its unique anonymous access tier, allowing developers to test models without registration, and a registration-based free tier with higher limits.
models: DeepSeek V4 Flash, kimi-k2.6, kimi-k3 +3
up · 273msfree tier · not quantifiedconnectMedium risk - Lambda AIlambda.aiPaid API
Lambda AI (Lambda Labs) is a premier GPU cloud provider offering serverless inference for large open-source models like Llama 3.1 405B. It provides a robust, developer-centric platform for high-performance AI tasks with usage-based billing.
models: Meta: Llama 3.1 405B Instruct, Mistral Large
unknownfree tier unknownconnectLow risk - LaoZhang AIapi.laozhang.aiPaid API
LaoZhang AI is an enterprise-grade AI API integration platform reselling over 200 models from OpenAI, Anthropic, and Google. It provides an OpenAI-compatible interface, invoice support for businesses, and a small signup credit ($0.05) for testing.
models: gpt-5.6-luna, gemini-3.6-flash, deepseek-v4-flash-0731 +1
unknownfree tier · not quantifiedconnectMedium risk - LiteRouterliterouter.comPaid API
LiteRouter is an AI API aggregator that provides access to hundreds of models from various providers via a single OpenAI-compatible interface. It supports free models with specific daily limits and pay-as-you-go billing for premium variants.
models: Gemma 4 31B, mimo-v2.5-pro, Llama-3.1-8B-Instruct +3
unknownfree tier · not quantifiedconnectMedium risk - LlamaGatellamagate.aiPaid API
LlamaGate is a unified AI model API gateway reselling access to Llama 3, Qwen, and DeepSeek models. It provides an OpenAI-compatible endpoint and focuses on simplified integration for developers. Details were checked against the provider's own page at llamagate.ai.
models: Qwen3.5-9B, Muse Spark 1.3, DeepSeek-R1 +3
unknownfree tier unknownconnectMedium risk - Logfarelogfare.aiPaid API
Logfare is an AI API router providing serverless inference for a wide range of open-source models. It includes access to many free-to-use models from providers like Liquid, Cohere, and NVIDIA, managed through a unified API dashboard.
models: inclusionAI: Ling 3.0 Flash Sante (free), LiquidAI: LFM2.5-2.6B (free), Cohere: North Mini Code (free) +3
unknownfree tier · not quantifiedconnectMedium risk - MNN AImnnai.ruCommercial aggregator
MNN AI (mnnai.ru) is a Russian AI API aggregator offering access to models like GPT-5.6 and Claude 5. It features a free entry tier with $1 monthly credits and strict rate limits (10 RPM), designed for low-volume testing and regional access.
models: GPT-5.6 Luna Pro, glm-5.3, DeepSeek V4 Pro 0813 +3
unknownfree tier · not quantifiedconnectMedium risk - Magnificwww.magnific.comPaid API
Magnific (formerly Freepik) is a professional AI creative platform offering image, video, and audio generation and upscaling. Its API provides credit-based access to proprietary models like Nano Banana 2 and Seedream 5.0, with subscription tiers ranging from a limited free plan to professional unlimited tiers.
unknownfree tier · not quantifiedconnectLow risk - Maritalkwww.maritaca.aiPaid API
Maritalk (by Maritaca AI) is a Brazilian AI provider specializing in Portuguese language models (Sabiá family). It offers an OpenAI-compatible API with inference hosted in Brazil for data sovereignty, featuring a daily free tier for chat and competitive PAYG token rates for the Sabiá-4 flagship model.
unknownfree tier · not quantifiedconnectMedium risk - MegaNova AImeganova.aiCommercial aggregator
MegaNova AI is a comprehensive AI API aggregator and infrastructure provider, offering access to 200+ models including GPT-5.6, Claude 4, and proprietary Manta models. It features a tiered free quota system based on account status and spend, plus OpenAI-compatible serverless and dedicated GPU endpoints.
models: MiniMax-M3, mimo-v2.5-pro, glm-5.3 +3
unknownfree tier · not quantifiedconnectMedium risk - Meta Llama APIllama.developer.meta.comPaid API
Meta Model API (developer.meta.com) is the official hosted inference service for Llama and Muse models. It provides OpenAI- and Anthropic-compatible endpoints with pay-as-you-go pricing for frontier models like Muse Spark and specialized models for transcription and reasoning.
models: Muse Glimmer 30B, Muse Spark 1.3, Muse Spark 1.2 +3
unknownfree tier unknownconnectLow risk - Minimax (China)www.minimaxi.comPaid API
MiniMax China (minimaxi.com) is the domestic-facing platform for MiniMax's foundation models. It mirrors the global platform's model offering and OpenAI-compatible API but optimizes for mainland China compliance and billing workflows.
models: MiniMax-M3, minimax-m2.7, MiniMax M2 +1
unknownfree tier · not quantifiedconnectMedium risk - Minimax Codingwww.minimax.ioPaid API
MiniMax (minimax.io) is a leading Chinese AI unicorn offering a multi-modal foundation model API. Its platform provides OpenAI-compatible access to the M3 and M2.7 model families, with a mix of monthly subscription plans (Plus/Max/Ultra) and permanent 50% discounts on pay-as-you-go token rates.
models: MiniMax-M3, minimax-m2.7, MiniMax M2 +2
unknownfree tier · not quantifiedconnectMedium risk - Mixedbread AIwww.mixedbread.comPaid API
Mixedbread AI is a specialized provider of knowledge-retrieval models and agents. Its platform offers an OpenAI-compatible API for its 'Toast' agent models, alongside managed search and indexing services billed by content tokens. It features a $5 one-time free credit for new users.
unknownfree tier · not quantifiedconnectLow risk - Mixlayerwww.mixlayer.comCommercial aggregator
Mixlayer is an AI API router and aggregator (mixlayer.com) offering transparent pay-as-you-go pricing for frontier models like GLM-5.3 and Qwen 3.5. It provides a free tier for small models like Qwen 3.5 4B and dedicated GPU hosting (H100/H200) billed by the hour.
models: Qwen3.5-Flash, deepseek-v4-flash-0731, Qwen3.5-9B +1
unknownfree tier · not quantifiedconnectMedium risk - Modelocmodeloc.comMonitor / Directory
Modeloc (modeloc.com) is an evaluation and compute-sharing platform for LLM relays. It features a 'compute pool' where users can share spare model quota to earn credits or spend credits to call pooled models, integrated with a trust-score leaderboard.
up · 891msfree tier · not quantifiedconnectHigh risk - MonsterAPImonsterapi.aiPaid API
MonsterAPI (monsterapi.ai) was an AI compute and model deployment platform. The service permanently shuttered operations on June 30, 2026, and its domains are no longer active. Details were checked against the provider's own page at www.crunchbase.com.
unknownno free tierconnectHigh risk - Morphmorphllm.comOfficial free tier
Morph (morphllm.com) is an AI provider focusing on agentic coding models. Its API offers high-speed code editing and search via proprietary models (Compact, Reflexes) and optimized frontier models (GLM, Kimi), with a limited free tier of 200 requests per month.
models: deepseek-v4-flash-0731, Qwen3.5-122B-A10B, glm-5.3 +3
unknownfree tier · not quantifiedconnectMedium risk - Muse Code (Meta)github.comPaid API
Muse Code is Meta's specialized AI coding environment and agent service, built on the Meta Model API. It provides a terminal-based interface for code generation and transformation, featuring both a monthly subscription plan for high usage and credit-based PAYG options via Muse Spark.
models: Llama 3.3 70B Instruct, Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct
unknownfree tier unknownconnectMedium risk - NVIDIA NIMintegrate.api.nvidia.comOfficial free tier
NVIDIA NIM (Build) is NVIDIA's official hosted model-inference service, providing API access to over 100 frontier open-source models (Llama 3.3, Nemotron, DeepSeek). It features a permanent free tier with a 40 RPM limit, authenticated via a standard API key and accessible through an OpenAI-compatible endpoint.
models: Nemotron 3 Nano 30B A3B, Nemotron 3.5 Lightning, Nemotron 3 Super +3
up · 14msfree tier · not quantifiedconnectLow risk - Naga.acnaga.acCommercial aggregator
Naga.ac is a unified AI aggregator providing access to over 240 models via one OpenAI-compatible API. It offers a free tier with 100 requests per day and a pay-as-you-go model with competitive rates, occasionally lower than direct provider pricing for select models.
models: Gemini 2.5 Flash Lite, Qwen3.8 Flash, Gemini 3.1 Flash Lite Preview +3
unknownfree tier · not quantifiedconnectMedium risk - NanoGPTnano-gpt.comCommercial aggregator
NanoGPT (nano-gpt.com) is a large-scale AI API aggregator providing access to 1,000+ models with strict 'list price' pay-as-you-go billing. It supports OpenAI-compatible endpoints for text, image, and video models, adding a flat 5% fee only for opt-in features like BYOK or pinned providers.
models: Gemini 2.5 Pro, GPT-4o, claude-sonnet-4-6
unknownfree tier unknownconnectMedium risk - NaraRouterbynara.idFree relay
NaraRouter (bynara.id) is a regional AI relay providing generous free token allowances (up to 7M tokens/day) for students and hobbyists. It supports an OpenAI-compatible API and Google login, resetting quotas daily at 07:00 WIB.
models: DeepSeek V4 Flash, glm-5.3-flash
unknownfree tier · not quantifiedconnectHigh risk - NavyAIapi.navyCommercial aggregator
NavyAI (api.navy) is a unified AI API gateway providing access to 150+ models through flat-rate subscription plans with daily token quotas. It offers an OpenAI-compatible interface, 'Login with Navy' OAuth for BYOP workflows, and plans ranging from a basic free tier to high-volume 'Apex' tiers.
models: DeepSeek V3.2, Llama 3.3 70B Instruct, Llama 4 Maverick +3
unknownfree tier · not quantifiedconnectMedium risk - NexusVAIapi.nexusvai.xyzCommercial aggregator
NexusVAI is a multi-model API platform providing a unified, OpenAI-compatible endpoint for frontier LLMs. It features a developer-centric playground (CancriCode) and offers pay-as-you-go access to a broad range of open and proprietary models with an emphasis on ease of integration.
up · 238msfree tier · not quantifiedconnectMedium risk - Nous Researchportal.nousresearch.comPaid API
Nous Research is an AI research lab offering the Hermes family of models via the Nous Portal. It provides an OpenAI-compatible endpoint for pay-as-you-go inference with no complex configuration required.
models: Hermes 3 405B Instruct, Hermes 3 70B Instruct
unknownfree tier unknownconnectMedium risk - OfoxAIofox.aiCommercial aggregator
OfoxAI is an enterprise AI gateway aggregating 100+ models from OpenAI, Anthropic, Google, and more. It offers official provider pricing with a 0% platform fee and includes built-in budget controls and audit logs.
models: DeepSeek V4 Flash Vision Exp, glm-5.3, gemini-3.7-flash +3
unknownfree tier unknownconnectMedium risk - OnAllwaysllm.onallways.topFree relay
OnAllways (llm.onallways.top) is a 'subscription-to-API' relay that pools various AI subscription accounts (Claude, GPT) and redistributes them as API keys. It is anonymously run and targets users looking for low-cost access to restricted models.
models: Claude series, Codex/GPT series, Gemini series (forwarded via pooled subscription accounts, dubious authenticity/stability)
up · 223msfree tier · not quantifiedconnectHigh risk - OpenAdapteropenadapter.devPaid API
OpenAdapter is an AI model API providing access to models like DeepSeek, Qwen, and GLM through an OpenAI-compatible endpoint. It focuses on easy integration by allowing standard clients to switch base URLs.
models: DeepSeek Chat, [次]kimi-k2.5, glm-5.3-flash +3
unknownfree tier unknownconnectMedium risk - OpenVectaopenvecta.comOfficial free tier
OpenVecta provides an AI model API with an OpenAI-compatible endpoint. It focuses on simplicity and ease of use, with reports of a free tier or trial credits available for new users. Details were checked against the provider's own page at openvecta.com.
models: Gemma 4 31B, GLM-4.7-Flash, DeepSeek V4 Flash +3
unknownfree tier · not quantifiedconnectMedium risk - OrcaRouterorcarouter.aiCommercial aggregator
OrcaRouter is an LLM gateway offering zero token markup and adaptive routing across 200+ models. It is OpenAI-compatible and supports both pay-as-you-go and subscription plans with a handful of free-tier models.
models: orcarouter/free, deepseek/deepseek-v4-flash-free, qwen/qwen3.8-27b-free +1
up · 152msfree tier · not quantifiedconnectMedium risk - PLaMoplamo.preferredai.jpPaid API
PLaMo is a Japanese LLM platform developing the PLaMo 3.0 Prime flagship model. It offers an OpenAI-compatible API optimized for high-performance Japanese language processing at a very low cost. Details were checked against the provider's own page at plamo.preferredai.jp.
unknownfree tier unknownconnectMedium risk - Perplexityperplexity.aiOfficial free tier
Perplexity provides the Sonar API for usage-based AI search and reasoning. It offers a limited free tier for end-users and a usage-based API for developers, optimized for real-time grounded answers. Details were checked against the provider's own page at perplexity.ai.
models: Perplexity default free search model (basic Sonar class)
up · 39msfree tier · not quantifiedconnectLow risk - Proxai.prox.us.ciCommercial aggregator
Prox (Aiproxy) is a closed AI API relay based on New-API, restricted to users with OAuth access from the external forum dc.hhhl.cc. It offers Stripe integration for paid top-ups and primarily serves a private community with aggregated access to frontier models.
up · 582msno free tierconnectHigh risk - Qwen Cloudwww.qwencloud.comPaid API
Qwen Cloud is the official developer platform for Alibaba's Qwen foundation models. It offers high-performance inference for the Qwen series (Max, Plus, Turbo) via an industry-standard OpenAI-compatible API, featuring credits-based billing and flexible subscription plans.
models: Qwen2.5 72B Instruct, Qwen-Plus, Qwen: Qwen-Turbo +3
unknownfree tier · not quantifiedconnectMedium risk - RRRapirrrapi.comCommercial aggregator
RRRapi is a third-party AI API relay built on the New API (One API) framework. It provides an OpenAI-compatible endpoint that aggregates multiple upstream models, featuring a balance-based top-up model with Stripe integration for credits.
up · 121msno free tierconnectMedium risk - Regolo AIregolo.aiPaid API
Regolo AI provides scalable serverless AI infrastructure with a focus on privacy and zero data retention. It serves popular open-source models (Llama, Mistral, Qwen) via an OpenAI-compatible API, offering monthly token capacity plans and a flexible free trial.
models: Qwen3.8 27B, Gemma 4 31B, gpt-oss-120b +3
unknownfree tier · not quantifiedconnectLow risk - Rekadocs.reka.aiPaid API
Reka AI develops natively multimodal foundation models (Spark, Edge, Flash, Core) capable of processing text, image, audio, and video. Their API offers high-performance inference with competitive usage-based pricing for both compact on-device models and large flagship models.
models: Reka Edge, Reka Flash 3
unknownfree tier unknownconnectMedium risk - RenAIrenai.unoCommercial aggregator
RenAI (renai.uno) is an OpenAI-compatible AI API gateway that aggregates multiple global and Chinese models, including OpenAI, Anthropic, Google, and DeepSeek. It operates on a balance top-up model and is designed for users needing high-compatibility access to a broad range of upstreams.
up · 232msno free tierconnectMedium risk - Routewayrouteway.aiPaid API
Routeway is an AI model router and gateway that provides a unified Bearer-auth interface for multiple global and domestic models. It simplifies integration for developers by handling model-specific protocols and providing a consistent management dashboard for keys and usage.
models: Muse Glimmer 30B, Gemma 4 26B A4B , Qwen3.8 27B +3
unknownfree tier · not quantifiedconnectMedium risk - SearchAPIwww.searchapi.ioSearch / Audio API
SearchAPI is a real-time SERP scraping API providing structured data from Google, Bing, Baidu, and more. It is designed for AI agents and developers, offering an OpenAI-compatible interface and a focus on 'pay-per-success' results.
unknownfree tier · not quantifiedconnectLow risk - SumoPodai.sumopod.comPaid API
SumoPod (ai.sumopod.com) is an AI model API aggregator, often associated with the GAIB compute ecosystem. It typically exposes an OpenAI-compatible gateway for multiple frontier models. Details were checked against the provider's own page at ai.sumopod.com.
models: Claude Haiku 4.5, Gemini 2.5 Pro, kimi-k2.6 +3
unknownfree tier unknownconnectMedium risk - SupXHapi.supxh.xinCommercial aggregator
SupXH (api.supxh.xin) is a generic AI API relay service providing a direct connection endpoint for various LLMs. It functions as a third-party aggregator with a minimal web presence, primarily serving as a host for OpenAI-compatible client configurations.
up · 54msno free tierconnectHigh risk - SupXH Freefree.supxh.xinFree relay
Shawn AI (肖恩AI) is a community-driven API relay station aimed at roleplay and tavern users. It aggregates hundreds of models including Claude 4.7 and GPT-5.5, offering free daily points and signing bonuses. It features a transparent point-deduction system and recovery for failed generations.
models: claude-opus-4-5/4-6, gpt-5.2/5.4/5.5, qwen3.7-max-t +3
up · 274msfree tier · not quantifiedconnectHigh risk - Syntheticsynthetic.newPaid API
Synthetic runs open-source AI models in private, secure datacenters with an emphasis on privacy and security. It offers an OpenAI-compatible API for integration with standard tools. Access is available via a monthly subscription or standard usage-based pricing.
models: glm-5.3-flash, Qwen3.8 27B, GLM-4.7-Flash +2
unknownfree tier unknownconnectMedium risk - TGhanapi.tghan.xyzCommercial aggregator
TGhan (Xiao Yu Yun API) is a feature-rich AI API relay providing unified access to over 30 LLM providers including OpenAI, Claude, and Grok. It offers professional features like empty-response compensation, transparent usage logs, and multi-protocol support optimized for high stability.
models: gemini-3-flash, gemini-3.1-pro-preview-thinking, [企业]gemini-2.5-pro-thinking +3
up · 685msno free tierconnectMedium risk - TabiTokentabitoken.comCommercial aggregator
TabiToken is a unified AI API platform built on the NewAPI protocol, providing low-latency access to a wide selection of models including OpenAI, Claude, and Gemini. It features real-time usage monitoring and multi-region deployment for global access.
unknownfree tier unknownconnectMedium risk - TheB.AItheb.aiCommercial aggregator
TheB.AI is a versatile AI model aggregator and chatbot platform providing access to multiple frontier models (GPT-4, Claude 3, Gemini) via a unified API. It features a pay-as-you-go model and an OpenAI-compatible gateway.
models: Gemini 2.5 Pro, claude-sonnet-5, GPT-4o
unknownfree tier · not quantifiedconnectMedium risk - Together AIwww.together.aiPaid API
Together AI is a leading cloud provider for open-source AI models, offering serverless inference, provisioned throughput, and dedicated clusters. It hosts over 100 models, including Llama, Qwen, and Mistral, with highly competitive per-token pricing.
models: Meta-Llama-3.1-70B-Instruct, DeepSeek Chat, Qwen2.5 72B Instruct +2
unknownfree tier unknownconnectLow risk - Token Kioskagent-router.gaib.aiCommercial aggregator
Token Kiosk (gaib.ai) is an AI model router and aggregator part of the GAIB AI infrastructure capital platform. it provides access to various open and proprietary models, potentially linked to compute-backed financial assets.
models: [次]kimi-k2.5, MiniMax-01, gemini-3.1-pro-preview
unknownfree tier unknownconnectMedium risk - TokenReplywww.tokenreply.comCommercial aggregator
TokenReply is a high-volume AI model aggregator offering access to advanced models (GPT-6 Astra, Claude 4.6, DeepSeek-V4 Pro) via an OpenAI-compatible gateway. It features pay-as-you-go billing with various tier-based model rates.
unknownno free tierconnectMedium risk - TokenRoutertokenrouter.comCommercial aggregator
TokenRouter (tokenrouter.com) is an AI model API aggregator providing a unified gateway to multiple frontier LLMs. It exposes an OpenAI-compatible endpoint for integration with standard SDKs and tools.
models: Claude 3 Haiku, DeepSeek Chat, Qwen-Plus +3
unknownfree tier unknownconnectMedium risk - UniExpandapi.uniexpand.comCommercial aggregator
UniExpand (Knowledge PRO API) is a unified AI application infrastructure provider. It offers a powerful API management platform to connect various frontier LLMs, focusing on enterprise-ready security, millisecond response times, and multi-region deployment for global access.
downfree tier · not quantifiedconnectMedium risk - UnoRouterunorouter.aiCommercial aggregator
UnoRouter is a unified AI API platform providing access to hundreds of models through a single intelligent endpoint. It is fully OpenAI-compatible and supports various clients including coding agents, character chat, and CLI tools. Features both pay-as-you-go and subscription plans with credit bonuses.
models: gemini-3.1-flash-lite, zai-glm-4.7, MiniMax-M3 +3
unknownfree tier · not quantifiedconnectMedium risk - Venice.aivenice.aiPaid API
Venice AI provides private, unrestricted access to leading AI models (text, image, audio, video) via an OpenAI-compatible API. It emphasizes zero data retention and hardware-verified privacy (TEE). Features a wide catalog of 300+ models including uncensored versions.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Meta: Llama 3.1 405B Instruct
unknownfree tier · not quantifiedconnectMedium risk - Void AIvoidai.appPaid API
Void AI is an AI API aggregator compatible with the OpenAI SDK. It uses a credit-based billing system where model costs are calculated using multipliers (e.g., GPT-5.1 is 0.75, GPT-4o is 1.25). Offers both flat-rate plans and usage-based top-ups.
models: GPT-5.1, gemini-3-flash-preview, GPT-4o +1
unknownfree tier · not quantifiedconnectMedium risk - Volcenginewww.volcengine.comPaid API
Volcengine is ByteDance's enterprise cloud platform. Its Ark (Huoshan Fangzhou) platform provides access to the Doubao (Doubao) model family and other high-performance LLMs. Supports OpenAI-compatible API calls and various billing plans including token-based monthly packages.
models: Qwen-Plus, zai-glm-4.7
unknownfree tier · not quantifiedconnectLow risk - Weights & Biases Inferencewandb.aiPaid API
Weights & Biases Weave provides an inference router and observability tools for AI development. It integrates with OpenRouter and other providers to offer a unified, OpenAI-compatible endpoint for model testing and production deployments.
models: gpt-oss-120b, gpt-oss:20b, Qwen3 235B A22B Thinking 2507 +2
unknownfree tier · not quantifiedconnectLow risk - Wuqiongyunai.wuqiongyun.topCommercial aggregator
Wuqiongyun is an anonymous Chinese AI API relay aggregating GPT, Claude, and Gemini models using the New-API panel. It specializes in reselling access with model-specific multipliers, often including 'anti-reverse' and 'kiro channel' model variants.
models: doubao-seedream-4-5-251128, doubao-seedream-5-0-260128, grok-imagine-image +3
up · 446msno free tierconnectHigh risk - X5Labx5lab.devPaid API
X5Lab provides an OpenAI-compatible API for accessing various large models. It allows users to authenticate using an API key (format: x5-...) via standard OpenAI clients. Access is based on a pay-as-you-go model with a documented dashboard for key management.
models: DeepSeek Chat, Gemini 2.5 Pro, claude-sonnet-5 +1
unknownfree tier unknownconnectMedium risk - Xiaolangwangyuexiaolangwangyue.gettoken.devCommercial aggregator
Xiaolangwangyue (integrated with GetToken.dev) is an AI API aggregator platform offering highly discounted balance top-up plans and subscriptions for ChatGPT (GPT-6/5.6/5.4) and other models. It follows a 1:1 official rate deduction from a prepaid USD balance bought with CNY at significant discounts.
up · 2957msfree tier · not quantifiedconnectMedium risk - Xiaomi MiMomimo.xiaomi.comOfficial free tier
Xiaomi MiMo (mimo.mi.com) is the official AI open platform for Xiaomi's in-house large models. It offers an OpenAI-compatible API for the MiMo-V2.5 series, featuring a high-concurrency token plan (¥60 for 50B tokens) and competitive per-million token pricing for Pro and Flash variants.
models: MiMo-V2.5, MiMo-V2-Flash, MiMo-V2-Pro (time-limited/quota-limited free) +1
up · 795msfree tier · not quantifiedconnectLow risk - Yi (01.AI)01.aiPaid API
Yi is the model family by 01.AI, offering high-performance LLMs including Yi-Lightning and Yi-Large. The developer platform provides OpenAI-compatible API access with usage-based billing. It is optimized for both Chinese and English language tasks.
unknownfree tier unknownconnectLow risk - Yolo-Autoyolo-auto.comSubscription router
Yolo-Auto offers flat-rate LLM API pricing specifically designed for coding agents and high-context developer workflows. It provides an OpenAI-compatible endpoint with no per-token billing on paid plans, emphasizing predictable costs and private inference (no routine prompt logging).
models: Qwen3.8 27B
unknownfree tier · not quantifiedconnectMedium risk - ZenMuxzenmux.aiCommercial aggregator
ZenMux is a commercial unified API gateway providing access to 100+ mainstream AI models. It features a unique 'LLM insurance' mechanism that compensates users for slow or failed requests. The platform aggregates models from official providers and authorized cloud partners.
models: moonshotai/kimi-k3-free, z-ai/glm-4.7-flash-free, z-ai/glm-4.6v-flash-free +2
up · 102msfree tier · not quantifiedconnectMedium risk - ZeroLimitAIwww.zerolimitai.comPaid API
ZeroLimitAI is an AI API aggregator offering free AI chat and a developer API with automatic model routing. It advertises a lifetime access plan for a one-time fee and a free tier for testing. The API is OpenAI-compatible.
models: Llama 4 Scout, Qwen3-235B-A22B, DeepSeek-R1
unknownfree tier · not quantifiedconnectMedium risk - Zylo APIzyloai.netPaid API
Zylo API provides a unified OpenAI-compatible gateway to models from Anthropic, OpenAI, Google, and more with zero token markup. Billing is based on base per-token rates deducted from a prepaid balance. A flat platform fee applies when adding credits.
models: Grok 4.20, deepseek-v4-pro, qwen3.7-MAX +3
unknownfree tier · not quantifiedconnectMedium risk - b.aib.aiCommercial aggregator
b.ai is an OpenAI-compatible LLM gateway (distinct from TheB.AI) focusing on unified multi-model access and privacy. It supports native crypto payments and anonymous access via the https://api.b.ai/v1 base URL.
models: Gemini 3.8 Flash, Qwen3.8 Max (0902), GPT-6 Astra +1
unknownfree tier · not quantifiedconnectMedium risk - g4f.space — Geminig4f.spaceFree relay
g4f.space is an AI API router providing access to multiple frontier models including Gemini. It emphasizes unrestricted, open-source access but now requires proof-of-work credits for anonymous use to maintain service quality.
models: Gemini 3.8 Flash, gemini-3.7-flash
unknownfree tier · not quantifiedconnectHigh risk - iFlyTek Spark Openspark-api-open.xf-yun.comOfficial free tier
spark-api-open.xf-yun.com is the OpenAI-compatible HTTP endpoint for iFlyTek's Spark (讯飞星火) LLM. It provides access to multimodal models via the iFlyTek Open Platform. A permanently free 'Lite' version is available for developers.
models: Spark Lite (model=lite), Spark Pro/Max/4.0 Ultra (one-off free token packs only; paid after depletion)
up · 822msfree tier · not quantifiedconnectLow risk
free-credit values are normalized USD estimates from public sources — see methodology · ordering is data (free-credit value), never sponsorship
What is an AI API router?
An AI API router is a service that exposes AI model APIs (often OpenAI- or Claude-compatible endpoints) so you can call many models — from different vendors or free and paid tiers — through one interface, with a single key and one billing surface.
How is the free-credit value calculated?
Free-credit values are normalized estimates in USD from public sources (signup bonuses, monthly quotas and usage allowances). Estimates are flagged on each entry; values we could not quantify are shown honestly as "Unknown" or "Not quantified" instead of invented numbers.
Is the ranking sponsored?
No. Every list on UPROUTER.ONLINE is ordered by data (free-credit value, price, risk rating or community score). There is no pay-to-rank — affiliate and referral links never influence data, scores or sorting.
How do I know a router is safe to use?
Each entry carries a risk rating (low / medium / high / unrated) with the reasoning behind it, a data-confidence score and live status. Risk is an evidence rating from community research — treat it as a starting point and always review the provider terms yourself.