$ uprouter — directory
Paid API AI API routers & providers
The Paid API category of the UPROUTER.ONLINE directory tracks 105 AI API routers and providers, each with normalized pricing, a free-tier value where one exists, a risk rating and a live status probe.
Across the category, 50 of them run a free tier worth $365.40 in free credits in total — the largest belong to Vertex AI, Amazon Bedrock, Voyage AI; 58 are Connect-ready (OpenAI- and Claude-compatible); 1 were up in the latest status check. These figures are recomputed from the live directory, so the counts shift as providers change their tiers and pricing.
Entries are ordered by free-credit value — a data sort, never sponsored — and every row links to a full entry with source URLs, a verification date and a change history. Inclusion is informational only; always verify current terms with the provider before you purchase.
- Vertex AIcloud.google.com
Google Cloud's enterprise-grade AI platform (Vertex AI) providing managed access to Gemini models and third-party models like Llama and Mistral. Offers advanced features like grounding with Google Search, model tuning, and provisioned throughput for predictable performance.
models: Gemini 2.5 Flash, Gemini 2.5 Pro, Meta: Llama 3.1 405B Instruct +1
unknown$300.00Low risk - Amazon Bedrockaws.amazon.com
Amazon Bedrock is a fully managed service that offers a choice of high-performing foundation models from leading AI companies like Anthropic, Meta, and Mistral via a single API. It is primarily pay-as-you-go with various service tiers (Standard, Flex, Priority).
models: Nova Lite 1.0, Nova Pro 1.0, Claude family +2
up · 65msno free tierLow risk - Voyage AIwww.voyageai.com
Voyage AI specializes in state-of-the-art embedding models and rerankers. It provides a managed API for high-quality vector embeddings (voyage-4 family) and efficient reranking, optimized for RAG applications. Offers a highly generous free tier for developers.
unknown$24.00Low risk - DeepInfradeepinfra.com
DeepInfra is a high-speed inference cloud for open-source AI models. It provides an OpenAI-compatible API for text generation, embeddings, and image generation. Authentication uses a standard API key from its dashboard.
models: Meta: Llama 3.1 405B Instruct, Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct +3
unknown$20.00estconnectLow risk - Sarvam AIdocs.sarvam.ai
Sarvam AI is a leading Indian AI startup focused on large-scale models for Indic languages. Their API platform provides access to the Sarvam series models for text and voice, targeting applications in the Indian market with high-performance regional language support.
models: Sarvam: Sarvam-1 (2B Instruct)
unknown$12.00Medium risk - v0 (Vercel)v0.dev
v0 by Vercel is an AI-powered full-stack web application builder that generates React components and backend logic. Usage is metered on input and output tokens which convert to credits. It provides access to various models including v0 Mini, Pro, and Max tiers.
unknown$5.00Low risk - Difydify.ai
Dify is an open-source LLM app development platform. Its cloud service (Dify Cloud) provides hosted model access via an API, with credits consumable across OpenAI, Anthropic, Gemini, and others. Details were checked against the provider's own page at dify.ai.
unknown$3.00estconnectLow risk - DeepAIdeepai.org
DeepAI offers a comprehensive suite of AI APIs for text-to-image, text generation, and audio. It features a $9.99/month Pro subscription that bundles substantial monthly allowances with affordable pay-as-you-go overage rates.
unknown$1.40Medium risk - PiAPIpiapi.ai
PiAPI offers a model aggregation platform with native API formats. It provides subscription tiers that include bonus credits and pay-as-you-go usage. New users receive a $0.50 starting credit. Details were checked against the provider's own page at piapi.ai.
models: GPT-4o, claude-sonnet-4-6
unknownfree tier unknownMedium risk - 360 AIai.360.cn
360 AI (360智脑) is a Chinese LLM API provider by 360 Group, offering conversational models with support for multiple file formats and LlamaIndex compatibility. Access is managed through a dashboard-generated API key, primarily serving enterprise and developer needs in China.
unknownfree tier unknownMedium risk - AI/ML APIaimlapi.com
AI/ML API is a major AI model aggregator providing unified access to 1000+ models from OpenAI, Anthropic, Google, DeepSeek, and more. It offers a single API key for LLM, image, voice, and video models with a focus on production-ready pay-as-you-go access.
models: Phi 4, Claude 3 Haiku, Gemini 2.5 Flash +3
unknownfree tier unknownconnectMedium risk - AINative Studioainative.studio
AINative Studio specializes in AI agent infrastructure, providing multimodal APIs, persistent agent memory, and MCP server hosting. It offers tiered subscription plans that include 'ZeroTime' credits for agentic workflows and specialized model access.
models: DeepSeek Chat, glm-5.3, kimi-k3
unknownfree tier unknownMedium risk - Ant Ling / Ring (inclusionAI)developer.ant-ling.com
Ant Ling (by inclusionAI) provides an OpenAI-compatible API for its proprietary Ling model family. It supports high-speed inference and is designed to integrate seamlessly with standard SDKs by switching the base URL, serving both developer and enterprise needs.
models: Ling 3.0 Flash, Ling 3.0 Flash Fin, inclusionAI: Ling 3.0 Flash Sante (free)
unknownfree tier unknownconnectMedium risk - Anthropicplatform.claude.com
Anthropic is a leading AI research company and official provider of the Claude model family. Its API platform (platform.claude.com) offers pay-as-you-go access to frontier models like Claude 3.7 Sonnet and Opus, emphasizing safety and steerability for enterprise applications.
models: Claude 3 Haiku, Claude Sonnet 4.5, Claude Opus 4.5 +2
unknownfree tier unknownMedium risk - Arcee AIarcee.ai
Arcee AI provides a domain-specific model platform (Arcee Conductor) and inference API for open-weight models. It uses a native API format but Conductor provides a unified interface. Access is pay-as-you-go with no added premium on 3rd party model inference costs.
models: Trinity Large Thinking
unknownfree tier unknownMedium risk - Aurikowww.auriko.ai
Auriko is an AI API router and aggregator that provides a unified interface to multiple LLM providers. It offers both platform-managed and Bring Your Own Key (BYOK) modes. It publishes high rate limits for both modes.
unknownfree tier unknownconnectMedium risk - Azure AI Foundrylearn.microsoft.com
Azure AI Foundry (formerly Azure AI Studio) provides access to a vast catalog of models including OpenAI, Anthropic, and Llama. It features a unified API surface and enterprise-grade management. Pricing is unified under Azure's pay-as-you-go model.
models: Claude 3 Haiku, DeepSeek Chat, Phi 4 +1
unknownfree tier unknownconnectLow risk - Azure OpenAIazure.microsoft.com
Azure OpenAI Service provides REST API access to OpenAI's powerful language models including the GPT-4 and GPT-5 series with Azure's enterprise capabilities. It uses a strictly pay-as-you-go or provisioned throughput model.
models: gpt-5.6-luna, GPT-4o-mini, GPT-4o +2
unknownfree tier unknownconnectLow risk - Baichuanwww.baichuan-ai.com
Baichuan is a leading Chinese AI company offering large-scale language models. Its platform provides model inference via native and OpenAI-compatible APIs, focusing on high-quality Chinese language performance.
unknownfree tier · not quantifiedconnectMedium risk - Baidu (ERNIE)ernie.baidu.com
Baidu's Qianfan platform provides access to the ERNIE Bot model family. It is a full-stack AI development platform offering model training, fine-tuning, and inference. New users typically receive a small trial credit.
models: Baidu: ERNIE Lite, Baidu: ERNIE Speed, Baidu: ERNIE 3.5 +2
unknownfree tier · not quantifiedconnectMedium risk - Baidu Qianfancloud.baidu.com
Baidu Qianfan is the official Model-as-a-Service (MaaS) platform of Baidu Intelligent Cloud. It serves the ERNIE (Wenxin Yiyan) series models and other third-party models via an OpenAI-compatible v2 API, offering high-performance inference for enterprise and individual developers.
models: Baidu: ERNIE 4.0 Turbo, ERNIE 4.5 VL 424B A47B , Baidu: ERNIE 4.0 +1
unknownfree tier · not quantifiedconnectMedium risk - Blackbox AIblackbox.ai
Blackbox AI provides a unified endpoint for 300+ models with end-to-end encryption. The service has migrated to an enterprise-focused gated model (enterprise.blackbox.ai) after deprecating its public inference surface (api.blackbox.ai).
models: Anthropic: Claude 3.5 Sonnet, Google: Gemini Pro 1.5, GPT-4o +1
unknownno free tierconnectHigh risk - BytePlus ModelArkconsole.byteplus.com
BytePlus ModelArk is ByteDance's international enterprise AI platform providing access to the Doubao model family. It features a tiered pricing structure and integrates with BytePlus's global cloud ecosystem.
models: ByteDance: Doubao Lite 32k, ByteDance: Doubao Pro 32k, Seed 2.1 Turbo +1
unknownfree tier unknownMedium risk - Charm Hyperhyper.charm.land
Hyper (by Charm) is an AI inference solution purpose-built for coding. It manages its own infrastructure to optimize open-source coding models. It provides 100 free Hypercredits monthly on signup. Details were checked against the provider's own page at hyper.charm.land.
models: DeepSeek: DeepSeek Coder V2 Instruct, Qwen2.5 72B Instruct, Meta-Llama-3.1-70B-Instruct +1
unknownfree tier · not quantifiedconnectMedium risk - Chat Oripeapi.oriper.com
Chat Oripe is an AI API aggregator offering access to multiple LLM families via a unified interface. It targets low-cost inference with standardized pricing across different upstream models. Details were checked against the provider's own page at api.oriper.com.
models: Llama-3.1-8B-Instruct, Mistral Nemo, DeepSeek Chat
unknownfree tier unknownconnectMedium risk - Clarifaidocs.clarifai.com
Clarifai provides a production-ready AI API for developers, offering an OpenAI-compatible endpoint to run inferences on Clarifai-hosted models from labs like Meta, Google, and Mistral, alongside their own specialized vision and NLP models.
models: Mistral Large, Anthropic: Claude 3.5 Sonnet, Meta: Llama 3.1 405B Instruct
unknownfree tier · not quantifiedconnectMedium risk - Cozecoze.com
Coze is an all-in-one AI bot building platform that allows users to create and deploy agents with multiple models. It provides a production API for developers to integrate these bots into external applications with usage-based billing.
models: Llama-3.1-8B-Instruct, Anthropic: Claude 3.5 Sonnet, GPT-4o
unknownfree tier unknownMedium risk - Cursor APIcursor.com
Cursor API provides the backend intelligence for the Cursor AI editor, offering access to frontier models like Claude 3.5 Sonnet and GPT-4o. It features a tiered subscription model with on-demand usage billed for overages.
models: gpt-3.5-turbo, Anthropic: Claude 3.5 Sonnet, GPT-4o
unknownfree tier · not quantifiedconnectMedium risk - DGriddgrid.ai
DGrid AI is a decentralized AI network that routes requests across a distributed mesh of nodes. It features an OpenAI-compatible gateway and is integrated into tools like Chatbox and Claude Code. Operates with on-chain transparency.
models: Meta: Llama 3.1 405B Instruct, DeepSeek Chat, claude-sonnet-5 +2
unknownfree tier unknownconnectMedium risk - Databrickswww.databricks.com
Databricks Foundation Model Serving provides enterprise-grade access to open foundation models like Llama 3.3 and DBRX. It is billed per token using Databricks Units (DBUs), integrated into the broader Databricks Data Intelligence Platform.
models: Meta: Llama 3.1 405B Instruct, Meta-Llama-3.1-70B-Instruct, Mistral: Mixtral 8x7B Instruct +1
unknownfree tier unknownconnectLow risk - DigitalOceandocs.digitalocean.com
DigitalOcean Inference provides a unified control plane for AI model inference. It offers serverless access to foundation models (Anthropic, OpenAI, DeepSeek, Kimi) at provider-aligned rates, plus dedicated GPU deployments.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Mistral: Mixtral 8x7B Instruct
unknownfree tier unknownconnectLow risk - Doubaodoubao.com
Doubao (by Bytedance/Volcengine) provides a suite of models including the high-performance Doubao-Pro and Lite series. It is a major Chinese cloud provider with native API endpoints. Details were checked against the provider's own page at www.volcengine.com.
models: ByteDance: Doubao Lite 32k, ByteDance: Doubao Pro 32k
unknownfree tier unknownLow risk - Empowerdocs.empower.dev
Empower provides an AI API router with an OpenAI-compatible endpoint. It focuses on bridging multiple model providers (Llama, etc.) through a single gateway for standard clients and SDKs. Details were checked against the provider's own page at docs.empower.dev.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct
unknownfree tier unknownconnectMedium risk - Fal.aifal.ai
Fal.ai provides serverless inference and GPU compute for media generation models (Wan, Kling, Flux). It offers output-based pricing for image/video and time-based pricing for dedicated GPU fleets. Details were checked against the provider's own page at fal.ai.
unknownfree tier · not quantifiedLow risk - Featherless AIfeatherless.ai
Featherless AI provides serverless access to LLMs and agent runtimes. It offers subscription-based 'Chat' plans with unlimited tokens and credit-based 'Developer' plans for API usage. Details were checked against the provider's own page at featherless.ai.
models: Ling 3.0 Flash, Laguna S 2.1, Qwen3.8 27B +3
unknownfree tier · not quantifiedconnectLow risk - FenayAIfenayai.com
FenayAI is an AI API router and aggregator providing unified access to major models like GPT-6 Astra and Claude Fable. It features an OpenAI-compatible gateway with BYOK support. Details were checked against the provider's own page at fenayai.com.
models: DeepSeek V4 Flash, Gemini 3.8 Flash, GPT-6 Astra +1
unknownfree tier unknownconnectMedium risk - GigaChat (Sber)developers.sber.ru
GigaChat is an AI model series developed by Sberbank, optimized for the Russian language and multimodal tasks. It provides a robust API for developers featuring both chat and reasoning capabilities, primarily targeting the Russian AI ecosystem.
unknownfree tier unknownMedium risk - GitLab Duo PATdocs.gitlab.com
GitLab Duo is a suite of AI-powered features integrated into the GitLab DevSecOps platform. It uses models like Claude and Gemini to provide code suggestions and chat. API access is powered by 'GitLab Credits', which can be purchased or included in premium plans.
models: claude-sonnet-5, gemini-3.1-pro-preview
unknownfree tier unknownLow risk - Google Julesjules.google
Google Jules is an AI-driven cloud coding agent platform. It allows developers to automate complex coding tasks and manage cloud resources via a dedicated API. It is part of Google's broader AI ecosystem for developers.
models: Gemini 2.5 Pro, gemini-3.1-pro-preview
unknownfree tier unknownLow risk - Haiperhaiper.ai
Haiper AI is an advanced generative video platform offering text-to-video, image-to-video, and video-to-video capabilities. Its API provides programmatic access to Haiper Video 2.x and 1.5 models with support for multiple resolutions, billed per second of generated video.
unknownfree tier unknownMedium risk - Heroku AIwww.heroku.com
Heroku AI is a suite of managed AI services on the Salesforce Heroku platform. It enables developers to integrate and scale AI features using vector databases and third-party AI add-ons. It focuses on simplifying the transition from prototype to production for AI applications.
models: Nova Lite 1.0, MiniMax M2, Qwen: Qwen3 72B
unknownfree tier unknownLow risk - IBM watsonx.ai Gatewaywww.ibm.com
IBM watsonx.ai Gateway provides managed access to foundation models from IBM and third parties. It offers an OpenAI-compatible endpoint and enterprise-grade tools for model governance and scaling. Access is available via pay-as-you-go billing per million tokens.
models: Meta-Llama-3.1-70B-Instruct, Mistral Large
unknownfree tier · not quantifiedconnectLow risk - Ideogramideogram.ai
Ideogram is a specialized AI image generation platform offering a production-grade API. It supports text-to-image, remixing, and advanced editing features like layer separation (Layerize). The API uses a token-based system for both text and image modalities, compatible with standard integration patterns.
unknownfree tier · not quantifiedMedium risk - Jina AI (Foundation API)jina.ai
Jina AI (Foundation API) provides specialized AI models for search and RAG, including high-quality embeddings, reranking, and a Reader API that converts web content to LLM-friendly markdown. The Reader API is currently free for developers, while other APIs follow token-based usage billing.
unknownfree tier · not quantifiedLow risk - KIE.AIkie.ai
KIE.AI is an AI model API relay providing access to frontier models like Claude 3.5 Sonnet and GPT-4o. It operates on a pay-as-you-go basis and provides an integrated dashboard for API key management and usage tracking.
models: gemini-3.1-pro-preview, claude-opus-4-7, gpt-5.5
unknownfree tier unknownconnectMedium risk - Kenarikenari.id
Kenari is an Indonesian AI gateway providing a unified OpenAI-compatible API to 38+ models from Claude, GPT, and Gemini. It is uniquely tailored for the Indonesian market, accepting local payment methods like QRIS and billing in Rupiah (IDR).
models: Muse Spark 1.3 Contributor, glm-5.3-flash, deepseek-v4-flash-0731 +3
unknownfree tier · not quantifiedconnectMedium risk - Kilo Gatewaykilo.ai
Kilo Gateway is a developer-focused AI model router integrated with Kilo Code. It provides a unified gateway to multiple providers (OpenRouter, etc.) at their original rates with no added markup. It supports free models for initial development and pay-as-you-go for production usage.
models: GPT-4o, GPT-6 Astra
unknownfree tier · not quantifiedconnectMedium risk - LLM Gatewayllmgateway.io
LLM Gateway is a unified AI API router that allows access to over 200 models from multiple providers. It features a 'Hosted Free' plan where free models are accessible with a 5 requests per 10 minutes limit when no credits are available in the account.
models: MiniMax-M3, glm-5.3, Gemini 3.8 Flash +3
unknownfree tier · not quantifiedconnectMedium risk - LLM.Kiwillm.kiwi
LLM.Kiwi is an AI model router offering access to various LLMs via an OpenAI-compatible endpoint. It provides a free tier with a rate limit of 40 requests per hour for specific high-resource models and higher limits for pro subscribers.
models: Auto Router
unknownfree tier · not quantifiedconnectMedium risk - Lambda AIlambda.ai
Lambda AI (Lambda Labs) is a premier GPU cloud provider offering serverless inference for large open-source models like Llama 3.1 405B. It provides a robust, developer-centric platform for high-performance AI tasks with usage-based billing.
models: Meta: Llama 3.1 405B Instruct, Mistral Large
unknownfree tier unknownconnectLow risk - LaoZhang AIapi.laozhang.ai
LaoZhang AI is an enterprise-grade AI API integration platform reselling over 200 models from OpenAI, Anthropic, and Google. It provides an OpenAI-compatible interface, invoice support for businesses, and a small signup credit ($0.05) for testing.
models: gpt-5.6-luna, gemini-3.6-flash, deepseek-v4-flash-0731 +1
unknownfree tier · not quantifiedconnectMedium risk - Leonardo AIleonardo.ai
Leonardo AI offers a high-performance image generation API for creative workflows. It supports advanced features like text-to-image, image-to-image, upscaling, and motion generation, all accessible via a unified production-ready API.
unknownfree tier unknownMedium risk - Liquid AIliquid.ai
Liquid AI develops Liquid Foundation Models (LFMs), a new class of efficient, hybrid AI models. The platform offers serverless inference for models like LFM-2.5-2.6B, which are free for use by companies with annual revenue under $10 million.
models: LiquidAI: LFM2.5-2.6B (free)
unknownfree tier · not quantifiedLow risk - LiteRouterliterouter.com
LiteRouter is an AI API aggregator that provides access to hundreds of models from various providers via a single OpenAI-compatible interface. It supports free models with specific daily limits and pay-as-you-go billing for premium variants.
models: Gemma 4 31B, mimo-v2.5-pro, Llama-3.1-8B-Instruct +3
unknownfree tier · not quantifiedconnectMedium risk - LlamaGatellamagate.ai
LlamaGate is a unified AI model API gateway reselling access to Llama 3, Qwen, and DeepSeek models. It provides an OpenAI-compatible endpoint and focuses on simplified integration for developers. Details were checked against the provider's own page at llamagate.ai.
models: Qwen3.5-9B, Muse Spark 1.3, DeepSeek-R1 +3
unknownfree tier unknownconnectMedium risk - Logfarelogfare.ai
Logfare is an AI API router providing serverless inference for a wide range of open-source models. It includes access to many free-to-use models from providers like Liquid, Cohere, and NVIDIA, managed through a unified API dashboard.
models: inclusionAI: Ling 3.0 Flash Sante (free), LiquidAI: LFM2.5-2.6B (free), Cohere: North Mini Code (free) +3
unknownfree tier · not quantifiedconnectMedium risk - Magnificwww.magnific.com
Magnific (formerly Freepik) is a professional AI creative platform offering image, video, and audio generation and upscaling. Its API provides credit-based access to proprietary models like Nano Banana 2 and Seedream 5.0, with subscription tiers ranging from a limited free plan to professional unlimited tiers.
unknownfree tier · not quantifiedconnectLow risk - Maritalkwww.maritaca.ai
Maritalk (by Maritaca AI) is a Brazilian AI provider specializing in Portuguese language models (Sabiá family). It offers an OpenAI-compatible API with inference hosted in Brazil for data sovereignty, featuring a daily free tier for chat and competitive PAYG token rates for the Sabiá-4 flagship model.
unknownfree tier · not quantifiedconnectMedium risk - Meta Llama APIllama.developer.meta.com
Meta Model API (developer.meta.com) is the official hosted inference service for Llama and Muse models. It provides OpenAI- and Anthropic-compatible endpoints with pay-as-you-go pricing for frontier models like Muse Spark and specialized models for transcription and reasoning.
models: Muse Glimmer 30B, Muse Spark 1.3, Muse Spark 1.2 +3
unknownfree tier unknownconnectLow risk - Minimax (China)www.minimaxi.com
MiniMax China (minimaxi.com) is the domestic-facing platform for MiniMax's foundation models. It mirrors the global platform's model offering and OpenAI-compatible API but optimizes for mainland China compliance and billing workflows.
models: MiniMax-M3, minimax-m2.7, MiniMax M2 +1
unknownfree tier · not quantifiedconnectMedium risk - Minimax Codingwww.minimax.io
MiniMax (minimax.io) is a leading Chinese AI unicorn offering a multi-modal foundation model API. Its platform provides OpenAI-compatible access to the M3 and M2.7 model families, with a mix of monthly subscription plans (Plus/Max/Ultra) and permanent 50% discounts on pay-as-you-go token rates.
models: MiniMax-M3, minimax-m2.7, MiniMax M2 +2
unknownfree tier · not quantifiedconnectMedium risk - Mixedbread AIwww.mixedbread.com
Mixedbread AI is a specialized provider of knowledge-retrieval models and agents. Its platform offers an OpenAI-compatible API for its 'Toast' agent models, alongside managed search and indexing services billed by content tokens. It features a $5 one-time free credit for new users.
unknownfree tier · not quantifiedconnectLow risk - MonsterAPImonsterapi.ai
MonsterAPI (monsterapi.ai) was an AI compute and model deployment platform. The service permanently shuttered operations on June 30, 2026, and its domains are no longer active. Details were checked against the provider's own page at www.crunchbase.com.
unknownno free tierconnectHigh risk - Muse Code (Meta)github.com
Muse Code is Meta's specialized AI coding environment and agent service, built on the Meta Model API. It provides a terminal-based interface for code generation and transformation, featuring both a monthly subscription plan for high usage and credit-based PAYG options via Muse Spark.
models: Llama 3.3 70B Instruct, Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct
unknownfree tier unknownconnectMedium risk - Naver CLOVA Studioapi.ncloud-docs.com
Naver CLOVA Studio is a hyperscale AI service platform using the HyperCLOVA X model. It provides APIs for text generation, chat, and specialized tuning, primarily serving the Korean language market via Naver Cloud Platform.
models: Naver: HyperCLOVA X
unknownfree tier unknownMedium risk - Nous Researchportal.nousresearch.com
Nous Research is an AI research lab offering the Hermes family of models via the Nous Portal. It provides an OpenAI-compatible endpoint for pay-as-you-go inference with no complex configuration required.
models: Hermes 3 405B Instruct, Hermes 3 70B Instruct
unknownfree tier unknownconnectMedium risk - Nube.shnube.sh
Nube Cloud is a specialized inference provider focused on cutting costs for open-source models like GLM, DeepSeek, and Qwen. It offers a zero-data-retention policy and deeply optimized routing for lightning-fast latency.
models: DeepSeek V4 Flash, glm-5.3-flash, Qwen3.8 Flash +1
unknownfree tier unknownMedium risk - OCI Generative AIwww.oracle.com
OCI Generative AI is Oracle's enterprise-grade inference service hosting Llama and Cohere models. It provides on-demand and dedicated hosting with native integration into the Oracle Cloud ecosystem. Details were checked against the provider's own page at www.oracle.com.
models: Command R (08-2024), Meta-Llama-3.1-70B-Instruct, Command R+ (08-2024)
unknownfree tier unknownLow risk - OpenAIplatform.openai.com
OpenAI is the developer of GPT-4o, o1, and other leading LLMs. It offers a usage-based API for developers with tiered processing (Fast/Batch) and extensive multimodal capabilities including audio and video generation.
models: GPT-4o-mini, o3 Mini, o4 Mini +3
unknownfree tier unknownLow risk - OpenAdapteropenadapter.dev
OpenAdapter is an AI model API providing access to models like DeepSeek, Qwen, and GLM through an OpenAI-compatible endpoint. It focuses on easy integration by allowing standard clients to switch base URLs.
models: DeepSeek Chat, [次]kimi-k2.5, glm-5.3-flash +3
unknownfree tier unknownconnectMedium risk - OpenCode Goopencode.ai
OpenCode Go is a $10/month subscription service providing generous limits to curated open coding models like GLM-5.3-Flash and DeepSeek V4 Flash. It is optimized for agentic coding and works with standard agents.
models: glm-5.3-flash, Qwen2.5 Coder 32B Instruct
unknownfree tier unknownMedium risk - PLaMoplamo.preferredai.jp
PLaMo is a Japanese LLM platform developing the PLaMo 3.0 Prime flagship model. It offers an OpenAI-compatible API optimized for high-performance Japanese language processing at a very low cost. Details were checked against the provider's own page at plamo.preferredai.jp.
unknownfree tier unknownconnectMedium risk - Poolsidepoolside.ai
Poolside develops foundation models optimized for software engineering. Their API provides access to the Laguna series models (Laguna S, Laguna XS) with an emphasis on code generation and technical reasoning.
models: Laguna S 2.1, Laguna XS 2.1
unknownfree tier · not quantifiedMedium risk - Predibasepredibase.com
Predibase is an enterprise-grade low-code platform for fine-tuning and serving small language models. As of mid-2026, evidence suggests their managed public serving API has been discontinued or moved behind a private VPC-only model.
unknownno free tierHigh risk - PublicAIpublicai.co
PublicAI is a decentralized AI data network and API gateway. It facilitates access to LLMs by leveraging a community-governed data infrastructure, offering an API platform for developers to integrate models with a focus on data transparency.
models: Llama 3.3 70B Instruct, DeepSeek Chat, Mistral Large 2407
unknownfree tier unknownMedium risk - Qwen Cloudwww.qwencloud.com
Qwen Cloud is the official developer platform for Alibaba's Qwen foundation models. It offers high-performance inference for the Qwen series (Max, Plus, Turbo) via an industry-standard OpenAI-compatible API, featuring credits-based billing and flexible subscription plans.
models: Qwen2.5 72B Instruct, Qwen-Plus, Qwen: Qwen-Turbo +3
unknownfree tier · not quantifiedconnectMedium risk - Recraftrecraft.ai
Recraft is a specialized AI design platform and API for high-quality vector and raster graphics generation. It allows developers to generate illustrations, icons, and 3D assets with consistent style controls and branding presets.
unknownfree tier · not quantifiedMedium risk - Regolo AIregolo.ai
Regolo AI provides scalable serverless AI infrastructure with a focus on privacy and zero data retention. It serves popular open-source models (Llama, Mistral, Qwen) via an OpenAI-compatible API, offering monthly token capacity plans and a flexible free trial.
models: Qwen3.8 27B, Gemma 4 31B, gpt-oss-120b +3
unknownfree tier · not quantifiedconnectLow risk - Rekadocs.reka.ai
Reka AI develops natively multimodal foundation models (Spark, Edge, Flash, Core) capable of processing text, image, audio, and video. Their API offers high-performance inference with competitive usage-based pricing for both compact on-device models and large flagship models.
models: Reka Edge, Reka Flash 3
unknownfree tier unknownconnectMedium risk - Routewayrouteway.ai
Routeway is an AI model router and gateway that provides a unified Bearer-auth interface for multiple global and domestic models. It simplifies integration for developers by handling model-specific protocols and providing a consistent management dashboard for keys and usage.
models: Muse Glimmer 30B, Gemma 4 26B A4B , Qwen3.8 27B +3
unknownfree tier · not quantifiedconnectMedium risk - Runwaydocs.dev.runwayml.com
Runway is a leading AI research company focusing on generative video and image tools. Their developer API provides programmatic access to Gen-3 Alpha and Gen-2 models, enabling creators to integrate advanced video generation into their own workflows and products.
unknownfree tier · not quantifiedMedium risk - SAP Generative AI Hubhelp.sap.com
SAP Generative AI Hub provides enterprise-grade access to foundation models within the SAP AI Core environment. It ensures secure and compliant integration of LLMs like GPT-4 and Claude 3.5 Sonnet into SAP business applications, featuring unified management and data residency controls.
models: gpt-3.5-turbo, GPT-4o, Anthropic: Claude 3.5 Sonnet +2
unknownfree tier unknownMedium risk - SEA-LIONsea-lion.ai
SEA-LION (Southeast Asian Languages In One Network) is a suite of LLMs specifically optimized for Southeast Asian languages and cultural contexts. The API platform allows developers to integrate these regional models into localized applications with easy Google-auth signup.
unknownfree tier · not quantifiedMedium risk - Segmindsegmind.com
Segmind is a serverless AI platform providing access to a wide range of models via API, including image generation (Stable Diffusion, Flux), LLMs, and more. It offers a pay-as-you-go model with monthly credits included in paid plans. Access is configured via an API key created in the dashboard.
unknownfree tier · not quantifiedMedium risk - Snowflake Cortexwww.snowflake.com
Snowflake Cortex is a managed service within the Snowflake Data Cloud that provides access to LLMs (Llama, Mistral, Gemma) and AI functions. It uses a credit-based system where usage is billed per million tokens processed, independent of the Snowflake edition.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Gemma 2 27B +1
unknownfree tier unknownLow risk - Speka AIspeka.me
Speka is a unified AI API gateway providing access to frontier models (DeepSeek, Nemotron, Llama) via a single API key. It uses a subscription model with monthly usage allowances and standard per-token overage rates.
models: DeepSeek V4 Flash, Llama 3.3 70B Instruct, gpt-oss-120b +2
unknownfree tier · not quantifiedMedium risk - Stability AIstability.ai
Stability AI provides a developer platform for its generative models, including Stable Diffusion 3.5 and Stable Audio. The API uses a credit-based system where different tasks (image generation, upscaling, editing) consume varying amounts of credits.
unknownfree tier · not quantifiedLow risk - SumoPodai.sumopod.com
SumoPod (ai.sumopod.com) is an AI model API aggregator, often associated with the GAIB compute ecosystem. It typically exposes an OpenAI-compatible gateway for multiple frontier models. Details were checked against the provider's own page at ai.sumopod.com.
models: Claude Haiku 4.5, Gemini 2.5 Pro, kimi-k2.6 +3
unknownfree tier unknownconnectMedium risk - Sunosuno.ai
Suno is a leading generative AI platform for music and audio. It offers subscription-based access to its music generation models (v3.5, v4) with daily or monthly credit allowances. Credits are used to generate songs, extend tracks, and extract stems.
unknownfree tier · not quantifiedLow risk - Syntheticsynthetic.new
Synthetic runs open-source AI models in private, secure datacenters with an emphasis on privacy and security. It offers an OpenAI-compatible API for integration with standard tools. Access is available via a monthly subscription or standard usage-based pricing.
models: glm-5.3-flash, Qwen3.8 27B, GLM-4.7-Flash +2
unknownfree tier unknownconnectMedium risk - Tencent Hunyuanhunyuan.tencent.com
Tencent Hunyuan (混元) is a series of large-scale AI models integrated with the Tencent ecosystem (WeChat, Tencent Cloud). It offers strong Chinese language support and multimodal capabilities, including vision and code generation.
models: Hunyuan A13B Instruct, Hy4 preview, Hy3
unknownfree tier · not quantifiedLow risk - Together AIwww.together.ai
Together AI is a leading cloud provider for open-source AI models, offering serverless inference, provisioned throughput, and dedicated clusters. It hosts over 100 models, including Llama, Qwen, and Mistral, with highly competitive per-token pricing.
models: Meta-Llama-3.1-70B-Instruct, DeepSeek Chat, Qwen2.5 72B Instruct +2
unknownfree tier unknownconnectLow risk - Topaztopazlabs.com
Topaz Labs specializes in AI-powered image and video enhancement software (Photo AI, Video AI). While primarily a desktop application suite, it offers cloud-based rendering and is associated with AI model APIs for media processing.
unknownfree tier unknownLow risk - Udioudio.com
Udio is an AI music generation platform that allows users to create high-fidelity audio tracks from text descriptions. It offers both a free tier with daily credits and subscription plans with higher quotas and advanced features (Voice Control, stem extraction).
unknownfree tier · not quantifiedLow risk - Venice.aivenice.ai
Venice AI provides private, unrestricted access to leading AI models (text, image, audio, video) via an OpenAI-compatible API. It emphasizes zero data retention and hardware-verified privacy (TEE). Features a wide catalog of 300+ models including uncensored versions.
models: Meta-Llama-3.1-70B-Instruct, Llama-3.1-8B-Instruct, Meta: Llama 3.1 405B Instruct
unknownfree tier · not quantifiedconnectMedium risk - Void AIvoidai.app
Void AI is an AI API aggregator compatible with the OpenAI SDK. It uses a credit-based billing system where model costs are calculated using multipliers (e.g., GPT-5.1 is 0.75, GPT-4o is 1.25). Offers both flat-rate plans and usage-based top-ups.
models: GPT-5.1, gemini-3-flash-preview, GPT-4o +1
unknownfree tier · not quantifiedconnectMedium risk - Volcenginewww.volcengine.com
Volcengine is ByteDance's enterprise cloud platform. Its Ark (Huoshan Fangzhou) platform provides access to the Doubao (Doubao) model family and other high-performance LLMs. Supports OpenAI-compatible API calls and various billing plans including token-based monthly packages.
models: Qwen-Plus, zai-glm-4.7
unknownfree tier · not quantifiedconnectLow risk - Volcengine Ark Agent Planconsole.volcengine.com
The Volcengine Ark Agent Plan is a subscription-based billing option for the Ark platform, offering monthly token quotas for Doubao and other models. It is designed for developers requiring predictable monthly costs and higher quotas than standard PAYG.
models: Qwen-Plus, zai-glm-4.7
unknownfree tier unknownLow risk - Wafer AIwafer.ai
Wafer AI provides fast, serverless AI model inference and agent infrastructure. It allows developers to configure AI models with an API key from its dashboard. Features the 'Wafer Pass' for predictable model access and specialized 'Wafer' models optimized for speed.
models: DeepSeek Chat, qwen/qwen3.5-397b-a17b, GLM-5.1 +1
unknownfree tier unknownMedium risk - Weights & Biases Inferencewandb.ai
Weights & Biases Weave provides an inference router and observability tools for AI development. It integrates with OpenRouter and other providers to offer a unified, OpenAI-compatible endpoint for model testing and production deployments.
models: gpt-oss-120b, gpt-oss:20b, Qwen3 235B A22B Thinking 2507 +2
unknownfree tier · not quantifiedconnectLow risk - Writerdev.writer.com
Writer is an enterprise generative AI platform featuring its own family of Palmyra models. It provides a consumption-based API for Palmyra X6, X5, and X4 models, optimized for business writing, data security, and precision. It does not natively use OpenAI-compatible request shapes.
models: Palmyra X5
unknownfree tier unknownLow risk - X5Labx5lab.dev
X5Lab provides an OpenAI-compatible API for accessing various large models. It allows users to authenticate using an API key (format: x5-...) via standard OpenAI clients. Access is based on a pay-as-you-go model with a documented dashboard for key management.
models: DeepSeek Chat, Gemini 2.5 Pro, claude-sonnet-5 +1
unknownfree tier unknownconnectMedium risk - Yi (01.AI)01.ai
Yi is the model family by 01.AI, offering high-performance LLMs including Yi-Lightning and Yi-Large. The developer platform provides OpenAI-compatible API access with usage-based billing. It is optimized for both Chinese and English language tasks.
unknownfree tier unknownconnectLow risk - ZeroLimitAIwww.zerolimitai.com
ZeroLimitAI is an AI API aggregator offering free AI chat and a developer API with automatic model routing. It advertises a lifetime access plan for a one-time fee and a free tier for testing. The API is OpenAI-compatible.
models: Llama 4 Scout, Qwen3-235B-A22B, DeepSeek-R1
unknownfree tier · not quantifiedconnectMedium risk - Zylo APIzyloai.net
Zylo API provides a unified OpenAI-compatible gateway to models from Anthropic, OpenAI, Google, and more with zero token markup. Billing is based on base per-token rates deducted from a prepaid balance. A flat platform fee applies when adding credits.
models: Grok 4.20, deepseek-v4-pro, qwen3.7-MAX +3
unknownfree tier · not quantifiedconnectMedium risk
free-credit values are normalized USD estimates from public sources — see methodology · ordering is data (free-credit value), never sponsorship
What is an AI API router?
An AI API router is a service that exposes AI model APIs (often OpenAI- or Claude-compatible endpoints) so you can call many models — from different vendors or free and paid tiers — through one interface, with a single key and one billing surface.
How is the free-credit value calculated?
Free-credit values are normalized estimates in USD from public sources (signup bonuses, monthly quotas and usage allowances). Estimates are flagged on each entry; values we could not quantify are shown honestly as "Unknown" or "Not quantified" instead of invented numbers.
Is the ranking sponsored?
No. Every list on UPROUTER.ONLINE is ordered by data (free-credit value, price, risk rating or community score). There is no pay-to-rank — affiliate and referral links never influence data, scores or sorting.
How do I know a router is safe to use?
Each entry carries a risk rating (low / medium / high / unrated) with the reasoning behind it, a data-confidence score and live status. Risk is an evidence rating from community research — treat it as a starting point and always review the provider terms yourself.