Skip to content

$ 60 entries match — showing 25–48 · sort: Free credits ↓

// the index — expansive, community-regulatedno affiliate ordering · no pay-to-rank · uprouter.online

388 routers & providers · 813 models tracked · $3.8k free credits

  • Official free tierAPI key
    up207msupdated 27m ago
    fireworks.ai
    source · docs.fireworks.ai/serverless/pricing
    overview

    Fireworks AI is a self-serve inference platform: per-token serverless (Standard/Priority/Fast tiers) with $1 in free credits and postpaid billing, fine-tuning (LoRA and full-param, priced per 1M training tokens), a shared-pool Serverless Training API, and on-demand GPUs from $8/hr (H100, from Sep 1 2026) to $20/hr (GB300).

    free quota

    Official pricing page: 'Get started with $1 in free credits' (confirmed 2026-09-16).

    free models
    Llama familyDeepSeek familyQwen family
    see fireworks.ai’s free quota / risk review →
  • Official free tierAPI key
    up94msupdated 26m ago
    hyperbolic.xyz
    source · hyperbolic.xyz
    overview

    Hyperbolic (hyperbolic.xyz) is the API-specific endpoint for Hyperbolic's AI cloud. It provides serverless inference for over 25 open-source models through an OpenAI-compatible /v1 endpoint. The platform supports text, image, and audio generation with pay-as-you-go pricing.

    free quota

    ~$1 trial credit for new users (DeepSeek V3, Llama 3.1 405B, Qwen); pay-as-you-go after it is used (Llama 405B ~$4/M, DeepSeek R1 ~$3/M tokens).

    free models
    DeepSeek V3Llama 3.1 405BQwen familyDeepSeek-R1
    see hyperbolic.xyz’s free quota / risk review →
  • Official free tierAPI key
    up246msupdated 26m ago
    api.studio.nebius.com
    source · nebius.com/pricing
    overview

    Nebius AI Studio is a European-based GPU cloud platform providing managed inference for top open-source LLMs like Llama 3.1 and DeepSeek R1. It leverages its own EU data centers to offer high-performance, privacy-compliant inference with transparent pay-as-you-go pricing tailored for both developers and enterprises.

    free quota

    AI Studio grants new accounts ~$1 one-time trial credit (startup/promo programs can grant far more); Nebius AI Cloud free-trial program suspended as of 2026-07-13; studio models otherwise pay-as-you-go.

    free models
    DeepSeek familyQwen familyLlama-3.1-8B-Instruct
    see api.studio.nebius.com’s free quota / risk review →
  • Official free tierAPI key
    upupdated 13d ago
    console.scaleway.com
    source · scaleway.com/en/docs/generative-apis/faq
    overview

    Scaleway Generative APIs is a serverless, OpenAI-compatible inference service hosted only in Paris and billed per 1M tokens in euros. Published rates run from EUR 0.15/1M input and EUR 0.60/1M output for gpt-oss-120b, and EUR 0.15/EUR 0.35 for mistral-small-3.2-24b-instruct, up to EUR 1.80/EUR 5.50 for glm-5.2; whisper-large-v3 is metered per audio minute at EUR 0.003 with free output tokens, and embedding models are input-only at EUR 0.10/1M. Every new customer gets a free tier of 1,000,000 tokens, with billing starting at token 1,000,001.

    free quota

    1,000,000 free tokens per new customer, stated on the Generative APIs page; valued at the cheapest published input rate of EUR 0.15/1M (mistral-small-3.2-24b-instruct or gpt-oss-120b), which is roughly USD 0.16 - the grant is token-denominated, so actual dollar value varies by model.

    free models
    Mistral familyQwen familyPixtral 12B (2409)Whisper Large v3Gemma 3 27B InstructLlama 3.3 70B Instruct
    see console.scaleway.com’s free quota / risk review →
  • Official free tierNo auth
    unknownupdated 13d ago
    playground.ai.cloudflare.com
    source · playground.ai.cloudflare.com
    overview

    The Cloudflare AI Playground at playground.ai.cloudflare.com is a browser environment for experimenting with, benchmarking and tuning AI models, with a Model mode for testing system prompts and run settings, a Compare mode for side-by-side evaluation of models under identical prompts, and an experimental Generative UI mode. It is the front end for models that Cloudflare hosts and serves, and the free access terms come from Workers AI: the free allocation allows anyone to use a total of 10,000 Neurons per day at no charge, with additional usage on the Workers Paid plan billed at $0.011 per 1,000 Neurons. The model catalog covers 65 models including Llama, Qwen, DeepSeek, Kimi, GLM, Gemma, Mistral, Nemotron, Whisper and FLUX families, and the playground itself states that output is unverified and does not reflect Cloudflare's views.

    free quota

    Free allocation is stated as 10,000 Neurons per day at no charge on the Workers AI pricing page, equivalent to about $0.11/day at the published rate of $0.011 per 1,000 Neurons (the dollar figure is a conversion, the allowance is official).

    free models
    gpt-oss-120bgpt-oss:20bQwen3.8 27BGemma 4 26B A4B Llama 4 ScoutNemotron 3 Super+3
    see playground.ai.cloudflare.com’s free quota / risk review →
  • Official free tierAPI key
    upupdated 13d ago
    developers.cloudflare.com
    source · developers.cloudflare.com/workers-ai/platform/pricing
    overview

    Cloudflare Workers AI serves open-weight models on Cloudflare's edge network, metered in Neurons ($0.011 per 1,000). Every account — including the Free plan — gets 10,000 Neurons per day free, resetting at 00:00 UTC; beyond that, the Paid plan bills usage. The catalog spans LLMs, embeddings, image and audio models and sits next to your Workers apps.

    free quota

    10,000 Neurons/day free on all plans (reset 00:00 UTC) ≈ $0.11/day ≈ $3.30/month at full use; overage $0.011 per 1,000 Neurons on the Paid plan.

    free models
    Gemma familyLlama familyQwen familyMistralLlama Guard 3DeepSeek R1 Distill
    see developers.cloudflare.com’s free quota / risk review →
  • Official free tierAPI key
    up179msupdated 26m ago
    nomic.ai
    source · nomic.ai/pricing
    overview

    Nomic AI provides the Atlas Embedding API and open-weight models like nomic-embed. It offers a monthly free tier for its embedding API and is transitioning to focus on document AI for specialized industries.

    free quota

    Atlas embedding API free tier ~1M embedding tokens/month for evaluation and prototyping; ~$0.10/mo equivalent at Nomic embed pricing (~$0.10/1M tokens); paid beyond that.

    free models
    nomic-embed-text-v1.5nomic-embed-textnomic-embed-vision
    see nomic.ai’s free quota / risk review →
  • Official free tierAPI key
    up312msupdated 26m ago
    cloud.sambanova.ai
    source · cloud.sambanova.ai/plans/pricing
    overview

    SambaNova Cloud serves open-weight models (DeepSeek-V3.1, DeepSeek-V3.2, Meta-Llama-3.3-70B-Instruct, gpt-oss-120b, gemma-4-31B-it, MiniMax-M2.7 and MiniMax-M3) through an OpenAI-compatible endpoint at api.sambanova.ai/v1, priced per 1M tokens from $0.22 input / $0.59 output for gpt-oss-120b up to $3 / $4.50 for DeepSeek-V3.1 and V3.2. The free tier is defined by the absence of a payment method: 20 requests per minute, 20 requests per day and 200,000 tokens per day per model, on the five models listed in the free-tier tables. Linking a card moves the account to the Developer tier, which raises per-model limits and caps the account at 20M tokens per day across all models.

    free quota

    Free tier is a recurring daily allowance, not a USD credit: 200,000 tokens/day (20 RPM, 20 RPD) per the official rate-limits docs; 200k tokens at the cheapest published input rate of $0.22/1M (gpt-oss-120b) is about $0.04/day, while the same 200k at DeepSeek-V3.1's $3/1M input rate would be about $0.60/day.

    free models
    Llama 3.1 405BDeepSeek-V3.1/V3.2gpt-oss-120bLlama 3.3 70B InstructQwen2.5 72B Instruct
    see cloud.sambanova.ai’s free quota / risk review →
  • Official free tierAPI key
    upupdated 13d ago
    bailian.console.alibabacloud.com
    source · alibabacloud.com/help/en/model-studio/model-pricing
    overview

    Alibaba Cloud Model Studio (Bailian, 百炼) is Alibaba Cloud's managed model platform, hosting Qwen and third-party models across text, image, audio and video with pay-as-you-go billing per million tokens in CNY. The cn-beijing list rates on the official price page include qwen-long at 0.5 CNY per million input tokens and 2 CNY per million output tokens, qwq-plus at 1.6 CNY input and 4 CNY output, and deepseek-v3.2-exp at 2 CNY input and 3 CNY output, while international-deployment rates are materially higher (qwq-plus international: 5.871 CNY input, 17.614 CNY output). The price page states it shows list prices only ('原价') and that promotional pricing lives in the console, and pricing is tiered so a single request's token volume sets the unit price for that whole request.

    free quota

    Token allowance, not cash: first activation of Bailian grants a per-model free inference quota of typically 1,000,000 tokens (input and output combined, cn-beijing only, valid 90 days) per help.aliyun.com/zh/model-studio/new-free-quota; no USD-convertible value is published (checked 2026-09-16).

    free models
    — none tracked
    see bailian.console.alibabacloud.com’s free quota / risk review →
  • Official free tierAPI key
    unknownupdated 23d ago
    api.airforce
    source · api.airforce
    overview

    Api.airforce is a unified AI model aggregator reselling access to 100+ models from Anthropic, OpenAI, and xAI. It offers a large selection of free models (including Grok-3 and Claude 3.7) alongside paid subscription tiers for higher throughput.

    free quota

    55 free tier models including Grok-3, Claude 3.7, Qwen3, Kimi-K2, Gemini 2.5 Flash, DeepSeek-V3

    free models
    Gemini 2.5 FlashDeepSeek ChatGrok 4.20Qwen3 Maxkimi-k2claude-sonnet-4-6
    see api.airforce’s free quota / risk review →
  • Official free tierAPI key
    upupdated 13d ago
    app.baseten.co
    source · baseten.co
    overview

    Baseten offers Model APIs, hosted open-weight models behind an OpenAI Chat Completions endpoint at inference.baseten.co with a beta Anthropic Messages endpoint, alongside dedicated GPU deployments and training billed per minute. Model API rates range from $0.10/1M input and $0.50/1M output for GPT OSS 120B to $3.00/1M input and $15.00/1M output for Kimi K3, with cached input billed at roughly a tenth of the input rate. The Basic plan is $0 per month pay-as-you-go, and the pricing FAQ states new accounts come with credits but publishes no amount.

    free quota

    The official pricing FAQ states that new Baseten accounts come with credits to experiment with deployments, but no dollar amount, expiry or conditions are published (checked 2026-09-16).

    free models
    Any supported model in their library (billed by compute; no fixed free-model list)
    see app.baseten.co’s free quota / risk review →
  • Official free tierAPI key
    up77msupdated 26m ago
    api.bfl.ai
    source · bfl.ai/pricing
    overview

    Black Forest Labs runs the hosted FLUX model API with a credit system where 1 credit equals USD $0.01, charged per image for stills and per second for video. FLUX 3 video costs $0.17/s at HD, $0.29/s FHD, $0.40/s QHD and $0.80/s UHD for text- and image-to-video, while video-to-video (continuation, editing, restyling) runs $0.41-$0.95/s; a draft tier renders HD at $0.06/s. FLUX.2 images start at $0.014 for [klein] 4B and rise to $0.07 for [max], with editing at $0.045 for [pro]. Clips are capped at 20 seconds and synchronized audio is included at no extra charge.

    free quota

    No free credit grant is stated on official pages; docs.bfl.ai/quick_start/get_started instructs users to add credits (suggesting $10-20 to start) with no signup bonus (checked 2026-09-16).

    free models
    — none tracked
    see api.bfl.ai’s free quota / risk review →
  • Official free tierAPI key
    down6003msupdated 27m ago
    api.bfl.ml
    source · docs.bfl.ml/quick_start/pricing
    overview

    BFL ML (api.bfl.ml) is the technical endpoint for Black Forest Labs' FLUX model API. It shares the same credit-based billing as the main bfl.ai domain, providing high-performance access to diffusion models for production workloads.

    free quota

    Black Forest Labs FLUX API is prepaid-credit only: 1 credit = $0.01, docs instruct 'create account, add credits' and to buy more on HTTP 402. No free signup credits found in official docs.

    free models
    — none tracked
    see api.bfl.ml’s free quota / risk review →
  • Official free tierNo auth
    unknownupdated 15d ago
    felo.ai
    source · felo.ai
    overview

    Felo (felo.ai) is a free, no-signup multilingual AI search and creation platform for chat, web search, and content generation. It exposes frontier models (e.g., GPT-6 Astra) and offers a paid Pro tier for heavier usage.

    free quota

    Free no-signup tier; Pro subscription pricing not quantified on fetched pages.

    free models
    GPT-6 Astra
    see felo.ai’s free quota / risk review →
  • Official free tierAPI key
    unknownupdated 13d ago
    freeinference.org
    source · freeinference.org
    overview

    FreeInference (freeinference.org) is an OpenAI- and Anthropic-compatible inference API offering frontier open models free of charge to the research community, with no credit card required at signup. It is presented as built at Harvard SEAS (MadSys Lab) and lists NVIDIA and Harvard SEAS as sponsors. The quota is described only as a "generous quota for research and prototyping"; no numeric limit is published, and the /pricing and /limits paths both return 404. The site carries a banner warning that GLM, Minimax and Kimi will be discontinued on Oct 1, and on Aug 11, 2026 it announced that new account onboarding is paused because the service is at capacity.

    free quota

    No free tier stated on official pages (checked 2026-09-16): freeinference.org offers a free account with a "Generous quota for research and prototyping" but publishes no dollar or token figure, and freeinference.org/pricing returns 404.

    free models
    Qwen3 CoderMiniMax-M3DeepSeek V4 FlashGLM-5.1
    see freeinference.org’s free quota / risk review →
  • Official free tierAPI key
    unknownupdated 23d ago
    friendli.ai
    source · friendli.ai/pricing
    overview

    FriendliAI is a high-performance AI inference platform specializing in serving frontier open-weight models like GLM, DeepSeek, and Llama with industry-leading speed. It offers serverless Model APIs, dedicated GPU endpoints, and on-premise containers.

    free quota

    Free tier for serverless inference — no credit card required

    free models
    Meta-Llama-3.1-70B-InstructLlama-3.1-8B-InstructMistral NemoGemma 4 31B
    see friendli.ai’s free quota / risk review →
  • Official free tierAPI key
    unknownupdated 13d ago
    open.bigmodel.cn
    source · open.bigmodel.cn/pricing
    overview

    Zhipu AI's GLM platform (bigmodel.cn, also served at open.bigmodel.cn) sells model access per token in CNY and bundles agentic coding into a separate GLM Coding Plan. The official pricing page lists GLM-5.3 at 8 yuan per million input tokens, 2 yuan per million cached tokens and 28 yuan per million output tokens with a 1M-token context, and GLM-5.3-Flash at 0.8 / 0.23 / 2.8 yuan. New accounts receive 20 million free tokens on registration, and the text model GLM-4.7-Flash (200K context) and vision model GLM-4.6V-Flash are listed as free for input, output and cache. The GLM Coding Plan shows 118 yuan/month (Lite), 538 yuan/month (Pro) and 1078 yuan/month (Max), with discounted continuous-subscription prices of 94.4 / 430.4 / 862.4 yuan and quarterly billing at 8-fold and annual at 7-fold discounts.

    free quota

    Free grant is 20,000,000 GLM tokens for new registrations (tokens, not currency), stated on bigmodel.cn/pricing and bigmodel.cn/claude-code; GLM-4.7-Flash and GLM-4.6V-Flash are additionally free per token.

    free models
    glm-5.3-flashglm-5.3GLM 5 Turbo
    see open.bigmodel.cn’s free quota / risk review →
  • Official free tierAPI key
    up54msupdated 26m ago
    models.github.ai
    source · github-models
    overview

    GitHub Models was GitHub's hosted model playground, model catalog, inference API and bring-your-own-key service, and GitHub previously presented it as separate from GitHub Copilot. GitHub's official documentation now states that as of July 30, 2026, GitHub Models has been fully retired: the playground, model catalog, inference API and BYOK are no longer available to any customer. The same page directs users needing model access to Azure AI Foundry and users wanting AI workflows on GitHub to GitHub Copilot.

    free quota

    No free tier stated on official pages; GitHub Models was fully retired on July 30, 2026 with no playground, catalog, inference API or BYOK access remaining (checked 2026-09-16).

    free models
    Llama familyGPT-4oGPT-4o-mini
    see models.github.ai’s free quota / risk review →
  • Official free tierAPI key
    up89msupdated 26m ago
    hyperbolic.ai
    source · hyperbolic.ai
    overview

    Hyperbolic (hyperbolic.ai) is an AI cloud that sells on-demand, reserved and private GPU capacity and, separately, a serverless inference API described as OpenAI-compatible at base URL https://api.hyperbolic.xyz/v1 with 25+ open-source models. Its official rate-limit table defines a Basic tier of 60 requests per minute available to a free account with a $0 minimum deposit, and a Pro tier of 600 requests per minute that unlocks automatically once $5 or more is deposited. Serverless inference is pay-as-you-go: text generation runs from $0.10 per 1M tokens for small models, $0.20-$0.40 for 32B-72B models and $0.30-$4.00 for 120B-480B models, with images at a $0.01 base rate per 1024x1024 image at 25 steps and audio at $5.00 per 1M characters.

    free quota

    No free tier stated on official pages (checked 2026-09-16): Hyperbolic documents a free "Basic" tier as 60 requests/minute with a Free account, but no dollar credit or trial allowance is published on the pricing or billing pages.

    free models
    DeepSeek V3Llama 3.1 405B (BF16, currently the only public 405B Base)Qwen familyFLUX / Stable Diffusion (image)Melo TTS (audio)
    see hyperbolic.ai’s free quota / risk review →
  • Official free tierAPI key
    upupdated 13d ago
    inference.net
    source · inference.net
    overview

    Inference.net is a unified LLM gateway, observability and fine-tuning platform rather than an image or video vendor. Its pay-as-you-go tier is $0 + usage and includes 1M gateway requests and 1M tracing spans per month, 14-day data retention, one seat and a 30 req/min rate limit. Growth is $250/month with a $50 one-time opening credit, 50M gateway requests and tracing spans per month, unlimited seats and a 250 req/min limit. The catalog lists 60 models from OpenAI, Anthropic, Google and open-source labs billed per token, from Schematron V2 Turbo at $0.03/$0.15 to Claude Opus 5 and GPT-6 Astra at $5.00-$10.00 input and $25.00-$50.00 output per 1M tokens.

    free quota

    No free dollar credit: the $0 pay-as-you-go tier includes 1M gateway requests and 1M tracing spans per month, while the $50 one-time opening credit is listed only on the $250/month Growth plan (checked 2026-09-16).

    free models
    Meta-Llama-3.1-70B-Instruct
    see inference.net’s free quota / risk review →
  • Official free tierWeb cookie
    unknownupdated 13d ago
    kimi.ai
    source · kimi.com/en
    overview

    Kimi is Moonshot AI's consumer assistant, used at kimi.com or through the Kimi app, with K2.6, K3 and K3 Cluster model options. Kimi's official help centre documents four paid tiers - Andante at 49 yuan/month, Moderato at 99 yuan/month, Allegretto at 199 yuan/month and Allegro at 699 yuan/month - and states that annual billing saves up to 1,680 yuan. The help centre also states that K2.6, K3 and K3 Cluster conversations are all billed against a single shared quota pool that refreshes monthly, and that Kimi Code carries an additional 5-hour/weekly limit. No numeric free-tier allowance is published on any official Kimi page fetched.

    free quota

    No free tier stated on official pages (checked 2026-09-16).

    free models
    MoonshotAI: Kimi K2.5MoonshotAI: Kimi K2 0711MoonshotAI: Kimi K2 0905MoonshotAI: Kimi K2 ThinkingMoonshotAI: Kimi K2.7 CodeMoonshotAI: Kimi K2.6+5
    see kimi.ai’s free quota / risk review →
  • Official free tierAPI key
    up1961msupdated 26m ago
    modelscope.cn
    source · modelscope.ai/docs/model-service/API-Inference/limits
    overview

    ModelScope (魔搭, modelscope.cn) is an open-source model and dataset community founded in June 2022 by the Institute for Intelligent Computing together with the CCF Open Source Development Committee, organised around 'Model-as-a-Service'. The platform publishes models, datasets, Studios (application display space), MCP and Skills listings, and maintains the ModelScope, Swift, EvalScope and ModelScope-Agent frameworks. Its stated free offering is platform credits: 'Sign up / Log in to get 200 Magicubes' plus 'Link Alibaba Cloud account, get 50 more daily', with Studios described as 'a free and flexible AI application display space'.

    free quota

    No cash value stated: the official home page says 'Sign up / Log in to get 200 Magicubes' and 'Link Alibaba Cloud account, get 50 more daily' (modelscope.cn, checked 2026-09-16); Magicubes have no published CNY or USD rate.

    free models
    Qwen family (most stable)DeepSeek-V3.1KimiGLM family (per the current console availability)DeepSeek-R1
    see modelscope.cn’s free quota / risk review →
  • Official free tierAPI key
    unknownupdated 23d ago
    morphllm.com
    source · morphllm.com/pricing
    overview

    Morph (morphllm.com) is an AI provider focusing on agentic coding models. Its API offers high-speed code editing and search via proprietary models (Compact, Reflexes) and optimized frontier models (GLM, Kimi), with a limited free tier of 200 requests per month.

    free quota

    Free tier: 250K credits/month, $0

    free models
    deepseek-v4-flash-0731Qwen3.5-122B-A10Bglm-5.3Morph V3 LargeMorph V3 Fastkimi-k3
    see morphllm.com’s free quota / risk review →
  • Official free tierAPI key
    up13msupdated 27m ago
    integrate.api.nvidia.com
    source · build.nvidia.com
    overview

    NVIDIA NIM's hosted API (build.nvidia.com / integrate.api.nvidia.com) serves a broad catalog of open-weight models — Nemotron, Llama, Qwen, DeepSeek, GLM and more — in an OpenAI-compatible shape, with a permanent free tier: API key at sign-up, no card, 40 requests/minute by default (200 RPM granted on request, per NVIDIA's developer forums).

    free quota

    Free tier is rate-limited (40 RPM default, 200 on request) rather than credit-based — no USD value to quantify.

    free models
    Nemotron 3 Nano 30B A3BNemotron 3.5 LightningNemotron 3 SuperNemotron 3 Ultranvidia/llama-3.1-nemotron-ultra-253b-v1gpt-oss-120b+6
    see integrate.api.nvidia.com’s free quota / risk review →