$ 60 entries match — showing 25–48 · sort: Free credits ↓
388 routers & providers · 813 models tracked · $3.8k free credits
- Official free tierAPI keyup207msupdated 27m agofireworks.aisource · docs.fireworks.ai/serverless/pricingoverview
Fireworks AI is a self-serve inference platform: per-token serverless (Standard/Priority/Fast tiers) with $1 in free credits and postpaid billing, fine-tuning (LoRA and full-param, priced per 1M training tokens), a shared-pool Serverless Training API, and on-demand GPUs from $8/hr (H100, from Sep 1 2026) to $20/hr (GB300).
free quotaOfficial pricing page: 'Get started with $1 in free credits' (confirmed 2026-09-16).
free modelsLlama familyDeepSeek familyQwen familysee fireworks.ai’s free quota / risk review → - Official free tierAPI keyup94msupdated 26m agohyperbolic.xyzsource · hyperbolic.xyzoverview
Hyperbolic (hyperbolic.xyz) is the API-specific endpoint for Hyperbolic's AI cloud. It provides serverless inference for over 25 open-source models through an OpenAI-compatible /v1 endpoint. The platform supports text, image, and audio generation with pay-as-you-go pricing.
free quota~$1 trial credit for new users (DeepSeek V3, Llama 3.1 405B, Qwen); pay-as-you-go after it is used (Llama 405B ~$4/M, DeepSeek R1 ~$3/M tokens).
free modelsDeepSeek V3Llama 3.1 405BQwen familyDeepSeek-R1see hyperbolic.xyz’s free quota / risk review → - Official free tierAPI keyup246msupdated 26m agoapi.studio.nebius.comsource · nebius.com/pricingoverview
Nebius AI Studio is a European-based GPU cloud platform providing managed inference for top open-source LLMs like Llama 3.1 and DeepSeek R1. It leverages its own EU data centers to offer high-performance, privacy-compliant inference with transparent pay-as-you-go pricing tailored for both developers and enterprises.
free quotaAI Studio grants new accounts ~$1 one-time trial credit (startup/promo programs can grant far more); Nebius AI Cloud free-trial program suspended as of 2026-07-13; studio models otherwise pay-as-you-go.
free modelsDeepSeek familyQwen familyLlama-3.1-8B-Instructsee api.studio.nebius.com’s free quota / risk review → - Official free tierAPI keyupupdated 13d agoconsole.scaleway.comsource · scaleway.com/en/docs/generative-apis/faqoverview
Scaleway Generative APIs is a serverless, OpenAI-compatible inference service hosted only in Paris and billed per 1M tokens in euros. Published rates run from EUR 0.15/1M input and EUR 0.60/1M output for gpt-oss-120b, and EUR 0.15/EUR 0.35 for mistral-small-3.2-24b-instruct, up to EUR 1.80/EUR 5.50 for glm-5.2; whisper-large-v3 is metered per audio minute at EUR 0.003 with free output tokens, and embedding models are input-only at EUR 0.10/1M. Every new customer gets a free tier of 1,000,000 tokens, with billing starting at token 1,000,001.
free quota1,000,000 free tokens per new customer, stated on the Generative APIs page; valued at the cheapest published input rate of EUR 0.15/1M (mistral-small-3.2-24b-instruct or gpt-oss-120b), which is roughly USD 0.16 - the grant is token-denominated, so actual dollar value varies by model.
free modelsMistral familyQwen familyPixtral 12B (2409)Whisper Large v3Gemma 3 27B InstructLlama 3.3 70B Instructsee console.scaleway.com’s free quota / risk review → - Official free tierNo authunknownupdated 13d agoplayground.ai.cloudflare.comsource · playground.ai.cloudflare.comoverview
The Cloudflare AI Playground at playground.ai.cloudflare.com is a browser environment for experimenting with, benchmarking and tuning AI models, with a Model mode for testing system prompts and run settings, a Compare mode for side-by-side evaluation of models under identical prompts, and an experimental Generative UI mode. It is the front end for models that Cloudflare hosts and serves, and the free access terms come from Workers AI: the free allocation allows anyone to use a total of 10,000 Neurons per day at no charge, with additional usage on the Workers Paid plan billed at $0.011 per 1,000 Neurons. The model catalog covers 65 models including Llama, Qwen, DeepSeek, Kimi, GLM, Gemma, Mistral, Nemotron, Whisper and FLUX families, and the playground itself states that output is unverified and does not reflect Cloudflare's views.
free quotaFree allocation is stated as 10,000 Neurons per day at no charge on the Workers AI pricing page, equivalent to about $0.11/day at the published rate of $0.011 per 1,000 Neurons (the dollar figure is a conversion, the allowance is official).
free modelsgpt-oss-120bgpt-oss:20bQwen3.8 27BGemma 4 26B A4B Llama 4 ScoutNemotron 3 Super+3see playground.ai.cloudflare.com’s free quota / risk review → - Official free tierAPI keyupupdated 13d agodevelopers.cloudflare.comsource · developers.cloudflare.com/workers-ai/platform/pricingoverview
Cloudflare Workers AI serves open-weight models on Cloudflare's edge network, metered in Neurons ($0.011 per 1,000). Every account — including the Free plan — gets 10,000 Neurons per day free, resetting at 00:00 UTC; beyond that, the Paid plan bills usage. The catalog spans LLMs, embeddings, image and audio models and sits next to your Workers apps.
free quota10,000 Neurons/day free on all plans (reset 00:00 UTC) ≈ $0.11/day ≈ $3.30/month at full use; overage $0.011 per 1,000 Neurons on the Paid plan.
free modelsGemma familyLlama familyQwen familyMistralLlama Guard 3DeepSeek R1 Distillsee developers.cloudflare.com’s free quota / risk review → - Official free tierAPI keyup179msupdated 26m agonomic.aisource · nomic.ai/pricingoverview
Nomic AI provides the Atlas Embedding API and open-weight models like nomic-embed. It offers a monthly free tier for its embedding API and is transitioning to focus on document AI for specialized industries.
free quotaAtlas embedding API free tier ~1M embedding tokens/month for evaluation and prototyping; ~$0.10/mo equivalent at Nomic embed pricing (~$0.10/1M tokens); paid beyond that.
free modelsnomic-embed-text-v1.5nomic-embed-textnomic-embed-visionsee nomic.ai’s free quota / risk review → - Official free tierAPI keyup312msupdated 26m agocloud.sambanova.aisource · cloud.sambanova.ai/plans/pricingoverview
SambaNova Cloud serves open-weight models (DeepSeek-V3.1, DeepSeek-V3.2, Meta-Llama-3.3-70B-Instruct, gpt-oss-120b, gemma-4-31B-it, MiniMax-M2.7 and MiniMax-M3) through an OpenAI-compatible endpoint at api.sambanova.ai/v1, priced per 1M tokens from $0.22 input / $0.59 output for gpt-oss-120b up to $3 / $4.50 for DeepSeek-V3.1 and V3.2. The free tier is defined by the absence of a payment method: 20 requests per minute, 20 requests per day and 200,000 tokens per day per model, on the five models listed in the free-tier tables. Linking a card moves the account to the Developer tier, which raises per-model limits and caps the account at 20M tokens per day across all models.
free quotaFree tier is a recurring daily allowance, not a USD credit: 200,000 tokens/day (20 RPM, 20 RPD) per the official rate-limits docs; 200k tokens at the cheapest published input rate of $0.22/1M (gpt-oss-120b) is about $0.04/day, while the same 200k at DeepSeek-V3.1's $3/1M input rate would be about $0.60/day.
free modelsLlama 3.1 405BDeepSeek-V3.1/V3.2gpt-oss-120bLlama 3.3 70B InstructQwen2.5 72B Instructsee cloud.sambanova.ai’s free quota / risk review → - Official free tierAPI keyupupdated 13d agobailian.console.alibabacloud.comsource · alibabacloud.com/help/en/model-studio/model-pricingoverview
Alibaba Cloud Model Studio (Bailian, 百炼) is Alibaba Cloud's managed model platform, hosting Qwen and third-party models across text, image, audio and video with pay-as-you-go billing per million tokens in CNY. The cn-beijing list rates on the official price page include qwen-long at 0.5 CNY per million input tokens and 2 CNY per million output tokens, qwq-plus at 1.6 CNY input and 4 CNY output, and deepseek-v3.2-exp at 2 CNY input and 3 CNY output, while international-deployment rates are materially higher (qwq-plus international: 5.871 CNY input, 17.614 CNY output). The price page states it shows list prices only ('原价') and that promotional pricing lives in the console, and pricing is tiered so a single request's token volume sets the unit price for that whole request.
free quotaToken allowance, not cash: first activation of Bailian grants a per-model free inference quota of typically 1,000,000 tokens (input and output combined, cn-beijing only, valid 90 days) per help.aliyun.com/zh/model-studio/new-free-quota; no USD-convertible value is published (checked 2026-09-16).
free models— none trackedsee bailian.console.alibabacloud.com’s free quota / risk review → - Official free tierAPI keyunknownupdated 23d agoapi.airforcesource · api.airforceoverview
Api.airforce is a unified AI model aggregator reselling access to 100+ models from Anthropic, OpenAI, and xAI. It offers a large selection of free models (including Grok-3 and Claude 3.7) alongside paid subscription tiers for higher throughput.
free quota55 free tier models including Grok-3, Claude 3.7, Qwen3, Kimi-K2, Gemini 2.5 Flash, DeepSeek-V3
free modelsGemini 2.5 FlashDeepSeek ChatGrok 4.20Qwen3 Maxkimi-k2claude-sonnet-4-6see api.airforce’s free quota / risk review → - Official free tierAPI keyupupdated 13d agoapp.baseten.cosource · baseten.cooverview
Baseten offers Model APIs, hosted open-weight models behind an OpenAI Chat Completions endpoint at inference.baseten.co with a beta Anthropic Messages endpoint, alongside dedicated GPU deployments and training billed per minute. Model API rates range from $0.10/1M input and $0.50/1M output for GPT OSS 120B to $3.00/1M input and $15.00/1M output for Kimi K3, with cached input billed at roughly a tenth of the input rate. The Basic plan is $0 per month pay-as-you-go, and the pricing FAQ states new accounts come with credits but publishes no amount.
free quotaThe official pricing FAQ states that new Baseten accounts come with credits to experiment with deployments, but no dollar amount, expiry or conditions are published (checked 2026-09-16).
free modelsAny supported model in their library (billed by compute; no fixed free-model list)see app.baseten.co’s free quota / risk review → - Official free tierAPI keyup77msupdated 26m agoapi.bfl.aisource · bfl.ai/pricingoverview
Black Forest Labs runs the hosted FLUX model API with a credit system where 1 credit equals USD $0.01, charged per image for stills and per second for video. FLUX 3 video costs $0.17/s at HD, $0.29/s FHD, $0.40/s QHD and $0.80/s UHD for text- and image-to-video, while video-to-video (continuation, editing, restyling) runs $0.41-$0.95/s; a draft tier renders HD at $0.06/s. FLUX.2 images start at $0.014 for [klein] 4B and rise to $0.07 for [max], with editing at $0.045 for [pro]. Clips are capped at 20 seconds and synchronized audio is included at no extra charge.
free quotaNo free credit grant is stated on official pages; docs.bfl.ai/quick_start/get_started instructs users to add credits (suggesting $10-20 to start) with no signup bonus (checked 2026-09-16).
free models— none trackedsee api.bfl.ai’s free quota / risk review → - Official free tierAPI keydown6003msupdated 27m agoapi.bfl.mlsource · docs.bfl.ml/quick_start/pricingoverview
BFL ML (api.bfl.ml) is the technical endpoint for Black Forest Labs' FLUX model API. It shares the same credit-based billing as the main bfl.ai domain, providing high-performance access to diffusion models for production workloads.
free quotaBlack Forest Labs FLUX API is prepaid-credit only: 1 credit = $0.01, docs instruct 'create account, add credits' and to buy more on HTTP 402. No free signup credits found in official docs.
free models— none trackedsee api.bfl.ml’s free quota / risk review → - Official free tierNo authunknownupdated 15d agofelo.aisource · felo.aioverview
Felo (felo.ai) is a free, no-signup multilingual AI search and creation platform for chat, web search, and content generation. It exposes frontier models (e.g., GPT-6 Astra) and offers a paid Pro tier for heavier usage.
free quotaFree no-signup tier; Pro subscription pricing not quantified on fetched pages.
free modelsGPT-6 Astrasee felo.ai’s free quota / risk review → - Official free tierAPI keyunknownupdated 13d agofreeinference.orgsource · freeinference.orgoverview
FreeInference (freeinference.org) is an OpenAI- and Anthropic-compatible inference API offering frontier open models free of charge to the research community, with no credit card required at signup. It is presented as built at Harvard SEAS (MadSys Lab) and lists NVIDIA and Harvard SEAS as sponsors. The quota is described only as a "generous quota for research and prototyping"; no numeric limit is published, and the /pricing and /limits paths both return 404. The site carries a banner warning that GLM, Minimax and Kimi will be discontinued on Oct 1, and on Aug 11, 2026 it announced that new account onboarding is paused because the service is at capacity.
free quotaNo free tier stated on official pages (checked 2026-09-16): freeinference.org offers a free account with a "Generous quota for research and prototyping" but publishes no dollar or token figure, and freeinference.org/pricing returns 404.
free modelsQwen3 CoderMiniMax-M3DeepSeek V4 FlashGLM-5.1see freeinference.org’s free quota / risk review → - Official free tierAPI keyunknownupdated 23d agofriendli.aisource · friendli.ai/pricingoverview
FriendliAI is a high-performance AI inference platform specializing in serving frontier open-weight models like GLM, DeepSeek, and Llama with industry-leading speed. It offers serverless Model APIs, dedicated GPU endpoints, and on-premise containers.
free quotaFree tier for serverless inference — no credit card required
free modelsMeta-Llama-3.1-70B-InstructLlama-3.1-8B-InstructMistral NemoGemma 4 31Bsee friendli.ai’s free quota / risk review → - Official free tierAPI keyunknownupdated 13d agoopen.bigmodel.cnsource · open.bigmodel.cn/pricingoverview
Zhipu AI's GLM platform (bigmodel.cn, also served at open.bigmodel.cn) sells model access per token in CNY and bundles agentic coding into a separate GLM Coding Plan. The official pricing page lists GLM-5.3 at 8 yuan per million input tokens, 2 yuan per million cached tokens and 28 yuan per million output tokens with a 1M-token context, and GLM-5.3-Flash at 0.8 / 0.23 / 2.8 yuan. New accounts receive 20 million free tokens on registration, and the text model GLM-4.7-Flash (200K context) and vision model GLM-4.6V-Flash are listed as free for input, output and cache. The GLM Coding Plan shows 118 yuan/month (Lite), 538 yuan/month (Pro) and 1078 yuan/month (Max), with discounted continuous-subscription prices of 94.4 / 430.4 / 862.4 yuan and quarterly billing at 8-fold and annual at 7-fold discounts.
free quotaFree grant is 20,000,000 GLM tokens for new registrations (tokens, not currency), stated on bigmodel.cn/pricing and bigmodel.cn/claude-code; GLM-4.7-Flash and GLM-4.6V-Flash are additionally free per token.
free modelsglm-5.3-flashglm-5.3GLM 5 Turbosee open.bigmodel.cn’s free quota / risk review → - Official free tierAPI keyup54msupdated 26m agomodels.github.aisource · github-modelsoverview
GitHub Models was GitHub's hosted model playground, model catalog, inference API and bring-your-own-key service, and GitHub previously presented it as separate from GitHub Copilot. GitHub's official documentation now states that as of July 30, 2026, GitHub Models has been fully retired: the playground, model catalog, inference API and BYOK are no longer available to any customer. The same page directs users needing model access to Azure AI Foundry and users wanting AI workflows on GitHub to GitHub Copilot.
free quotaNo free tier stated on official pages; GitHub Models was fully retired on July 30, 2026 with no playground, catalog, inference API or BYOK access remaining (checked 2026-09-16).
free modelsLlama familyGPT-4oGPT-4o-minisee models.github.ai’s free quota / risk review → - Official free tierAPI keyup89msupdated 26m agohyperbolic.aisource · hyperbolic.aioverview
Hyperbolic (hyperbolic.ai) is an AI cloud that sells on-demand, reserved and private GPU capacity and, separately, a serverless inference API described as OpenAI-compatible at base URL https://api.hyperbolic.xyz/v1 with 25+ open-source models. Its official rate-limit table defines a Basic tier of 60 requests per minute available to a free account with a $0 minimum deposit, and a Pro tier of 600 requests per minute that unlocks automatically once $5 or more is deposited. Serverless inference is pay-as-you-go: text generation runs from $0.10 per 1M tokens for small models, $0.20-$0.40 for 32B-72B models and $0.30-$4.00 for 120B-480B models, with images at a $0.01 base rate per 1024x1024 image at 25 steps and audio at $5.00 per 1M characters.
free quotaNo free tier stated on official pages (checked 2026-09-16): Hyperbolic documents a free "Basic" tier as 60 requests/minute with a Free account, but no dollar credit or trial allowance is published on the pricing or billing pages.
free modelsDeepSeek V3Llama 3.1 405B (BF16, currently the only public 405B Base)Qwen familyFLUX / Stable Diffusion (image)Melo TTS (audio)see hyperbolic.ai’s free quota / risk review → - Official free tierAPI keyupupdated 13d agoinference.netsource · inference.netoverview
Inference.net is a unified LLM gateway, observability and fine-tuning platform rather than an image or video vendor. Its pay-as-you-go tier is $0 + usage and includes 1M gateway requests and 1M tracing spans per month, 14-day data retention, one seat and a 30 req/min rate limit. Growth is $250/month with a $50 one-time opening credit, 50M gateway requests and tracing spans per month, unlimited seats and a 250 req/min limit. The catalog lists 60 models from OpenAI, Anthropic, Google and open-source labs billed per token, from Schematron V2 Turbo at $0.03/$0.15 to Claude Opus 5 and GPT-6 Astra at $5.00-$10.00 input and $25.00-$50.00 output per 1M tokens.
free quotaNo free dollar credit: the $0 pay-as-you-go tier includes 1M gateway requests and 1M tracing spans per month, while the $50 one-time opening credit is listed only on the $250/month Growth plan (checked 2026-09-16).
free modelsMeta-Llama-3.1-70B-Instructsee inference.net’s free quota / risk review → - Official free tierWeb cookieunknownupdated 13d agokimi.aisource · kimi.com/enoverview
Kimi is Moonshot AI's consumer assistant, used at kimi.com or through the Kimi app, with K2.6, K3 and K3 Cluster model options. Kimi's official help centre documents four paid tiers - Andante at 49 yuan/month, Moderato at 99 yuan/month, Allegretto at 199 yuan/month and Allegro at 699 yuan/month - and states that annual billing saves up to 1,680 yuan. The help centre also states that K2.6, K3 and K3 Cluster conversations are all billed against a single shared quota pool that refreshes monthly, and that Kimi Code carries an additional 5-hour/weekly limit. No numeric free-tier allowance is published on any official Kimi page fetched.
free quotaNo free tier stated on official pages (checked 2026-09-16).
free modelsMoonshotAI: Kimi K2.5MoonshotAI: Kimi K2 0711MoonshotAI: Kimi K2 0905MoonshotAI: Kimi K2 ThinkingMoonshotAI: Kimi K2.7 CodeMoonshotAI: Kimi K2.6+5see kimi.ai’s free quota / risk review → - Official free tierAPI keyup1961msupdated 26m agomodelscope.cnsource · modelscope.ai/docs/model-service/API-Inference/limitsoverview
ModelScope (魔搭, modelscope.cn) is an open-source model and dataset community founded in June 2022 by the Institute for Intelligent Computing together with the CCF Open Source Development Committee, organised around 'Model-as-a-Service'. The platform publishes models, datasets, Studios (application display space), MCP and Skills listings, and maintains the ModelScope, Swift, EvalScope and ModelScope-Agent frameworks. Its stated free offering is platform credits: 'Sign up / Log in to get 200 Magicubes' plus 'Link Alibaba Cloud account, get 50 more daily', with Studios described as 'a free and flexible AI application display space'.
free quotaNo cash value stated: the official home page says 'Sign up / Log in to get 200 Magicubes' and 'Link Alibaba Cloud account, get 50 more daily' (modelscope.cn, checked 2026-09-16); Magicubes have no published CNY or USD rate.
free modelsQwen family (most stable)DeepSeek-V3.1KimiGLM family (per the current console availability)DeepSeek-R1see modelscope.cn’s free quota / risk review → - Official free tierAPI keyunknownupdated 23d agomorphllm.comsource · morphllm.com/pricingoverview
Morph (morphllm.com) is an AI provider focusing on agentic coding models. Its API offers high-speed code editing and search via proprietary models (Compact, Reflexes) and optimized frontier models (GLM, Kimi), with a limited free tier of 200 requests per month.
free quotaFree tier: 250K credits/month, $0
free modelsdeepseek-v4-flash-0731Qwen3.5-122B-A10Bglm-5.3Morph V3 LargeMorph V3 Fastkimi-k3see morphllm.com’s free quota / risk review → - Official free tierAPI keyup13msupdated 27m agointegrate.api.nvidia.comsource · build.nvidia.comoverview
NVIDIA NIM's hosted API (build.nvidia.com / integrate.api.nvidia.com) serves a broad catalog of open-weight models — Nemotron, Llama, Qwen, DeepSeek, GLM and more — in an OpenAI-compatible shape, with a permanent free tier: API key at sign-up, no card, 40 requests/minute by default (200 RPM granted on request, per NVIDIA's developer forums).
free quotaFree tier is rate-limited (40 RPM default, 200 on request) rather than credit-based — no USD value to quantify.
free modelsNemotron 3 Nano 30B A3BNemotron 3.5 LightningNemotron 3 SuperNemotron 3 Ultranvidia/llama-3.1-nemotron-ultra-253b-v1gpt-oss-120b+6see integrate.api.nvidia.com’s free quota / risk review →