Skip to content

Free cloud APIs ​

MiniTavern is BYOK (bring your own key): it does not include cloud credits. You register with a vendor or inference platform and paste your own API Key.
This page summarizes common free / trial quotas and how to fill them into the App. Quotas and policies change — always trust what each vendor’s console shows in real time.

Security note: an API Key is like a password. Store it only in the App’s on-device secure storage. Do not paste it into group chats, public screenshots, or public repos.

Fastest path ​

Your situationSuggestionApp platform
Mainland China, want to chat ASAPZhipu Flash free models or SiliconFlow free modelsCustom (OpenAI)
One Key for text + image (multimodal)Agnes AICustom (OpenAI) / Images
One Key for GPT / Gemini *-free + imageAIHubMixCustom (OpenAI) / Images
Coding-oriented gateway, limited Free textOpenCode ZenCustom (OpenAI)
Official Chinese LLM trialQwen / Alibaba Cloud BailianQwen
Reasoning / code-orientedDeepSeekDeepSeek
Overseas account, multimodal textGoogle GeminiGemini(Google)
One Key for many free open-source modelsOpenRouter :freeOpenRouter
Not sure which free vendors existFreeLLM directory (awesome-freellm-apis)See directory
Ultra-fast open-source inferenceGroqCustom (OpenAI)
Prefer no cloudRun local modelsCustom (OpenAI)

General Key steps: Model link. Local deploy: Run local models · 30-minute quickstart.

How to fill in MiniTavern ​

Path: Settings → Models → Model link → New

  1. Pick the correct model platform (see “App platform” below).
  2. Paste the API Key; for custom platforms also fill Base URL.
  3. Select or type the model ID (must match vendor docs; case-sensitive).
  4. Tap Test → on success, set as Default.
VendorApp platformBase URL (common)Notes
GeminiGemini(Google)PresetGemini protocol — do not fill as OpenAI
OpenRouterOpenRouterPresetFree model IDs often end with :free
Qwen / BailianQwenPreset (intl switchable)Use a general API Key, not a Coding Plan–only Key
DeepSeekDeepSeekPreset
Zhipu / SiliconFlow / Agnes / AIHubMix / Zen / Groq / Mistral, etc.Custom (OpenAI)See each sectionOpenAI-compatible gateways

China-focused (prefer mainland direct access) ​

Zhipu AI (GLM) ​

Best for: long-term free chat (Flash series); friendly for Chinese roleplay.

ItemDetails
Site / Keyopen.bigmodel.cn → Console → API Keys
Free tierPermanent free models such as GLM-*-Flash; new users often get extra trial credit (see console)
Card requiredUsually no
Base URLhttps://open.bigmodel.cn/api/paas/v4
Model examplesglm-4.5-flash, glm-4.7-flash (names follow the model plaza)
AppCustom (OpenAI)

Steps: register → create API Key → in App choose Custom (OpenAI) → Base URL + Key → Flash-series model ID → test.

SiliconFlow ​

Best for: one Key across many open-source chat models; some image models also have free tiers.

ItemDetails
Site / Keycloud.siliconflow.cn → Account → API keys
Free tierFree models are not billed (name without Pro/ prefix); complete real-name verification for full free catalog
Rate limitsFixed RPM/TPM per model; see official Rate Limits
Base URLhttps://api.siliconflow.cn/v1
Model examplesIn the plaza, filter “Free”, e.g. Qwen/Qwen3-8B (list changes)
AppCustom (OpenAI)

Note: the same name may have a free version and a paid Pro/... version — wrong ID bills you. Free-model usage on the bill should be 0.

Qwen / Alibaba Cloud Bailian ​

Best for: official Qwen series; MiniTavern has a Qwen preset.

ItemDetails
ActivateBailian console (China North 2 / Beijing) — accept terms to auto-activate
Free tierNew-user free quota: most models ~1M tokens / model, ~90 days; quotas do not share across models
KeyBailian → API-Key; use a general Key (Token Plan / Coding Plan–only Keys do not consume free quota)
AppQwen (China preset; international site switchable)
Avoid chargesEnable “Stop when free quota is used up” so pay-as-you-go does not start after quota ends

Official docs: new-user free quota. Calls after quota exhaustion or expiry incur charges.

Image gen (Wanxiang / Qwen-Image, etc.) may use new-user quota — in Settings → Image gen → Image gen link pick Qwen; whether it is free depends on that model’s quota bar in the console.

DeepSeek ​

Best for: reasoning, long text, high value; cheap after credits, but not permanently free.

ItemDetails
Site / Keyplatform.deepseek.com → API Keys
Free tierNew users usually get signup credit / trial balance (amount and rules change; see console)
AppDeepSeek
Model examplesdeepseek-v4-flash, deepseek-v4-pro (see official docs)

After credits run out, usage is billed; requests fail with zero balance.

ModelScope API-Inference ​

Best for: Alibaba Cloud real-name already done; try community open-source inference.

ItemDetails
Tokenmodelscope.cn → Access token
Free tierRegistered users get daily API-Inference quotas (totals and per-model caps: official limits)
Base URLhttps://api-inference.modelscope.cn/v1
AppCustom (OpenAI)

International / aggregators ​

Agnes AI ​

Best for: one free Key for text chat and image gen (both Flash free tier). Video is on the site; the App has no built-in video entry yet.

Full intro, model tables, and Key steps: Agnes AI.

ItemDetails
Site / Keyagnes-ai.com · platform.agnes-ai.com
Free tierFree Access: Flash text / image callable for free (RPM limits)
Base URLhttps://apihub.agnes-ai.com/v1
Text exampleagnes-2.5-flash
Image exampleagnes-image-2.1-flash (image link → Custom (OpenAI Images))
AppCustom (OpenAI) / Custom (OpenAI Images)

OpenCode Zen ​

Best for: Free text models on the OpenCode curated gateway (e.g. mimo-v2.5-free). Most flagships are metered — watch auto-recharge.

Full guide: OpenCode Zen.

ItemDetails
Site / Keyopencode.ai/auth · docs
Free tierModels marked Free on the pricing table (within caps); others are metered
Base URLhttps://opencode.ai/zen/v1
Text examplesmimo-v2.5-free, hy3-free, big-pickle
AppCustom (OpenAI) (prefer chat/completions-compatible IDs)

AIHubMix ​

Best for: one Key for many platform-subsidized *-free models (GPT / Gemini / GLM, etc.) and GPT-Image-2-free image gen; usually no card required.

Full guide: AIHubMix.

ItemDetails
Site / Keyaihubmix.com · free models post
Free tierModel IDs ending in -free; per-model RPM / daily caps
Base URLhttps://aihubmix.com/v1
Text examplesgpt-4o-free, gpt-5.5-free, coding-glm-5.1-free
Image examplegpt-image-2-free (Custom (OpenAI Images))
AppCustom (OpenAI) / Custom (OpenAI Images)

Google Gemini ​

Best for: multimodal text, long context; text free tier is relatively generous.

ItemDetails
Keyaistudio.google.com/app/apikey
Free tierSome Gemini Flash / Lite text models have a Free tier (RPM / RPD as shown in AI Studio)
Card requiredText free tier usually no; latest image models mostly have no Free tier
RegionEU / UK / Switzerland may lack free tier
PrivacyFree-tier prompts may be used to improve products (see Google terms)
AppGemini(Google)

For chat, pick a text model (e.g. gemini-flash-latest). Do not pick IDs with tts or image: Test sends a text reply request; TTS only outputs AUDIO and errors with response modalities (TEXT) is not supported. Configure voice/image under those features.

If Test/send returns HTTP 404 with no longer available to new users, that model ID is retired or closed to new users (e.g. gemini-2.5-flash-lite). Refresh the model list and pick a live Flash / Lite (e.g. gemini-flash-latest). You do not need the Interactions API.

If send fails with HTTP 429 and generate_content_free_tier_requests (often limit: 20), free-tier RPM is exhausted — wait for Please retry in … then retry. See ai.dev/rate-limit; raise limits via billing in AI Studio. This is not “no money” or a monthly project spend cap.

Image gen: Gemini App web and Developer API quotas differ; API image pricing pages often mark Free tier unavailable. Chat on Gemini + local A1111/ComfyUI for images is a common split — see Image model deploy.

OpenRouter ​

Best for: one Key across a large open-source / third-party free pool; IDs often end with :free. Full lists, tables, and troubleshooting: OpenRouter.

ItemDetails
Keyopenrouter.ai/keys
Free tier:free / Free collections; default ~20 RPM, limited daily requests; topping up can raise free daily caps (official rules)
AppOpenRouter
Model examplesopenrouter/free, meta-llama/llama-3.3-70b-instruct:free, openai/gpt-oss-20b:free
CollectionsFree Models · summary openairouter.net/free-models

Note: free upstreams may log prompts for training. On HTTP 429: Provider returned error, try another :free, wait, or BYOK.

Groq ​

Best for: very low latency open-source inference.

ItemDetails
Keyconsole.groq.com/keys
Free tierPermanent Free tier (RPM / RPD / TPM per model; changes over time)
Card requiredUsually no
Base URLhttps://api.groq.com/openai/v1
Model examplesllama-3.1-8b-instant, llama-3.3-70b-versatile, openai/gpt-oss-20b
AppCustom (OpenAI)

Mistral AI ​

ItemDetails
Keyconsole.mistral.ai/api-keys
Free tierExperiment / trial plans (monthly token scale per console); prompts may be used to improve products
Base URLhttps://api.mistral.ai/v1
AppCustom (OpenAI)

Cerebras / SambaNova and other fast inference ​

PlatformKey entryBase URLNotes
Cerebrascloud.cerebras.aihttps://api.cerebras.ai/v1Free tier often has daily token caps; may require a payment method
SambaNovacloud.sambanova.aihttps://api.sambanova.ai/v1Permanent free tier + occasional credits; strict rate limits

App: Custom (OpenAI) for both; model names follow each console.

GitHub Models ​

ItemDetails
EntryGitHub Models
AuthGitHub Personal Access Token
Base URLhttps://models.github.ai/inference
AppCustom (OpenAI)
LimitsPer-model RPM / RPD tiers — light trials

NVIDIA NIM ​

ItemDetails
Entrybuild.nvidia.com
Base URLhttps://integrate.api.nvidia.com/v1
AppCustom (OpenAI) (or Nvidia preset if restored in App later)

Trial credits (not permanently free) ​

These are mostly time-limited credits — fine for evaluation, not as your only long-term backend:

VendorTypical situationApp
OpenAIOccasional console trial credit; main API has no stable permanent free tierChatGPT(OpenAI)
Anthropic ClaudeConsole starter credits, burn quicklyClaude(Anthropic)
DeepSeekSignup credit aboveDeepSeek
Alibaba Cloud BailianNew-user token quota (~90 days)Qwen

When credits run out, reduce auto-charge risk (delete Key, enable stop-on-quota-end, or switch to free models / local).


Free image generation ​

PathCostNotes
Local A1111 / ComfyUIPower + GPUUnlimited images; phone → PC LAN
SiliconFlow free image modelsFree models = ¥0Image link → Custom (OpenAI Images), same SiliconFlow Base URL …/v1, pick a free image ID
Agnes AI image FlashFree AccessImage link → Custom (OpenAI Images), Base URL https://apihub.agnes-ai.com/v1, e.g. agnes-image-2.1-flash
AIHubMix*-free image modelsImage link → Custom (OpenAI Images), Base URL https://aihubmix.com/v1, e.g. gpt-image-2-free
Wanxiang / Qwen imageDepends on Bailian free quota for that modelSettings → Image gen → Image gen link → Qwen
OpenAI / most Gemini image APIsUsually paidDo not treat as “free image” without a stable Free tier

Chat and image links can differ: e.g. chat on Zhipu Flash, images on home ComfyUI.


Tips and troubleshooting ​

  1. Test before chatting: on failure tap “View details” and check Error codes.
  2. 429 / rate limits: lighter model, shorter memory, off-peak, or a second free platform.
  3. Unexpected charges: check for Pro/ paid models, Bailian without stop-on-quota-end, or exhausted credits.
  4. Privacy: free tiers often allow vendors to use logs to improve models; for sensitive content prefer local models.
  5. Quota changes: numbers here were common at research time; trust the vendor site. Broader index: FreeLLM directory (awesome-freellm-apis · freellm.net, unofficial, updated daily).