title: 2026 年養蝦 (OpenClaw) 必備，25 個免費 AI API 總整理

---

**Category**

- [AI](https://www.soft4fun.net/category/tech/ai)
- [科技](https://www.soft4fun.net/category/tech)

**Tag**

- [免費 API](https://www.soft4fun.net/tag/%e5%85%8d%e8%b2%bb-api)
- [免費 LLM API](https://www.soft4fun.net/tag/%e5%85%8d%e8%b2%bb-llm-api)

**圖片清單**

- ![free-llm-api](https://cf.img.soft4fun.net/2026/02/free-llm-api-scaled.jpg "free-llm-api")

---

「**OpenClaw**」(龍蝦) 最近在國內外爆紅，很多人把它形容成真正能幫你「做事情」的 AI 代理人，而不是只會聊天的機器人。它可以掛在你的電腦上，接 Telegram、接各種指令，然後背後串接大型語言模型去幫你處理任務。

 

不過要養這池蝦的成本不低，根據手哥實際測試，在設定初期對 Tokens 的使用量非常高，一個比較完整的設定使用數百~上千萬 Token 發生機率都很高，手哥就曾一個晚上噴掉將近 30 美元的費用。為了幫大家省錢，手哥整理了這份超過 25 家有提供「**[免費 LLM API](https://www.soft4fun.net/tag/%e5%85%8d%e8%b2%bb-llm-api)**」廠商清單，幫大家省省荷包！

                   快速導覽           

1. [OpenRouter：多模型整合神器](#openrouter-多模型整合神器)
2. [Google AI Studio：Gemini 官方å
  ¥口](#google-ai-studio-gemini-官方å
  ¥口)
  1. [Google AI Studio å
  �費模型額度總表](#google-ai-studio-å
  �費模型額度總表)
3. [NVIDIA NIM：官方推論平台](#nvidia-nim-官方推論平台)
4. [Mistral：開源模型強è€
  ](#mistral-開源模型強è€
  )
5. [HuggingFace Inference](#huggingface-inference)
6. [Vercel AI Gateway](#vercel-ai-gateway)
7. [Cerebras 與 Groq：高效能推論](#cerebras-與-groq-高效能推論)
8. [Cohere](#cohere)
9. [GitHub Models](#github-models)
10. [Cloudflare Workers AI](#cloudflare-workers-ai)
11. [Google Cloud Vertex AI](#google-cloud-vertex-ai)
12. [試用金型平台](#試用金型平台)
  1. [Hyperbolic 可用模型æ¸
  單](#hyperbolic-可用模型æ¸
  單)
  2. [SambaNova Cloud 可用模型æ¸
  單](#sambanova-cloud-可用模型æ¸
  單)
  3. [Scaleway Generative APIs 可用模型æ¸
  單](#scaleway-generative-apis-可用模型æ¸
  單)

     

> 💡溫馨提醒：å
> �費通常都有速率或額度限制，你不能當它是無限資源。但用來測試、做 PoC、甚至跑小型服務，å
> ¶實非常夠。

 

## OpenRouter：多模型整合神器

 

如果你想一次測多種模型，我會先推薦 OpenRouter。它等於幫你把很多模型整合成一個 API 入口。

 

免費政策：

 

- 所有å
  �費模型å
  ±用額度池
- 約 20 RPM
- 約 50 requests/day
- 若å
  值 $10 lifetime，可提升到 1,000 requests/day
- 無 token 月額度，但受 RPM / RPD 限制
- 不需信用卡即可開始

 

官網：[https://openrouter.ai](https://openrouter.ai)  
額度說明：[https://openrouter.ai/docs/api-reference/limits](https://openrouter.ai/docs/api-reference/limits)

 

優點是方便切模型，缺點是免費層真的只是給你測試用。OpenClaw 接它非常適合做模型比較。

 

目前提供的免費模型 (隨時會變化)：

 

- [Gemma 3 12B Instruct](https://openrouter.ai/google/gemma-3-12b-it:free)
- [Gemma 3 27B Instruct](https://openrouter.ai/google/gemma-3-27b-it:free)
- [Gemma 3 4B Instruct](https://openrouter.ai/google/gemma-3-4b-it:free)
- [Hermes 3 Llama 3.1 405B](https://openrouter.ai/nousresearch/hermes-3-llama-3.1-405b:free)
- [Llama 3.1 405B Instruct](https://openrouter.ai/meta-llama/llama-3.1-405b-instruct:free)
- [Llama 3.2 3B Instruct](https://openrouter.ai/meta-llama/llama-3.2-3b-instruct:free)
- [Llama 3.3 70B Instruct](https://openrouter.ai/meta-llama/llama-3.3-70b-instruct:free)
- [Mistral Small 3.1 24B Instruct](https://openrouter.ai/mistralai/mistral-small-3.1-24b-instruct:free)
- [Qwen 2.5 VL 7B Instruct](https://openrouter.ai/qwen/qwen-2.5-vl-7b-instruct:free)
- [allenai/molmo-2-8b:free](https://openrouter.ai/allenai/molmo-2-8b:free)
- [arcee-ai/trinity-large-preview:free](https://openrouter.ai/arcee-ai/trinity-large-preview:free)
- [arcee-ai/trinity-mini:free](https://openrouter.ai/arcee-ai/trinity-mini:free)
- [cognitivecomputations/dolphin-mistral-24b-venice-edition:free](https://openrouter.ai/cognitivecomputations/dolphin-mistral-24b-venice-edition:free)
- [deepseek/deepseek-r1-0528:free](https://openrouter.ai/deepseek/deepseek-r1-0528:free)
- [google/gemma-3n-e2b-it:free](https://openrouter.ai/google/gemma-3n-e2b-it:free)
- [google/gemma-3n-e4b-it:free](https://openrouter.ai/google/gemma-3n-e4b-it:free)
- [liquid/lfm-2.5-1.2b-instruct:free](https://openrouter.ai/liquid/lfm-2.5-1.2b-instruct:free)
- [liquid/lfm-2.5-1.2b-thinking:free](https://openrouter.ai/liquid/lfm-2.5-1.2b-thinking:free)
- [moonshotai/kimi-k2:free](https://openrouter.ai/moonshotai/kimi-k2:free)
- [nvidia/nemotron-3-nano-30b-a3b:free](https://openrouter.ai/nvidia/nemotron-3-nano-30b-a3b:free)
- [nvidia/nemotron-nano-12b-v2-vl:free](https://openrouter.ai/nvidia/nemotron-nano-12b-v2-vl:free)
- [nvidia/nemotron-nano-9b-v2:free](https://openrouter.ai/nvidia/nemotron-nano-9b-v2:free)
- [openai/gpt-oss-120b:free](https://openrouter.ai/openai/gpt-oss-120b:free)
- [openai/gpt-oss-20b:free](https://openrouter.ai/openai/gpt-oss-20b:free)
- [qwen/qwen3-4b:free](https://openrouter.ai/qwen/qwen3-4b:free)
- [qwen/qwen3-coder:free](https://openrouter.ai/qwen/qwen3-coder:free)
- [qwen/qwen3-next-80b-a3b-instruct:free](https://openrouter.ai/qwen/qwen3-next-80b-a3b-instruct:free)
- [tngtech/deepseek-r1t-chimera:free](https://openrouter.ai/tngtech/deepseek-r1t-chimera:free)
- [tngtech/deepseek-r1t2-chimera:free](https://openrouter.ai/tngtech/deepseek-r1t2-chimera:free)
- [tngtech/tng-r1t-chimera:free](https://openrouter.ai/tngtech/tng-r1t-chimera:free)
- [upstage/solar-pro-3:free](https://openrouter.ai/upstage/solar-pro-3:free)
- [z-ai/glm-4.5-air:free](https://openrouter.ai/z-ai/glm-4.5-air:free)

 

## Google AI Studio：Gemini 官方入口

 

Google AI Studio 提供 Gemini、Gemma 等模型的 API。免費層會有 tokens/分鐘與每日請求上限，依模型不同而異。

 

免費政策：

 

- 每個模型各自有 RPM / TPM / RPD
- Flash 系列：RPD 只有 20
- Gemma 系列：RPD 14,400
- 額度以「每模型」計算
- 不需信用卡即可使用 API key
- 可能依地區政策變動

 

官網：[https://aistudio.google.com](https://aistudio.google.com)

 

這個比較適合想直接用 Google 生態系模型的人。要注意的是免費政策會調整，敏感資料請一定看官方條款。

 

### Google AI Studio 免費模型額度總表

 

| 模型名稱 | 每分鐘 Token 上限 (TPM) | 每分鐘請求數 (RPM) | 每日請求數 (RPD) |
| --- | --- | --- | --- |
| **Gemini 3 Flash** | 250,000 tokens / 分鐘 | 5 requests / 分鐘 | 20 requests / 天 |
| **Gemini 2.5 Flash** | 250,000 tokens / 分鐘 | 5 requests / 分鐘 | 20 requests / 天 |
| **Gemini 2.5 Flash-Lite** | 250,000 tokens / 分鐘 | 10 requests / 分鐘 | 20 requests / 天 |
| **Gemma 3 27B Instruct** | 15,000 tokens / 分鐘 | 30 requests / 分鐘 | 14,400 requests / 天 |
| **Gemma 3 12B Instruct** | 15,000 tokens / 分鐘 | 30 requests / 分鐘 | 14,400 requests / 天 |
| **Gemma 3 4B Instruct** | 15,000 tokens / 分鐘 | 30 requests / 分鐘 | 14,400 requests / 天 |
| **Gemma 3 1B Instruct** | 15,000 tokens / 分鐘 | 30 requests / 分鐘 | 14,400 requests / 天 |

 

## NVIDIA NIM：官方推論平台

 

NVIDIA 提供免費推論（需手機驗證），免費層通常是每分鐘請求上限限制。

 

免費政策：

 

- 約 40 RPM
- 需手機驗證
- 不強調每日總量
- 多數模型 context 有限制

 

官網：[https://build.nvidia.com/explore/discover](https://build.nvidia.com/explore/discover)

 

適合想測 NVIDIA 生態模型的人。不過長對話或大 context 要注意模型限制。

 

## Mistral：開源模型強者

 

Mistral 有兩條線。

 

第一是 [La Plateforme](https://console.mistral.ai/)，有實驗性免費方案，需要手機驗證，並且可能同意資料用於訓練。每模型會有速率與 tokens 上限。

 

- 每模型 1 request/秒
- 高 TPM
- 每月 token cap（依官方說明）
- 需手機驗證
- 需同意資料可用於訓練

 

第二是 [Codestral](https://codestral.mistral.ai/)，偏向程式碼模型，免費但有限速與每日請求上限。

 

- 約 30 RPM
- 約 2,000 requests/day
- 偏程式碼模型
- 需手機驗證

 

如果你想讓 OpenClaw 做程式碼任務，Codestral 是不錯的選擇。

 

## HuggingFace Inference

 

HuggingFace 的 serverless 推論服務會給少量免費額度（約 US$0.10 等值），支援大量社群模型。

 

免費政策：

 

- 每月約 $0.10 推論 credit
- 用完即停
- 不限 RPM，但 credit 很少
- 模型大小限制（serverless ≤ 10GB）

 

文件與定價：[https://huggingface.co/docs/inference-providers/en/pricing](https://huggingface.co/docs/inference-providers/en/pricing)

 

優點是模型多到爆，缺點是免費額度真的不多。

 

## Vercel AI Gateway

 

Vercel 提供每月 US$5 免費額度，可作為模型路由層。

 

免費政策：

 

- 每月 $5 額度
- Gateway 層統一管理
- è¶
  出即轉付費

 

官網：[https://vercel.com/docs/ai-gateway](https://vercel.com/docs/ai-gateway)  
定價：[https://vercel.com/docs/ai-gateway/pricing](https://vercel.com/docs/ai-gateway/pricing)

 

如果你本來就在 Vercel 生態，這會很好用。

 

## Cerebras 與 Groq：高效能推論

 

[Cerebras](https://cloud.cerebras.ai/) 與 [Groq](https://console.groq.com)都是高效能推論平台。它們的免費額度會依模型不同，有些小模型每日請求數很高，大模型就會降很多。如果你想跑大量請求，小模型反而比較划算。

 

Cerebras 免費政策：

 

- 每模型 RPM 不同
- 每日請求上限（常見 14,400）
- Token/分鐘限制

 

Grok 免費政策：

 

- 每模型有 RPM / RPD
- 小模型每日上限高
- 大模型每日上限低

 

## Cohere

 

Cohere 免費層約每分鐘 20 次、每月 1,000 次。

 

免費政策：

 

- 20 RPM
- 每月 1,000 requests
- 多模型å
  ±用

 

官網：[https://cohere.com](https://cohere.com)

 

適合輕量應用，做產品級服務很快會不夠用。

 

## GitHub Models

 

GitHub Models 依 Copilot 訂閱等級決定使用額度。

 

免費政策：

 

- å
  �費用戶 token 很少
- 依訂閱不同é
  �額不同
- 偏開發測試用途

 

官網：[https://github.com/marketplace/models](https://github.com/marketplace/models)

 

比較適合在 GitHub 裡做原型測試。

 

## Cloudflare Workers AI

 

每日 10,000 neurons 免費。如果你想在邊緣環境部署 AI 任務，這個很方便。

 

免費政策：

 

- 每日 10,000 neurons
- 按計算單位消耗
- 適合邊緣運行

 

官方說明：[https://developers.cloudflare.com/workers-ai/platform/pricing/#free-allocation](https://developers.cloudflare.com/workers-ai/platform/pricing/#free-allocation)

 

## Google Cloud Vertex AI

 

部分模型在 preview 期間提供免費請求上限。這種免費多半是「暫時性紅利」，不要當成長期免費來源。

 

官網：[https://console.cloud.google.com/vertex-ai/model-garden](https://console.cloud.google.com/vertex-ai/model-garden)

 

## 試用金型平台

 

| Provider | 免費 / 試用額度 | 使用期限 | 其他要求 | 可用模型 |
| --- | --- | --- | --- | --- |
| **Fireworks** | $1 | 未標示 | 無特別說明 | Various open models |
| **Baseten** | $30 | 未標示 | 依算力時間計費 | Any supported model (pay by compute time) |
| **Nebius** | $1 | 未標示 | 無特別說明 | Various open models |
| **Novita** | $0.5 | 1 年 | 無特別說明 | Various open models |
| **AI21** | $10 | 3 個月 | 無特別說明 | Jamba family of models |
| **Upstage** | $10 | 3 個月 | 無特別說明 | Solar Pro / Solar Mini |
| **NLP Cloud** | $15 | 未標示 | 需手機驗證 | Various open models |
| **Alibaba Cloud (International) Model Studio** | 每模型 1,000,000 tokens | 未標示 | 無特別說明 | Various open models + proprietary Qwen models |
| **Modal** | $5/月（註冊後）$30/月（綁定付款方式） | 每月 | 依算力時間計費 | Any supported model (pay by compute time) |
| **Inference.net** | $1$25（回覆 email survey） | 未標示 | 需回覆調查信件 | Various open models |
| **Hyperbolic** | $1 | 未標示 | 無特別說明 | 詳細模型如下（完整列出） |
| **SambaNova Cloud** | $5 | 3 個月 | 無特別說明 | 詳細模型如下（完整列出） |
| **Scaleway Generative APIs** | 1,000,000 free tokens | 未標示 | 無特別說明 | 詳細模型如下（完整列出） |

 

### Hyperbolic 可用模型清單

 

- DeepSeek V3
- DeepSeek V3 0324
- Llama 3.1 405B Base
- Llama 3.1 405B Instruct
- Llama 3.1 70B Instruct
- Llama 3.1 8B Instruct
- Llama 3.2 3B Instruct
- Llama 3.3 70B Instruct
- Pixtral 12B (2409)
- Qwen QwQ 32B
- Qwen2.5 72B Instruct
- Qwen2.5 Coder 32B Instruct
- Qwen2.5 VL 72B Instruct
- Qwen2.5 VL 7B Instruct
- deepseek-ai/deepseek-r1-0528
- openai/gpt-oss-120b
- openai/gpt-oss-120b-turbo
- openai/gpt-oss-20b
- qwen/qwen3-235b-a22b
- qwen/qwen3-235b-a22b-instruct-2507
- qwen/qwen3-coder-480b-a35b-instruct
- qwen/qwen3-next-80b-a3b-instruct
- qwen/qwen3-next-80b-a3b-thinking

 

### SambaNova Cloud 可用模型清單

 

- E5-Mistral-7B-Instruct
- Llama 3.1 8B
- Llama 3.3 70B
- Llama 3.3 70B
- Llama-4-Maverick-17B-128E-Instruct
- Qwen/Qwen3-235B
- Qwen/Qwen3-32B
- Whisper-Large-v3
- deepseek-ai/DeepSeek-R1-0528
- deepseek-ai/DeepSeek-R1-Distill-Llama-70B
- deepseek-ai/DeepSeek-V3-0324
- deepseek-ai/DeepSeek-V3.1
- deepseek-ai/DeepSeek-V3.1-Terminus
- deepseek-ai/DeepSeek-V3.2
- openai/gpt-oss-120b
- tbd

 

### Scaleway Generative APIs 可用模型清單

 

- BGE-Multilingual-Gemma2
- DeepSeek R1 Distill Llama 70B
- Gemma 3 27B Instruct
- Llama 3.1 8B Instruct
- Llama 3.3 70B Instruct
- Mistral Nemo 2407
- Pixtral 12B (2409)
- Whisper Large v3
- devstral-2-123b-instruct-2512
- gpt-oss-120b
- holo2-30b-a3b
- mistral-small-3.2-24b-instruct-2506
- qwen3-235b-a22b-instruct-2507
- qwen3-coder-30b-a3b-instruct
- qwen3-embedding-8b
- voxtral-small-24b-2507
