| Model | Provider | Free type | Modalities |
|---|---|---|---|
@cf/deepseek-ai/deepseek-r1-distill-qwen-32b | Cloudflare Workers AI | renewing-quota | text, embeddings, image, audio |
@cf/google/gemma-3-12b-it | Cloudflare Workers AI | renewing-quota | text, embeddings, image, audio |
@cf/meta/llama-3.1-8b-instruct | Cloudflare Workers AI | renewing-quota | text, embeddings, image, audio |
@cf/meta/llama-3.3-70b-instruct-fp8-fast | Cloudflare Workers AI | renewing-quota | text, embeddings, image, audio |
@cf/mistralai/mistral-small-3.1-24b-instruct | Cloudflare Workers AI | renewing-quota | text, embeddings, image, audio |
@cf/qwen/qwen2.5-coder-32b-instruct | Cloudflare Workers AI | renewing-quota | text, embeddings, image, audio |
01-ai/yi-large | NVIDIA NIM | trial-credit | text |
adept/fuyu-8b | NVIDIA NIM | trial-credit | text |
ai21labs/jamba-1.5-large-instruct | NVIDIA NIM | trial-credit | text |
aisingapore/sea-lion-7b-instruct | NVIDIA NIM | trial-credit | text |
baai/bge-m3 | NVIDIA NIM | trial-credit | text |
bigcode/starcoder2-15b | NVIDIA NIM | trial-credit | text |
cohere/cohere-command-a | GitHub Models | renewing-quota | text |
cohere/north-mini-code:free | OpenRouter | renewing-quota | text, vision |
command-a-03-2025 | Cohere | renewing-quota | text, embeddings, rerank |
command-r-08-2024 | Cohere | renewing-quota | text, embeddings, rerank |
command-r-plus-08-2024 | Cohere | renewing-quota | text, embeddings, rerank |
command-r7b-12-2024 | Cohere | renewing-quota | text, embeddings, rerank |
databricks/dbrx-instruct | NVIDIA NIM | trial-credit | text |
deepseek-ai/deepseek-coder-6.7b-instruct | NVIDIA NIM | trial-credit | text |
deepseek-r1-distill-llama-70b | Scaleway Generative APIs | trial-credit | text, audio |
DeepSeek-V3.1 | SambaNova Cloud | renewing-quota | text |
deepseek-v4-flash | Ollama Cloud | renewing-quota | text, vision |
deepseek-v4-pro | Ollama Cloud | renewing-quota | text, vision |
deepseek/deepseek-r1 | GitHub Models | renewing-quota | text |
deepseek/deepseek-r1-0528 | GitHub Models | renewing-quota | text |
embed-v4.0 | Cohere | renewing-quota | text, embeddings, rerank |
gemini-2.5-flash | Google Gemini API (AI Studio) | renewing-quota | text, vision, embeddings, audio |
gemini-2.5-flash-lite | Google Gemini API (AI Studio) | renewing-quota | text, vision, embeddings, audio |
gemini-2.5-pro | Google Gemini API (AI Studio) | renewing-quota | text, vision, embeddings, audio |
gemma-3-27b-it | Scaleway Generative APIs | trial-credit | text, audio |
gemma-4-31b | Cerebras | renewing-quota | text |
gemma4:31b | Ollama Cloud | renewing-quota | text, vision |
glm-4.5-flash | Z.ai (Zhipu AI / GLM) | perpetual | text, vision |
glm-4.6v-flash | Z.ai (Zhipu AI / GLM) | perpetual | text, vision |
glm-4.7-flash | Z.ai (Zhipu AI / GLM) | perpetual | text, vision |
glm-5.1 | Ollama Cloud | renewing-quota | text, vision |
glm-5.2 | Ollama Cloud | renewing-quota | text, vision |
google/gemma-4-26b-a4b-it:free | OpenRouter | renewing-quota | text, vision |
google/gemma-4-31b-it:free | OpenRouter | renewing-quota | text, vision |
gpt-oss-120b | Cerebras | renewing-quota | text |
gpt-oss-120b | SambaNova Cloud | renewing-quota | text |
gpt-oss:120b | Ollama Cloud | renewing-quota | text, vision |
gpt-oss:20b | Ollama Cloud | renewing-quota | text, vision |
inclusionai/ling-3.0-flash:free | OpenRouter | renewing-quota | text, vision |
kimi-k2.5 | Ollama Cloud | renewing-quota | text, vision |
llama-3.1-8b-instant | Groq | renewing-quota | text, audio |
llama-3.3-70b-instruct | Scaleway Generative APIs | trial-credit | text, audio |
llama-3.3-70b-versatile | Groq | renewing-quota | text, audio |
LLM-Research/Llama-4-Maverick-17B-128E-Instruct | ModelScope (API-Inference) | renewing-quota | text, vision |
MedAIBase/AntAngelMed | ModelScope (API-Inference) | renewing-quota | text, vision |
Meta-Llama-3.3-70B-Instruct | SambaNova Cloud | renewing-quota | text |
meta/llama-3.2-11b-vision-instruct | GitHub Models | renewing-quota | text |
meta/llama-3.2-90b-vision-instruct | GitHub Models | renewing-quota | text |
microsoft/phi-4 | GitHub Models | renewing-quota | text |
MiniMax/MiniMax-M1-80k | ModelScope (API-Inference) | renewing-quota | text, vision |
mistral-ai/codestral-2501 | GitHub Models | renewing-quota | text |
mistral-small-3.2-24b-instruct-2506 | Scaleway Generative APIs | trial-credit | text, audio |
MusePublic/Qwen-Image-Edit | ModelScope (API-Inference) | renewing-quota | text, vision |
nvidia/nemotron-3-nano-30b-a3b:free | OpenRouter | renewing-quota | text, vision |
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | OpenRouter | renewing-quota | text, vision |
openai-fast | Pollinations.ai | perpetual | image, text, audio |
openai/gpt-4.1 | GitHub Models | renewing-quota | text |
openai/gpt-oss-120b | Groq | renewing-quota | text, audio |
openai/gpt-oss-20b | Groq | renewing-quota | text, audio |
openai/gpt-oss-20b:free | OpenRouter | renewing-quota | text, vision |
OpenGVLab/InternVL3_5-241B-A28B | ModelScope (API-Inference) | renewing-quota | text, vision |
PaddlePaddle/ERNIE-4.5-0.3B-PT | ModelScope (API-Inference) | renewing-quota | text, vision |
poolside/laguna-s-2.1:free | OpenRouter | renewing-quota | text, vision |
Qwen/Qwen-Image-Edit | ModelScope (API-Inference) | renewing-quota | text, vision |
qwen2.5-coder-32b-instruct | Scaleway Generative APIs | trial-credit | text, audio |
qwen3-235b-a22b-instruct-2507 | Scaleway Generative APIs | trial-credit | text, audio |
rerank-v3.5 | Cohere | renewing-quota | text, embeddings, rerank |
Shanghai_AI_Laboratory/Intern-S1 | ModelScope (API-Inference) | renewing-quota | text, vision |
whisper-large-v3 | Groq | renewing-quota | text, audio |
whisper-large-v3-turbo | Groq | renewing-quota | text, audio |
zai-glm-4.7 | Cerebras | renewing-quota | text |
Model lists are a live sample and change often — always confirm against the provider. A provider missing here usually needs an API key to list its models; see how the list is refreshed.