

Models share a common quota.
Limits (shared across models)
20 requests/minute50 requests/dayUp to 1000 requests/day with $10 lifetime topup
Models
- Hermes 3 Llama 3.1 405B
- Llama 3.2 3B Instruct
- Llama 3.3 70B Instruct
- cognitivecomputations/dolphin-mistral-24b-venice-edition:free
- cohere/north-mini-code:free
- google/gemma-4-26b-a4b-it:free
- google/gemma-4-31b-it:free
- nvidia/nemotron-3-nano-30b-a3b:free
- nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
- nvidia/nemotron-3-super-120b-a12b:free
- nvidia/nemotron-3-ultra-550b-a55b:free
- nvidia/nemotron-3.5-content-safety:free
- nvidia/nemotron-nano-12b-v2-vl:free
- nvidia/nemotron-nano-9b-v2:free
- openai/gpt-oss-20b:free
- poolside/laguna-m.1:free
- poolside/laguna-xs-2.1:free
- qwen/qwen3-coder:free
- qwen/qwen3-next-80b-a3b-instruct:free
- tencent/hy3:free