The free-tier landscape moves fast. Here are strong models available free right now, pulled straight from the live catalog.
- MiniMax M3 (NV) — free on NVIDIA NIM, up to 197K context.
- Kimi K2.7 Code (NavyAI) — free on NavyAI, up to 262K context.
- Gemini 3.6 Flash — free on Google AI Studio, up to 1.0M context.
- Kimi K3 (NavyAI) — free on NavyAI, up to 262K context.
- Kimi K3 (HF) — free on HuggingFace Router, up to 262K context.
- Gemini 3.5 Flash — free on Google AI Studio, up to 1.0M context.
- Qwen3-Coder Next (HF) — free on HuggingFace Router, up to 262K context.
- Kimi K2.6 (HF) — free on HuggingFace Router, up to 262K context.
- GLM-5.2 (NV) — free on NVIDIA NIM, up to 200K context.
- Nemotron-3 Ultra 550B (NV) — free on NVIDIA NIM, up to 1.0M context.
This list regenerates from the live FreeLLMAPI catalog as providers add and retire models. Want them the day they land instead of on the free 30-day delay? Premium keeps your router live.
Use every one of these through one key. FreeLLMAPI is an open-source, self-hosted router that puts all these free tiers behind a single OpenAI-compatible endpoint and fails over when one is rate-limited. Browse the catalog or go live.