NVIDIA NIM

minimaxai/minimax-m3
Context window197K tokens
Free budgetfree · 40 RPM tokens/mo
Requests / min40 RPM
Recurring free, 40 RPM, eval-only ToS info
NVIDIA NIM replaced its depleting trial credits with a recurring per-account rate limit (40 RPM default, varies by model), verified June 2026. The trial ToS still scopes usage to evaluation/prototyping, not production.

HuggingFace Router

MiniMaxAI/MiniMax-M3
Context window1.0M tokens
Free budget$0.10/mo credit tokens/mo
Rate limitsnot published
Small $0.10/month routed credit warning
HuggingFace Inference Providers grants only ~$0.10/month of routed credit on the free tier (PRO is $2/month). Enough for light experimentation; exhausts quickly. Credits apply only to HF-routed requests.

ModelScope

MiniMax/MiniMax-M3
Context window197K tokens
Free budgetfree · 2000 req/day account-wide tokens/mo
Requests / day100 RPD

NavyAI

minimax-m3
Context window1.0M tokens
Free budget~2.3M/month shared · 2x tokens/mo
Requests / min20 RPM
Tokens / day75K TPD

Get this model the moment it changes

Limits move, models get replaced, better ones launch. Premium routers see this page's data live. Free routers see last month's.

زنده کنید · $19/yr →

Use it

FreeLLMAPI is a self-hosted router you run yourself. Install it, paste in your free NVIDIA NIM key, and minimaxai/minimax-m3 answers on an OpenAI-compatible endpoint at http://localhost:3001/v1. No credit card, no hosted middleman: your prompts and your provider keys never leave your machine.

Install the router (macOS, Linux, WSL)
curl -fsSL https://freellmapi.co/install.sh | bash
Install the router (Windows PowerShell)
iwr -useb https://freellmapi.co/install.ps1 | iex
Call minimaxai/minimax-m3 with curl
curl http://localhost:3001/v1/chat/completions \
  -H "Authorization: Bearer freellmapi-your-unified-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "minimaxai/minimax-m3",
    "messages": [{"role": "user", "content": "Say hi in five words."}]
  }'
The same request in Python (openai)
from openai import OpenAI

client = OpenAI(
    base_url="http://localhost:3001/v1",
    api_key="freellmapi-your-unified-key",
)

resp = client.chat.completions.create(
    model="minimaxai/minimax-m3",
    messages=[{"role": "user", "content": "Say hi in five words."}],
)
print(resp.choices[0].message.content)

The router answers on /v1/chat/completions and every other OpenAI surface, plus the Anthropic Messages API, so existing clients need only a new base_url. Swap the model id for auto and the router picks the best free model that is still under its limits. Full reference: docs/api.md.