LLM7
codestral-latest
Context window32K tokens
Free budget~2M (60-100/hr) tokens/mo
Requests / min40 RPM
No API key required info
Routes anonymously — the catalog ships a keyless sentinel row and calls work with no account or key.
Mistral
codestral-latest
Context window256K tokens
Free budget~50-100M tokens/mo
Requests / min2 RPM
Tokens / min500K TPM
xKiro
mistralai/codestral-2508
Context window256K tokens
Free budgetfree · 500K tokens/day (1M with Telegram) tokens/mo
Rate limitsnot published
Get this model the moment it changes
Limits move, models get replaced, better ones launch. Premium routers see this page's data live. Free routers see last month's.
लाइव करें · $19/yr →Use it
FreeLLMAPI is a self-hosted router you run yourself. Install it, paste in your free LLM7 key, and codestral-latest answers on an OpenAI-compatible endpoint at http://localhost:3001/v1. No credit card, no hosted middleman: your prompts and your provider keys never leave your machine.
Install the router (macOS, Linux, WSL)
curl -fsSL https://freellmapi.co/install.sh | bashInstall the router (Windows PowerShell)
iwr -useb https://freellmapi.co/install.ps1 | iexCall codestral-latest with curl
curl http://localhost:3001/v1/chat/completions \
-H "Authorization: Bearer freellmapi-your-unified-key" \
-H "Content-Type: application/json" \
-d '{
"model": "codestral-latest",
"messages": [{"role": "user", "content": "Say hi in five words."}]
}'The same request in Python (openai)
from openai import OpenAI
client = OpenAI(
base_url="http://localhost:3001/v1",
api_key="freellmapi-your-unified-key",
)
resp = client.chat.completions.create(
model="codestral-latest",
messages=[{"role": "user", "content": "Say hi in five words."}],
)
print(resp.choices[0].message.content)The router answers on /v1/chat/completions and every other OpenAI surface, plus the Anthropic Messages API, so existing clients need only a new base_url. Swap the model id for auto and the router picks the best free model that is still under its limits. Full reference: docs/api.md.