tashfeenahmed/freellmapi
FreeLLMAPI self-hosts one API over many providers’ free tiers, with routing, failover, quota tracking, and coding-agent setup. Its capacity figures are project-maintained claims.
FreeLLMAPI claims a catalog of **34 providers**, **635 free endpoints**, and roughly **7.4B listed tokens per month** behind OpenAI-, Anthropic-, Gemini-, and Ollama-compatible interfaces. It adds routing, failover, quota tracking, encrypted local key storage, and agent setup generators.
For low-cost experiments, point a non-critical Codex or Claude Code setup at the local router and inspect the routed-provider header. Test tool calls, structured output, model switching, and rate-limit behavior before trusting a fallback chain with repository work.
FreeLLMAPI claims a catalog of **34 providers**, **635 free endpoints**, and roughly **7.4B listed tokens per month** behind OpenAI-, Anthropic-, Gemini-, and Ollama-compatible interfaces. It adds routing, failover, quota tracking, encrypted local key storage, and agent setup generators. For low-cost experiments, point a non-critical Codex or Claude Code setup at the local router and inspect the routed-provider header. Test tool calls, structured output, model switching, and rate-limit behavior before trusting a fallback chain with repository work. The capacity is aggregated from documented free tiers, not a guaranteed pool, and provider quotas change frequently. Free installs receive catalog additions after a **30-day delay**, while the live feed is paid; upstream providers still receive requests sent through them.
FreeLLMAPI broadens low-cost model experimentation by aggregating many changing free tiers behind familiar APIs, but it is a local routing catalog rather than guaranteed capacity or a privacy boundary. Against managed gateways, its differentiator is breadth and client compatibility; its tradeoffs are delayed free catalog updates, upstream data exposure, volatile quotas, and the need to validate fallback behavior per workload.