Free AI models are exploding — here’s what changed this week
Free AI models are exploding — here’s what changed this week
OpenRouter’s free tier jumped to 23 models (+2 vs yesterday). New additions: Google Gemma 4 (26B and 31B), OpenAI GPT-OSS-20b, Nvidia Nemotron Nano V2 variants, Qwen3-Next-80B-A3B.
The free API provider list grew to 56 providers (+4). MincAPI, SixFingerAPI, Subaxis, UnoRouter joined.
Google’s Gemini 3.5 Flash (1M context, 1500 RPD) and 3.1 Flash-Lite (30 RPM) remain stable. NVIDIA NIM confirmed operational for hosted free inference. Cerebras free tier: gpt-oss-120b, zai-glm-4.7 (~2,600 tok/s). GitHub Models still runs gpt-5, gpt-4.1, Llama 4 Scout, Mistral Small 3.1 for prototyping.
Sources: OpenRouter API, awesome-free-ai-api GitHub repo, AI Studio, NVIDIA NIM, Cerebras, GitHub Models.
Why it matters: Lower-cost prototyping just got a lot more options. You can now test bleeding-edge models without hitting a paywall or managing your own GPUs.
Related
More from the blog
Anthropic's Opus 5 Is About Efficiency, Not Magic
Opus 5 costs the same as 4.8 but delivers more per token — a reminder that AI progress is increasingly about squeezing better outputs out of what we already have.
Platforms Are Splitting on What Counts as Authentic
Substack just launched AI authorship labels while Deezer reports over half of daily uploads are AI-generated — and no platform agrees on how to solve it.
The $750 Billion Question: AI CapEx Has Gone Off the Rails
AI infrastructure spending ballooned to $750B for OpenAI alone, and nobody can say what the end state looks like.
AI Chip Shortage Is Making Smartphones More Expensive in India
India's smartphone shipments dropped 10% in Q2 2026 — not from weak demand, but from an AI chip shortage that's pushing RAM and storage prices up.