
JustSayAIWeeklyReport
Kolibri-1 Goes Open Source, Gemini Cuts Free Tier: The Old LLM Era Is Over
This issue,
what to read.
The timeline on the left mirrors these chapters — read in order, or jump straight to any story.
Kolibri-1 Goes Open Source,Gemini Cuts Free Tier:
The Old LLM Era Is Over
The open-sourcing of Kolibri-1 stands in sharp contrast to Google's tightening of Gemini's free-tier allowances, signaling that the AI industry is moving beyond the old paradigm of monolithic super-models and unconditional free expansion. DeepMind's concept of symbiotic Agents further dismantles the singularity fantasy of all-capable models, as technical approaches begin to diverge. Notably, Google is simultaneously showcasing the engineering value of complex inference through Gemini 3.8 Flash while constraining its free layer—a tension that makes the interplay between commercialization and open-source ecosystems a critical variable in shaping what comes next.
“Kolibri, 78 billion parameters, only 3.46 billion active—that number just looks off.”
- Kolibri uses sparse activation: 3.46B of 78B parameters active, trading breadth for per-token cost efficiency in common scenarios.
- Apache 2.0 and million-context claims signal giving up universal dominance for cost-per-token competition.
- Industry logic shifted from leaderboard rankings to cost efficiency; Kolibri explicitly avoids chasing first place, taking a fundamentally different business path.
“Okay, second point — that "artificial symbiotic intelligence" thing from DeepMind? The moment I heard the term, something just felt off.”
- Artificial Symbiotic Intelligence isn't mere rhetoric—it acknowledges diminishing marginal returns for single LLMs, where scaling compute 10× may not yield 10× smarter capabilities.
- Five specialized Agents cross-checking and calling each other beats 10x parameters, cheaper too.
- Swapping individual IQ for organizational efficiency: five average people with a good process often beat one Einstein.
“Right after dropping Gemini Argon, they turned around and slashed the free tier.”
- Argon limits access to cybersecurity citing safety, but high inference costs prevent full opening.
- Safety is Google's cost-control valve, not just protection.
- New model launch cuts free tier, confirming LLM compute cost pressure is real.
worst session -1.70%(2026-10-01)
worst session -2.30%(2026-09-29)
worst session -1.49%(2026-10-02)
Kolibri, 78 billion parameters, only 3.46 billion active—that number just looks off.
- 01Kolibri uses sparse activation: 3.46B of 78B parameters active, trading breadth for per-token cost efficiency in common scenarios.
- 02Apache 2.0 and million-context claims signal giving up universal dominance for cost-per-token competition.
- 03Industry logic shifted from leaderboard rankings to cost efficiency; Kolibri explicitly avoids chasing first place, taking a fundamentally different business path.
Source ↗ huggingface.co
Okay, second point — that "artificial symbiotic intelligence" thing from DeepMind? The moment I heard the term, something just felt off.
- 01Artificial Symbiotic Intelligence isn't mere rhetoric—it acknowledges diminishing marginal returns for single LLMs, where scaling compute 10× may not yield 10× smarter capabilities.
- 02Five specialized Agents cross-checking and calling each other beats 10x parameters, cheaper too.
- 03Swapping individual IQ for organizational efficiency: five average people with a good process often beat one Einstein.
Source ↗ the-decoder.com
Right after dropping Gemini Argon, they turned around and slashed the free tier.
- 01Argon limits access to cybersecurity citing safety, but high inference costs prevent full opening.
- 02Safety is Google's cost-control valve, not just protection.
- 03New model launch cuts free tier, confirming LLM compute cost pressure is real.
Source ↗ nytimes.com
The full week — every daily brief's headline, linked to its issue:
09.29This week ran 5 headlines; 3 made the main thread; 12 daily briefs.
“Kolibri-1 Goes Open Source, Gemini Cuts Free Tier: The Old LLM Era Is Over”
Two issues a day. Ten minutes to turn AI noise into judgment — mornings for the world, evenings for China.

Two issues a day — AI noise into judgment.