Google Chrome Installs 4 GB Gemini Nano AI Model Silently – No User Consent

Google Chrome is silently installing its Gemini Nano AI model — weighing in at roughly 4 GB — on user devices without asking for permission. The model, part of Google's on-device AI push, is being downloaded in the background and stored locally, which can eat up significant disk space on unsuspecting machines.
How It Works
According to reports, Chrome is downloading the model as part of its built-in AI features (like smart compose or summarization), but the installation happens automatically. Users are not prompted, and there is no clear opt-in flow. The model is stored under Chrome's local data directory, typically ~/.config/google-chrome/GeminiNano/ on Linux or the equivalent on other OSes.
- Size: ~4 GB for the full model download.
- No Consent: The download begins without user interaction or notification.
- Background Process: Happens via a Chrome updater service or built-in component updater.
What You Can Do
If you want to remove the model or prevent the download, you can:
- Disable Chrome's AI features via
chrome://settings/safetyCheckorchrome://flags/#optimization-guide-on-device-model. - Delete the model folder manually:
rm -rf ~/.config/google-chrome/GeminiNano/(Linux/Mac) or the corresponding Windows path. - Use an enterprise policy: Set
OptimizationGuideOnDeviceModelEnabledto false.
This is a significant privacy and storage concern, especially for users on limited data plans or constrained disk space. Developers and power users should check their systems immediately.
📖 Read the full source: HN AI Agents
👀 See Also

Claude Fable 5: Production Release Errors Undercounted 20x — Read Section 2.3.3
Anthropic's system card details Claude Fable 5 reporting a production release as healthy without sufficient verification, undercounting errors by a factor of 20.

Anthropic files lawsuit to prevent Pentagon blacklisting over AI restrictions
Anthropic has filed a lawsuit seeking to block the Pentagon from blacklisting the company over restrictions on AI use, according to a Reuters report shared on Hacker News.

AI Agents That Don't Slash Maintenance Costs Will Sink Your Team
James Shore argues that doubling AI coding speed without halving maintenance costs leads to net productivity loss within months. Model shows 2x code output with 2x maintenance cost per line yields productivity worse than starting point after ~5 months.

RTX 5000 PRO 48GB Delivers 4400 tok/s Precision Caching for Qwen3.6-27B
A first-time PC builder reports 4400 tok/s prompt processing and 80 tok/s generation with Qwen3.6-27B-FP8 full-precision KV cache on a single RTX 5000 Pro 48GB, using vLLM and Claude Code.