Tech article

Mini-AGI – dynamic continual learning model trained from scratch on 8GB VRAM

No preview is available. Read the original article for the full story.

Hacker News | Sep 21, 2026 | volotat

Automated excerpt

Every language model you can actually own today is a model somebody else trained and then froze. It is the same path training uses: same chunking, same cache, same gradient step. An Empirical Model of Large-Batch Training - McCandlish et al. , 2018.

Selected automatically from source text; not independently written or fact-checked. Read the original for full context.

Read the original article

More tech news