AIToday
Large Language ModelsAI Business & IndustryHacker NewsPublished: Mar 25, 2026, 12:26 JST1 min read

Researchers compress a 24-million parameter language model into just 15MB using GPTQ-lite and Muon optimization techniques

Researchers compress a 24-million parameter language model into just 15MB using GPTQ-lite and Muon optimization techniques

3 Key Points

  1. GolfStudent v2 achieves extreme model compression, reducing a 24M-parameter LLM to only 15MB in size

  2. Uses GPTQ-lite quantization combined with Muon optimization to achieve the compression

  3. Contribution submitted to OpenAI's parameter-golf project on GitHub, focusing on efficient model design

  4. Demonstrates significant progress in making capable language models deployable on resource-constrained devices

Ask the AI about this article →

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Visko raises $10M, launches live AI video model OrbisSiliconANGLE AI · 1h ago
  • Runway unveils Solaris, an AI that generates app interfaces in real timeTHE DECODER · 1h ago
  • Google AI Search flags Facebook users as dangerTHE DECODER · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleBudget-conscious AI enthusiast explores Tesla P40 GPU as affordable alternative to expensive RTX 3090 for running 30B parameter local language models.