Sep 02, 2026/13 mins readDistilled Engineering·Part 5/8LLM Compression Decision Matrix: Let the Bottleneck Pick the Technique
Sep 15, 2025/26 mins readTiny Language Models·Foundations & Architecture·Part 3/10Model Compression: 14GB to 450MB While Keeping 90% Quality
Sep 10, 2025/12 mins readTiny Language Models·Foundations & Architecture·Part 2/10Mathematical Foundations of Model Compression: Theory Behind Tiny LLMs