llama-cpp 2 15 Repos Every AI Engineer Should Know to Run LLMs Faster (Without Burning GPU Budget) Apr 30, 2026 PrismML 1-bit Bonsai: Why 1-Bit LLMs Could Make On-Device AI Actually Practical Apr 2, 2026