AIDev.news
llama.cpp b11345: Adds q2_k and q3_k Quantization Support for Hexagon Backend · AIDev.news