Just saw Saluki 27B's 2-bit Qwen3.8 beating the original at tool calling—thanks to clever model compression in llama.cpp/GGUF. Curious how tiny bits boost agents? Dive in! #Saluki27B #Qwen3_8 #2bitQuantization

Just saw Saluki 27B's 2-bit Qwen3.8 beating the original at tool calling—thanks to clever model compression in llama.cpp/GGUF. Curious how tiny bits boost agents? Dive in! #Saluki27B #Qwen3_8 #2bitQuantization
PrismML just dropped Ternary Bonsai 2 – a 5.9 GB, ternary‑weight version of a 27B model that hits Qwen 3.8‑level performance. Perfect for local, multimodal AI without the storage nightmare. Curious? Dive in! #TernaryBonsai2 #PrismML #Qwen3_8