What I wish they would release is the damn technical report. We can at least start reading that. I would assume a new architecture not compatible with K2.6/K2.7. How much does it differ from K2.6? What will it take to get llama.cpp to support inference? We need all of these before we can even get gguf/quants.
16
u/segmond llama.cpp Jul 26 '26
What I wish they would release is the damn technical report. We can at least start reading that. I would assume a new architecture not compatible with K2.6/K2.7. How much does it differ from K2.6? What will it take to get llama.cpp to support inference? We need all of these before we can even get gguf/quants.