mirror of
https://github.com/ikawrakow/ik_llama.cpp.git
synced 2026-08-12 22:29:39 +04:00
* Allow concatenating quantized tensors * Missed this assert * Allow K to be f32 in ggml_cuda_op_indexer_topk
* Allow concatenating quantized tensors * Missed this assert * Allow K to be f32 in ggml_cuda_op_indexer_topk