llama.cpp b10984

cuda: support row-contiguous SUM_ROWS ( #26308 ) cuda: support row-contiguous SUM_ROWS organize the code and add GGML_OP_MEAN to support row-contiguous…

cuda: support row-contiguous SUM_ROWS ( #26308 ) cuda: support row-contiguous SUM_ROWS organize the code and add GGML_OP_MEAN to support row-contiguous tensors using the same shared kernel, and add a test to MEAN permute/slice Keep original comments and add if/else branch…

Read the original source — github.com

release · Shared by tscosj

0 comments

No comments yet.