Spaces:

natasa365
/

whisper.cpp

Sleeping

slaren commited on Mar 3, 2024

Commit

d1b60e4

unverified ·

1 Parent(s): 98e5c63

cuda : fix data race in soft max (llama/5853)

Files changed (1) hide show

ggml-cuda.cu CHANGED Viewed

@@ -6947,6 +6947,7 @@ static __global__ void soft_max_f32(const float * x, const float * mask, const f
     // find the sum of exps in the block
     tmp = warp_reduce_sum(tmp);
     if (block_size > WARP_SIZE) {
         if (warp_id == 0) {
             buf_iw[lane_id] = 0.0f;
         }

     // find the sum of exps in the block
     tmp = warp_reduce_sum(tmp);
     if (block_size > WARP_SIZE) {
+        __syncthreads();
         if (warp_id == 0) {
             buf_iw[lane_id] = 0.0f;
         }