whisper.cpp

Running

App Files Files Community

matteogeniaccio matteo serva

JohannesGaessler commited on Aug 1, 2024

Commit

686bb18

1 Parent(s): 0019ddb

ggml-cuda: Adding support for unified memory (llama/8035)

Browse files

* Adding support for unified memory

* adding again the documentation about unified memory

* refactoring: Moved the unified memory code in the correct location.

* Fixed compilation error when using hipblas

* cleaning up the documentation

* Updating the documentation

Co-authored-by: Johannes Gäßler <[email protected]>

* adding one more case where the PR should not be enabled

---------

Co-authored-by: matteo serva <[email protected]>
Co-authored-by: Johannes Gäßler <[email protected]>

Files changed (1) hide show

ggml/src/ggml-cuda.cu +15 -0

ggml/src/ggml-cuda.cu CHANGED Viewed

@@ -130,7 +130,22 @@ static cudaError_t ggml_cuda_device_malloc(void ** ptr, size_t size, int device)
     }
     return res;
 #else
     return cudaMalloc(ptr, size);
 #endif
 }

     }
     return res;
 #else
+#if !defined(GGML_USE_HIPBLAS) && !defined(GGML_USE_MUSA)
+    cudaError_t err;
+    if (getenv("GGML_CUDA_ENABLE_UNIFIED_MEMORY") != nullptr)
+    {
+        err = cudaMallocManaged(ptr, size);
+    }
+    else
+    {
+        err = cudaMalloc(ptr, size);
+    }
+    return err;
+#else
     return cudaMalloc(ptr, size);
+#endif // !defined(GGML_USE_HIPBLAS) && !defined(GGML_USE_MUSA)
 #endif
 }