I am writing a code to compute dot product of two vectors using CUBLAS

Question

0

Asked: June 12, 20262026-06-12T03:01:25+00:00 2026-06-12T03:01:25+00:00

I am writing a code to compute dot product of two vectors using CUBLAS

0

I am writing a code to compute dot product of two vectors using CUBLAS routine of dot product but it returns the value in host memory. I want to use the dot product for further computation on GPGPU only. How can I make the value reside on GPGPU only and use it for further computations without making an explicit copy from CPU to GPGPU?

Report

Leave an answer
Cancel reply

You must login to add an answer.

Need An Account,

1 Answer

Editorial Team · Answer 1 · 2026-06-12T03:01:26+00:00

~~You can’t, exactly, using CUBLAS.~~ As per talonmies’ answer, starting with the CUBLAS V2 api (CUDA 4.0) the return value can be a device pointer. Refer to his answer. But if you are using the V1 API it’s a single value, so it’s pretty trivial to pass it as an argument to a kernel that uses it—you don’t need an explicit cudaMemcpy (but there is one implied in order to return a host value).

Starting with the Tesla K20 GPU and CUDA 5, you will be able to call CUBLAS routines from device kernels using CUDA Dynamic Parallelism. This means you would be able to call cublasSdot (for example) from inside a __global__ kernel function, and your result would therefore be returned on the GPU.

Sign Up

Sign In

Forgot Password

The Archive Base Latest Questions

I am writing a code to compute dot product of two vectors using CUBLAS

Leave an answerCancel reply

1 Answer

Leave an answer
Cancel reply