According to the CUDA TOOLKIT DOCUMENTATION:
https://docs.nvidia.com/cuda/cuda-c-programming-guide/
Device memory can be allocated either as linear memory or as CUDA arrays.
Does this mean that the CUDA arrays are not stored linearly in GPU memory?
In my experiment, I successfully dumped my data from GPU memory based on the cudamemcpy function. If my data is allocated by cudaMallocArray, does it mean that the data are not physically linear in GPU memory and need to be extracted by other API?