When writing with CUDA you also have a lot of boilerplate and things to think about like streams, blocks, grids. The list goes on and on.
To get to that level of abstraction you'd have to use something like pytorchs cuda backend.
The CUDA machinery is still there, just not exposed to work with directly.
When writing with CUDA you also have a lot of boilerplate and things to think about like streams, blocks, grids. The list goes on and on.
To get to that level of abstraction you'd have to use something like pytorchs cuda backend.