cudaFuncSetAttribute and dynamic parallelism

A similar request was made here. The OP there filed a bug with NVIDIA (3503453), and that bug is in the late stages of being finalized. The development work is done, and based on my read of the bug, it should become available in CUDA 12.1 (if it is not already available in 12.0. I haven’t tested 12.0, but based on what I see in the bug it doesn’t appear to be in 12.0). There is no guarantee that it will be in 12.1, that is just my best guess based on what I see so far.