There are a few issues enabling this currently:
- It turns on parallel compiliation but nvlink is not bundled with Reactant's CUDA jll
- There appears to be invalidation issues where a cached kernel is re-used instead of invalidated on change. This leads to issues that appear like buffer aliasing. Need to investigate exactly what is going on here, less confident about the actual diagnosis
There are a few issues enabling this currently: