Skip to content

Compile on blackwell with CUDA 13+ - #787

Closed
atvasilopoulos wants to merge 3 commits into
IBM:masterfrom
atvasilopoulos:compile_on_blackwell
Closed

atvasilopoulos wants to merge 3 commits into
IBM:masterfrom
atvasilopoulos:compile_on_blackwell

Conversation

@atvasilopoulos

Copy link
Copy Markdown
Collaborator

Description

Compiling from source did not work properly when trying to compile on a Blackwell GPU with CUDA 13+. Found that cub::TransformInputIterator and cub::CountingInputIterator are getting deprecated and are moving to the new calls cuda::transform_iterator and cuda::counting_iterator. Changed them appropriately with CUDART_VERSION >= 13000 guards.

Details

Compiling on a Blackwell GPU using CUDA 13.2 failed due to errors with the cub iterators. Changed them to the new iterators in the CUDA namespace, and the repo compiles properly. Tests are passing.

When running tests found two warnings regarding the RangeEstimator class, which are now fixed (detaching before conversion to float).

GPU model:
RTX PRO 6000 Blackwell Workstation Edition

CUDA compiler version:
Cuda compilation tools, release 13.2, V13.2.86
Build cuda_13.2.r13.2/compiler.37953736_0

@PabloCarmona

Copy link
Copy Markdown
Collaborator

Hello @atvasilopoulos ! Thanks for the PR, I think we are aiming for the same, can you check/review my previous PR addressing this problem, that will help a lot! #783

@atvasilopoulos

Copy link
Copy Markdown
Collaborator Author

I will close this since #783 was merged with a more complete solution to the problem. Thanks @PabloCarmona!

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants