Set TORCH_CUDA_ARCH_LIST=8.0;8.6+PTX - #1
Conversation
Xref git@github.com:weiji14/flash-attn-feedstock.git
Expand to CUDA compatibility 8.0 and above, xref https://developer.nvidia.com/cuda-gpus. Also increase MAX_JOBS.
|
Hi! This is the friendly automated conda-forge-linting service. I just wanted to let you know that I linted all conda-recipes in your PR ( |
|
@conda-forge-admin, please rerender |
…nda-forge-pinning 2024.05.07.15.53.14
…nda-forge-pinning 2024.05.07.15.53.14
|
Probably, there isn't enough time on Azure to complete these builds. If it does build in the 6 hours, great! Let's publish at least one build. In parallel, if you feel there are significant performance to be gained by building for '8.0,9.0+PTX' or even more archs, then please start doing the steps in this checklist in order to get this feedstock and yourself added to the allowlist for the gpu-long-running runners which have a time limit longer than 6 hours. |
So I've tried:
Oo yes, I was keeping an eye on those, thanks for pointing to the docs! I'll take a look at that. For now, let's just stick with |
Compile for CUDA compatibility 8.0 and above (Ampere generation or newer), continuing on work from conda-forge/staged-recipes#26239 (comment)
Note that build number is kept at 0, since the initial commit's (c75ac8e) build failed due to Azure pipelines running out of disk space.
Checklist
0(if the version changed)conda-smithy(Use the phrase@conda-forge-admin, please rerenderin a comment in this PR for automated rerendering)Continuing from conda-forge/staged-recipes#26239