Chirag Jain
|
50896ec574
Make nvcc threads configurable via environment variable (#885)
|
9 kuukautta sitten |
Tri Dao
|
dc08ea1c33
Support H100 for other CUDA extensions
|
1 vuosi sitten |
Tri Dao
|
fa6d1ce44f
Add fused_dense and dropout_add_layernorm CUDA extensions
|
2 vuotta sitten |