Tri Dao b16d814c62 Revert to before Cutlass 3.6.0 update to investigate perf issue il y a 1 semaine
..
composable_kernel @ a9b170b541 e2182cc21d Support page kvcache in AMD ROCm (#1198) il y a 3 mois
cutlass @ e1cd8c7866 b16d814c62 Revert to before Cutlass 3.6.0 update to investigate perf issue il y a 1 semaine
flash_attn 83e41b3ca4 Add custom ops for compatibility with PT Compile (#1139) il y a 3 mois
flash_attn_ck 53a4f34163 Hotfix due to change of upstream api (#1239) il y a 3 mois
ft_attention 50896ec574 Make nvcc threads configurable via environment variable (#885) il y a 10 mois
fused_dense_lib 50896ec574 Make nvcc threads configurable via environment variable (#885) il y a 10 mois
fused_softmax 50896ec574 Make nvcc threads configurable via environment variable (#885) il y a 10 mois
layer_norm 50896ec574 Make nvcc threads configurable via environment variable (#885) il y a 10 mois
rotary 50896ec574 Make nvcc threads configurable via environment variable (#885) il y a 10 mois
xentropy 50896ec574 Make nvcc threads configurable via environment variable (#885) il y a 10 mois