Tri Dao
|
0e8c46ae08
Run isort and black on test files
|
1 år sedan |
Tri Dao
|
2a2a3c4bfd
[LayerNorm] Add test for randomness
|
1 år sedan |
Tri Dao
|
d2f4324f4c
[LayerNorm] Make sure memory addresses are aligned to 16 bytes
|
1 år sedan |
Tri Dao
|
393882bc08
[LayerNorm] Implement LN with parallel residual, support dim 8k
|
1 år sedan |
Tri Dao
|
6738d9477d
[LayerNorm] Implement RMS Norm
|
1 år sedan |
Tri Dao
|
5db330519a
[LayerNorm] Support taking subset of input or subset of output
|
2 år sedan |
Tri Dao
|
ae137ed17a
[LayerNorm] Fuse LayerScale
|
2 år sedan |
Tri Dao
|
8c6609ae1a
[LayerNorm] Support all dimensions up to 6k (if divisible by 8)
|
2 år sedan |
Tri Dao
|
fa6d1ce44f
Add fused_dense and dropout_add_layernorm CUDA extensions
|
2 år sedan |