flash-attention/flash_attn/losses
2022-11-13 17:27:26 -08:00
..
cross_entropy_apex.py Add fused cross entropy loss 2022-11-12 21:58:41 -08:00
cross_entropy_parallel.py Make nccl operations async in CrossEntropyLossParallel 2022-11-13 17:27:26 -08:00