AlpinDale 50c2434267 move megatron to a top-level directory 9 mesi fa
..
__init__.py c2aaaefd57 allow out-of-tree model registry 9 mesi fa
baichuan.py 50c2434267 move megatron to a top-level directory 9 mesi fa
bloom.py 50c2434267 move megatron to a top-level directory 9 mesi fa
chatglm.py 50c2434267 move megatron to a top-level directory 9 mesi fa
cohere.py 50c2434267 move megatron to a top-level directory 9 mesi fa
dbrx.py 50c2434267 move megatron to a top-level directory 9 mesi fa
decilm.py e31c6f0b45 feat: refactor modeling logic and support more models (#274) 11 mesi fa
deepseek.py 50c2434267 move megatron to a top-level directory 9 mesi fa
falcon.py 50c2434267 move megatron to a top-level directory 9 mesi fa
gemma.py 50c2434267 move megatron to a top-level directory 9 mesi fa
gpt2.py 50c2434267 move megatron to a top-level directory 9 mesi fa
gpt_bigcode.py 50c2434267 move megatron to a top-level directory 9 mesi fa
gpt_j.py 50c2434267 move megatron to a top-level directory 9 mesi fa
gpt_neox.py 50c2434267 move megatron to a top-level directory 9 mesi fa
internlm2.py 50c2434267 move megatron to a top-level directory 9 mesi fa
llama.py 50c2434267 move megatron to a top-level directory 9 mesi fa
llava.py 4d33ce60da feat: Triton flash attention backend for ROCm (#407) 9 mesi fa
mixtral.py 50c2434267 move megatron to a top-level directory 9 mesi fa
mpt.py 50c2434267 move megatron to a top-level directory 9 mesi fa
olmo.py 50c2434267 move megatron to a top-level directory 9 mesi fa
opt.py 50c2434267 move megatron to a top-level directory 9 mesi fa
phi.py 50c2434267 move megatron to a top-level directory 9 mesi fa
qwen.py 50c2434267 move megatron to a top-level directory 9 mesi fa
qwen2.py 50c2434267 move megatron to a top-level directory 9 mesi fa
qwen2moe.py 50c2434267 move megatron to a top-level directory 9 mesi fa
stablelm.py 50c2434267 move megatron to a top-level directory 9 mesi fa