AlpinDale a985143768 core: add cuda graph support for encoder-decoder models (#1051) 2 semanas atrás
..
__init__.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
cache_engine.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
cpu_model_runner.py 65a59bbb6b cpu: raise error if using encoder-decoder models (#1027) 2 semanas atrás
cpu_worker.py f2b6dc3872 cpu: add support for W8A8 quantization via compressed-tensor (#1017) 2 semanas atrás
embedding_model_runner.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
enc_dec_model_runner.py a985143768 core: add cuda graph support for encoder-decoder models (#1051) 2 semanas atrás
model_runner.py a985143768 core: add cuda graph support for encoder-decoder models (#1051) 2 semanas atrás
model_runner_base.py 304e1e5a8a core: dump model runner inputs during crash (#1023) 2 semanas atrás
multi_step_model_runner.py 1390915778 multi-step: add support for flashinfer attention backend (#1033) 2 semanas atrás
multi_step_tpu_worker.py 4b1b658855 tpu: implement multi-step scheduling (#1046) 2 semanas atrás
multi_step_worker.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
neuron_model_runner.py 145e554a4d neuron: add 8bit quantization for Neuron (#994) 2 semanas atrás
neuron_worker.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
openvino_model_runner.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
openvino_worker.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
tpu_model_runner.py 4b1b658855 tpu: implement multi-step scheduling (#1046) 2 semanas atrás
tpu_worker.py a50548c0b9 tpu: use XLA rank for persistent cache path (#989) 2 semanas atrás
utils.py a985143768 core: add cuda graph support for encoder-decoder models (#1051) 2 semanas atrás
worker.py a113309876 kernel: add meta functions for ops to prevent graph breaks (#1019) 2 semanas atrás
worker_base.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
xpu_model_runner.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás
xpu_worker.py 3bb0f07461 chore: rename `task_handler` to `worker` (#985) 2 semanas atrás