kliuae
|
ff5abcd746
|
[ROCm] Add support for Punica kernels on AMD GPUs (#3140)
Co-authored-by: miloice <jeffaw99@hotmail.com>
|
2024-05-09 09:19:50 -07:00 |
Alexei-V-Ivanov-AMD
|
9b5c9f9484
|
[CI/Build] AMD CI pipeline with extended set of tests. (#4267)
Co-authored-by: simon-mo <simon.mo@hey.com>
|
2024-05-02 12:29:07 -07:00 |
Hongxia Yang
|
95e5b087cf
|
[AMD][Hardware][Misc][Bugfix] xformer cleanup and light navi logic and CI fixes and refactoring (#4129)
|
2024-04-21 21:57:24 -07:00 |
Woosuk Kwon
|
bd3c144e0b
|
[Bugfix][ROCm] Add numba to Dockerfile.rocm (#3962)
|
2024-04-10 07:37:17 -07:00 |
Juan Villamizar
|
6c0b04515f
|
[ROCm][Hardware][AMD] Use Triton Kernel for default FA on ROCm (#3643)
Co-authored-by: jpvillam <jpvillam@amd.com>
Co-authored-by: Gregory Shtrasberg <Gregory.Shtrasberg@amd.com>
Co-authored-by: Woosuk Kwon <woosuk.kwon@berkeley.edu>
|
2024-04-09 15:10:47 -07:00 |
Hongxia Yang
|
9765b5c406
|
[ROCm][Bugfix] Fixed several bugs related to rccl path and attention selector logic (#3699)
|
2024-03-29 14:52:36 -07:00 |
youkaichao
|
8f44facddd
|
[Core] remove cupy dependency (#3625)
|
2024-03-27 00:33:26 -07:00 |
Hongxia Yang
|
05af6da8d9
|
[ROCm] enable cupy in order to enable cudagraph mode for AMD GPUs (#3123)
Co-authored-by: lcskrishna <lollachaitanya@gmail.com>
|
2024-03-04 18:14:53 -08:00 |
Hongxia Yang
|
5255d99dc5
|
[ROCm] Dockerfile fix for flash-attention build (#2885)
|
2024-02-15 10:22:39 -08:00 |
Hongxia Yang
|
0580aab02f
|
[ROCm] support Radeon™ 7900 series (gfx1100) without using flash-attention (#2768)
|
2024-02-10 23:14:37 -08:00 |
Hongxia Yang
|
c81dddb45c
|
[ROCm] Fix build problem resulted from previous commit related to FP8 kv-cache support (#2790)
|
2024-02-06 22:36:59 -08:00 |
Douglas Lehr
|
2ccee3def6
|
[ROCm] Fixup arch checks for ROCM (#2627)
|
2024-02-05 14:59:09 -08:00 |
Hongxia Yang
|
6b7de1a030
|
[ROCm] add support to ROCm 6.0 and MI300 (#2274)
|
2024-01-26 12:41:10 -08:00 |
TJian
|
05bdf4eaf3
|
Fix Dockerfile.rocm (#2101)
Co-authored-by: miloice <jeffaw99@hotmail.com>
|
2023-12-14 00:45:58 -08:00 |
TJian
|
f375ec8440
|
[ROCm] Upgrade xformers version for ROCm & update doc (#2079)
Co-authored-by: miloice <jeffaw99@hotmail.com>
|
2023-12-13 00:56:05 -08:00 |
TJian
|
6ccc0bfffb
|
Merge EmbeddedLLM/vllm-rocm into vLLM main (#1836)
Co-authored-by: Philipp Moritz <pcmoritz@gmail.com>
Co-authored-by: Amir Balwel <amoooori04@gmail.com>
Co-authored-by: root <kuanfu.liu@akirakan.com>
Co-authored-by: tjtanaa <tunjian.tan@embeddedllm.com>
Co-authored-by: kuanfu <kuanfu.liu@embeddedllm.com>
Co-authored-by: miloice <17350011+kliuae@users.noreply.github.com>
|
2023-12-07 23:16:52 -08:00 |