gxw
|
8ab2e9ec65
|
LoongArch: DGEMM small matrix opt
|
2 years ago |
gxw
|
6017ad7146
|
loongarch64: Update dgemm_kernel_16x4 to dgemm_kernel_16x6
|
1 year ago |
pengxu
|
6546600342
|
Optimized ssymv and dsymv kernel LASX for LoongArch
|
1 year ago |
gxw
|
990507e3b8
|
LoongArch64: Opt zgemv with LASX
|
1 year ago |
gxw
|
d51ffec3a2
|
LoongArch64: Opt cgemv with LASX
|
1 year ago |
pengxu
|
4787a55c64
|
Optimized cgemm kernel 16x4 LASX for LoongArch
|
1 year ago |
pengxu
|
fe3da43b7d
|
Optimized zgemm kernel 8*4 LASX, 4*4 LSX and cgemm kernel 8*4 LSX for LoongArch
|
1 year ago |
gxw
|
7bc93d95a1
|
LoongArch64: Opt {c/z}axpby
|
1 year ago |
gxw
|
276e3ebf9e
|
LoongArch64: Add dzamax and dzamin opt
|
1 year ago |
pengxu
|
a5d0d21378
|
loongarch64: Add zgemm and cgemm optimization
|
1 year ago |
gxw
|
546f13558c
|
loongarch64: Add {c/z}swap and {c/z}sum optimization
|
1 year ago |
Hao Chen
|
edabb93668
|
loongarch64: Refine axpby optimization functions.
|
1 year ago |
Hao Chen
|
1ec5dded43
|
loongarch64: Add c/zrot optimization functions.
Signed-off-by: Hao Chen <chenhao@loongson.cn>
|
1 year ago |
Hao Chen
|
3c53ded315
|
loongarch64: Add c/znrm2 optimization functions.
|
1 year ago |
Hao Chen
|
fbd612f8c4
|
loongarch64: Add ic/zamin optimization functions.
|
1 year ago |
Hao Chen
|
d97272cb35
|
loongarch64: Add c/zdot optimization functions.
|
1 year ago |
Hao Chen
|
65a0aeb128
|
loongarch64: Add c/zcopy optimization functions.
Signed-off-by: Hao Chen <chenhao@loongson.cn>
|
1 year ago |
Hao Chen
|
2a34fb4b80
|
loongarch64: Add and refine scal optimization functions.
Signed-off-by: Hao Chen <chenhao@loongson.cn>
|
1 year ago |
Hao Chen
|
8785e948b5
|
loongarch64: Add camin optimization function.
|
1 year ago |
Hao Chen
|
0753848e03
|
loongarch64: Refine and add axpy optimization functions.
Signed-off-by: Hao Chen <chenhao@loongson.cn>
|
1 year ago |
Hao Chen
|
06fd5b5995
|
loongarch64: Add and Refine asum optimization functions.
|
1 year ago |
Hao Chen
|
173a65d4e6
|
loongarch64: Add and refine iamax optimization functions.
|
1 year ago |
zhoupeng
|
ea70e165c7
|
loongarch64: Refine rot optimization.
|
1 year ago |
zhoupeng
|
116aee7527
|
loongarch64: Refine imin optimization.
|
1 year ago |
zhoupeng
|
8be2654193
|
loongarch64: Refine imax optimization.
|
1 year ago |
zhoupeng
|
154baad454
|
loongarch64: Refine iamin optimization.
|
1 year ago |
Shiyou Yin
|
36c12c4971
|
loongarch64: Refine copy,swap,nrm2,sum optimization.
|
1 year ago |
Shiyou Yin
|
c6996a80e9
|
loongarch64: Refine amax,amin,max,min optimization.
|
1 year ago |
yancheng
|
d32f38fb37
|
loongarch64: Add optimizations for nrm2.
|
1 year ago |
yancheng
|
f9b468990e
|
loongarch64: Add optimizations for rot.
|
1 year ago |
yancheng
|
c80e7e27d1
|
loongarch64: Add optimizations for sum and asum.
|
1 year ago |
yancheng
|
d4c96a35a8
|
loongarch64: Add optimizations for axpy and axpby.
|
1 year ago |
yancheng
|
360acc0a41
|
loongarch64: Add optimizations for swap.
|
1 year ago |
yancheng
|
174c25766b
|
loongarch64: Add optimizations for copy.
|
1 year ago |
yancheng
|
49829b2b7d
|
loongarch64: Add optimizations for iamin.
|
1 year ago |
yancheng
|
be83f5e4e0
|
loongarch64: Add optimizations for iamax.
|
1 year ago |
yancheng
|
e3fb2b5afa
|
loongarch64: Add optimizations for imin.
|
1 year ago |
yancheng
|
e46b48e372
|
loongarch64: Add optimizations for imax.
|
1 year ago |
yancheng
|
702fc1d56d
|
loongarch64: Add optimization for min.
|
1 year ago |
yancheng
|
346b384d1c
|
loongarch64: Add optimization for max.
|
1 year ago |
yancheng
|
ff2ecc6cda
|
loongarch64: Add optimization for amin.
|
1 year ago |
yancheng
|
265b5f2e80
|
loongarch64: Add optimizations for amax.
|
1 year ago |
yancheng
|
993ede7c70
|
loongarch64: Add optimizations for scal.
|
1 year ago |
Shiyou Yin
|
13b8c44b44
|
loongarch: Add optimization for dsdot kernel.
|
1 year ago |
Shiyou Yin
|
3def6a8143
|
loongarch: Add LASX optimization for dot.
|
1 year ago |
gxw
|
4670eb1462
|
LoongArch64: Add dtrsm kernel
|
2 years ago |
gxw
|
f2cf929374
|
LoongArch64: Add sgemv kernel
|
2 years ago |
gxw
|
553cc1372f
|
LoongArch64: Add sgemm_kernel
|
2 years ago |
gxw
|
e8b571d245
|
LoongArch64: Add dgemv_t_8_lasx.S and dgemv_n_8_lasx.S V2
|
2 years ago |
Martin Kroeker
|
41c31bc1d4
|
Revert "LoongArch64: Add dgemv_t_8_lasx.S and dgemv_n_8_lasx.S"
|
2 years ago |