wernsaar
|
410afda9b4
|
added cpu detection and target ARMV6, used in raspberry pi
|
12 years ago |
wernsaar
|
bf04544902
|
added gemv_n kernel for single and double precision
|
12 years ago |
wernsaar
|
86283c0be1
|
added gemv_t kernel for single and double precision
|
12 years ago |
wernsaar
|
f27cabfd08
|
added nrm2 kernel for all precisions
|
12 years ago |
wernsaar
|
23dd474cd0
|
added rot kernel for all precisions
|
12 years ago |
wernsaar
|
f1b452e160
|
added scal kernel for all precisions
|
12 years ago |
wernsaar
|
3dabd7e6e6
|
added swap-kernel for all precisions
|
12 years ago |
wernsaar
|
6f4a0ebe38
|
added max- und min-kernels for all precisions
|
12 years ago |
wernsaar
|
f750103336
|
small optimizations on dot-kernels
|
12 years ago |
wernsaar
|
00f33c0134
|
added asum_kernel for all precisions and complex
|
12 years ago |
wernsaar
|
5b36cc0f47
|
added blas level1 dot kernels for complex and double complex
|
12 years ago |
wernsaar
|
c8f1aeb154
|
added optimized blas level1 dot kernels for single and double precision
|
12 years ago |
wernsaar
|
8fa93be06e
|
added optimized blas level1 copy kernels
|
12 years ago |
wernsaar
|
1e8128f41c
|
added cgemm_tcopy_2_vfpv3.S and zgemm_tcopy_2_vfpv3.S
|
12 years ago |
wernsaar
|
80a2e901b1
|
added dgemm_tcopy_4_vfpv3.S and sgemm_tcopy_4_vfpv3.S
|
12 years ago |
wernsaar
|
ac50bccbd2
|
added cgemm_ncopy_2_vfpv3.S and made assembler labels unique
|
12 years ago |
wernsaar
|
82015beaef
|
added zgemm_ncopy_2_vfpv3.S and made assembler labels unique
|
12 years ago |
wernsaar
|
370e3834a9
|
added missing file kernel/arm/Makefile
|
12 years ago |
wernsaar
|
e31186efd4
|
deleted obsolete dgemm_kernel and dtrmm_kernel
|
12 years ago |
wernsaar
|
2b801a00a5
|
small optimizations on sgemm_kernel for ARMV7
|
12 years ago |
wernsaar
|
b3eab8fcb7
|
minor optimizations on zgemm_kernel for ARMV7
|
12 years ago |
wernsaar
|
02bc36ac79
|
added sgemm_ncopy routine and made some improvements on cgemm_kernel for ARMV7
|
12 years ago |
wernsaar
|
85484a42df
|
added kernels for cgemm, ctrmm, zgemm and ztrmm
|
12 years ago |
wernsaar
|
3983011f0b
|
added sgemm- and strmm_kernel
|
12 years ago |
wernsaar
|
2a1515c9dd
|
added dgemm_ncopy_4_vfpv3.S
|
12 years ago |
wernsaar
|
31f51e78bc
|
minor optimizations on dgemm_kernel
|
12 years ago |
wernsaar
|
e0b968c3a7
|
Changed kernels for dgemm and dtrmm
|
12 years ago |
wernsaar
|
1c63180bb6
|
updated dgemm_kernel_8x2_vfpv3.S
|
12 years ago |
wernsaar
|
4a474ea7dc
|
changed dgemm_kernel to use fused multiply add
|
12 years ago |
wernsaar
|
69ce737cc5
|
modified Makefile.L3 for ARM
|
12 years ago |
wernsaar
|
70411af888
|
initial checkin of kernel/arm
|
12 years ago |
wernsaar
|
067e8417fd
|
removed unnessesary instructions from zgemm_kernel_2x2_bulldozer.S
|
12 years ago |
wernsaar
|
a82da3d069
|
removed unnessesary instructions
|
12 years ago |
Zhang Xianyi
|
1569bf14f8
|
Refs #282. Fixed zgemv_n typo bug on Win64.
|
12 years ago |
Zhang Xianyi
|
c0159d44a3
|
Merge branch 'develop' of https://github.com/wernsaar/OpenBLAS into wernsaar-develop
|
12 years ago |
wernsaar
|
c17a850c1c
|
modified KERNEL.BULLDOZER
|
12 years ago |
wernsaar
|
099853fff6
|
added dtrsm_kernel_RN_8x2_bulldozer.S
|
12 years ago |
wernsaar
|
44d23881b5
|
dtrsm_kernel_LT_8x2_bulldozer.S performance optimization
|
12 years ago |
Zhang Xianyi
|
32fb6b9bb2
|
Merge branch 'develop' of https://github.com/wernsaar/OpenBLAS into wernsaar-develop
|
12 years ago |
wernsaar
|
aaeb8eaecd
|
modified dtrsm_kernel_LT_8x2_bulldozer.S
|
12 years ago |
wernsaar
|
8aeec32ea0
|
modified dtrsm_kernel_LT_8x2_bulldozer.S
|
12 years ago |
wernsaar
|
87fc9de572
|
added dtrsm_kernel_LT_8x2_bulldozer.S
|
12 years ago |
wernsaar
|
564aa60fec
|
removed dtrsm_kernel_LT_8x2_bulldozer.S
|
12 years ago |
wernsaar
|
f645665dd6
|
fixed bug in dgemv_t_bulldozer.S
|
12 years ago |
wernsaar
|
e45a347cd2
|
repaired trmm bug in sgemm_kernel_16x2_bulldozer.S
|
12 years ago |
wernsaar
|
99727ac013
|
repaired trmm bug in cgemm_kernel_4x2_bulldozer.S
|
12 years ago |
wernsaar
|
6e0a2fbc0c
|
repaired trmm bug in zgemm_kernel_2x2_bulldozer.S
|
12 years ago |
wernsaar
|
0a22f99c58
|
repaired trmm bug in dgemm_kernel_8x2_bulldozer.S
|
12 years ago |
wernsaar
|
cff70a666d
|
added generic trmm kernels and modified Makefile.L3
|
12 years ago |
wernsaar
|
84bd0aabaa
|
added dtrsm_kernel_LT_8x2_bulldozer.S
|
12 years ago |