Martin Kroeker
7c162b8a21
Update isamax_power8.S
6 years ago
Martin Kroeker
0544cbc806
Fix syntax of endianness conditional
6 years ago
Martin Kroeker
120d20731f
Fix syntax of endianness conditional
6 years ago
Martin Kroeker
dc345d84df
Fix syntax of endianness conditional and add gcc version check for workaround
6 years ago
Martin Kroeker
616921fd91
Merge pull request #27 from xianyi/develop
rebase
6 years ago
Martin Kroeker
8a9e9a82a1
Merge pull request #2410 from bartoldeman/fix-dscal-inline-asm
Fix inline asm in dscal: mark x, x1 as clobbered. Fixes #2408
6 years ago
Bart Oldeman
7ea5e07d1c
Fix inline asm in dscal: mark x, x1 as clobbered. Fixes #2408
The leaq instructions in dscal_kernel_inc_8 modify x and x1 so they
must be declared as input/output constraints, otherwise the compiler
may assume the corresponding registers are not modified.
6 years ago
Martin Kroeker
cb6ef49857
Merge pull request #2407 from susilehtola/patch-2
Patch out instances of Z15 in dynamic_zarch.c
6 years ago
Martin Kroeker
63994e1cdb
Merge pull request #2405 from susilehtola/patch-1
Fix typo in dynamic_zarch.c
6 years ago
Martin Kroeker
496e3019bc
Merge pull request #2404 from martin-frbg/issue2395
Fix spurious application of USE_TRMM in cmake builds
6 years ago
Martin Kroeker
169be3f097
Merge pull request #2403 from martin-frbg/issue2400
Fix coretype identification of Intel Cannon Lake, Ice Lake and Goldmont
6 years ago
Martin Kroeker
6ccbb089c2
Merge pull request #2402 from gxw-loongson/develop
Avoid printing the following information on mips and mips64 when check msa
6 years ago
Martin Kroeker
59ebe3636a
Merge pull request #2399 from martin-frbg/buffersize
Make BUFFER_SIZE configurable at build time
6 years ago
Susi Lehtola
5a6bba3061
Patch out instances of Z15 in dynamic_zarch.c
There does not appear to be a Z15 kernel yet, causing link errors from the code. This patch fixes the issue.
6 years ago
Susi Lehtola
dff173e50e
Fix typo in dynamic_zarch.c
6 years ago
Martin Kroeker
7e5cbb6f35
Fix bad conditional syntax that caused spurious application of USE_TRMM
6 years ago
Martin Kroeker
303bdb673b
Fix coretype detection for Intel extended models 6 and 7
affecting Goldmont, Cannon Lake, Ice Lake autodetection
6 years ago
gxw
754433f420
Avoid printing the following information on mips and mips64 when check msa:
"unrecognized command line option ‘-mmsa’"
6 years ago
Martin Kroeker
7f0d523b42
Make BUFFER_SIZE configurable
6 years ago
Martin Kroeker
c353d8b106
Make BUFFER_SIZE configurable
6 years ago
Martin Kroeker
579be3aa9d
Add configuration option for BUFFER_SIZE
6 years ago
Martin Kroeker
449e8ea443
Merge pull request #26 from xianyi/develop
rebase
6 years ago
Martin Kroeker
3bec250cf9
Increment version to 0.3.9.dev
6 years ago
Martin Kroeker
f03dd23e90
Increment version to 0.3.9.dev
6 years ago
Martin Kroeker
fb5eb47558
Merge pull request #2398 from xianyi/develop
Update from develop in preparation of the 0.3.8 release
6 years ago
Martin Kroeker
fa93d63365
Merge branch 'release-0.3.0' into develop
6 years ago
Martin Kroeker
90e6c66a57
Merge pull request #2397 from martin-frbg/038changes
Update Changelog with changes from 0.3.8
6 years ago
Martin Kroeker
32d97330b3
Update with changes from 0.3.8
6 years ago
Martin Kroeker
29eaf4b6d7
Merge pull request #25 from xianyi/develop
rebase
6 years ago
Martin Kroeker
47c1bf7f4d
typo fixes
6 years ago
Martin Kroeker
2b55f0ad30
Merge pull request #2393 from martin-frbg/issue2388
Provide more documentation in README.md
6 years ago
Martin Kroeker
a5b32ab06c
Merge pull request #2390 from martin-frbg/pgi
Small corrections for compilation with PGI compilers
6 years ago
Martin Kroeker
50545b19d0
Update CPU and OS support and document DYNAMIC_ARCH option in README.md
prompted by #2388
6 years ago
Martin Kroeker
b3cbd60d7a
Remove PGI from list again as it is actually still not capable
6 years ago
Martin Kroeker
70199d1905
Merge pull request #2389 from Zeyiii/develop
Fix bugs in benchmark of gemv
6 years ago
Martin Kroeker
cfe63d8cc2
Remove OpenMP libraries from link list
6 years ago
Martin Kroeker
d55b10830f
Remove OpenMP libraries from link list
6 years ago
Martin Kroeker
c1c10cbb21
Merge pull request #2384 from wjc404/develop
Optimize AVX512 DGEMM (& DTRMM)
6 years ago
Martin Kroeker
5989841524
Add PGI to avx512-supporting compilers
6 years ago
Martin Kroeker
68a43db358
Fix utest compilation with PGI
6 years ago
Martin Kroeker
9694037b23
Set SUFFIX in tempfile commands, fix bad architecture option for PGI compiler in avx512 test
6 years ago
Martin Kroeker
71faa1c1a7
Merge pull request #24 from xianyi/develop
rebase
6 years ago
wjc404
3447d04eaf
Update dgemm_kernel_16x2_skylakex.c
6 years ago
wjc404
8b5cdcc64c
Update sgemm_kernel_8x4_haswell.c
6 years ago
wjc404
4e00d96a78
Update dgemm_kernel_16x2_skylakex.c
6 years ago
w00421467
ce9ea8f826
Fix another branch
6 years ago
w00421467
0b909203cb
Fix bugs in benchmark of gemv
6 years ago
wjc404
096da2f51a
Update dgemm_kernel_16x2_skylakex.c
6 years ago
wjc404
2f96a2c55b
Update trmm_R.c
6 years ago
wjc404
833bd0f8ff
Update trmm_L.c
6 years ago