kseniyazaytseva
|
46de7c8a2b
|
Merge remote-tracking branch 'origin/risc-v-new-tests' into new-tests
|
1 year ago |
Martin Kroeker
|
b1ae777afb
|
Merge pull request #4497 from sergei-lewis/dev/slewis/zaxpy
Fix axpy test hangs when n==0. Reenable zaxpy_vector kernel for C910V.
|
2 years ago |
Sergei Lewis
|
ff1523163f
|
Fix axpy test hangs when n==0. Reenable zaxpy_vector kernel for C910V.
|
2 years ago |
Martin Kroeker
|
ba3bfe85ee
|
Merge pull request #4495 from martin-frbg/update-gensymbol
Update gensymbol with recently added CBLAS interfaces and LAPACK/LAPACKE functions
|
2 years ago |
Martin Kroeker
|
93872f4681
|
drop the ?laqz? symbols for now (not translatable by f2c)
|
2 years ago |
Martin Kroeker
|
83bec51355
|
Update with recently added CBLAS interfaces and LAPACK/LAPACKE functions
|
2 years ago |
Martin Kroeker
|
974f29c4e9
|
Merge pull request #4494 from ChipKerchner/fixPower10CPUID
Make sure CPU ID works for all POWER_10 conditions
|
2 years ago |
Chip Kerchner
|
d408ecedba
|
Add environment variable to display coretype for dynamic arch.
|
2 years ago |
Martin Kroeker
|
a96a04ee61
|
Merge pull request #4493 from martin-frbg/issue4475-3
Fix incompatible pointer types in the declarations of C/ZAXPBY
|
2 years ago |
Chip Kerchner
|
ac6b4b7aa4
|
Make sure CPU ID works for all POWER_10 conditions
|
2 years ago |
Martin Kroeker
|
500ac4de5e
|
fix incompatible pointer types
|
2 years ago |
Martin Kroeker
|
b3fa16345d
|
fix prototype for c/zaxpby
|
2 years ago |
kseniyazaytseva
|
cfabc48190
|
Update rotg tests
|
2 years ago |
kseniyazaytseva
|
ec5cfe3bc8
|
Fix invalid tests
|
2 years ago |
kseniyazaytseva
|
ff10e6b6dc
|
Fix zero step tests
|
2 years ago |
Martin Kroeker
|
e9cfb7fd30
|
Merge pull request #4491 from martin-frbg/fixup-4488
fix sbgemm bfloat16 conversion errors introduced in PR 4488
|
2 years ago |
Martin Kroeker
|
e9f480111e
|
fix sbgemm bfloat16 conversion errors introduced in PR 4488
|
2 years ago |
Martin Kroeker
|
22b487b622
|
Merge pull request #4488 from martin-frbg/issue4475-2
Separate the interface for SBGEMMT from GEMMT
|
2 years ago |
Martin Kroeker
|
818bf30628
|
Merge pull request #4490 from ChipKerchner/missingCPUIDsForAIX
Add missing CPU ID definitions for old versions of AIX.
|
2 years ago |
Martin Kroeker
|
344763331a
|
Merge pull request #4484 from martin-frbg/lapack981
Rescale input vector more often in C/ZLARFGP (Reference-LAPACK PR 981)
|
2 years ago |
Chip Kerchner
|
08ce6b1c1c
|
Add missing CPU ID definitions for old versions of AIX.
|
2 years ago |
Martin Kroeker
|
fb99fc2e6e
|
fix type conversion warnings
|
2 years ago |
Martin Kroeker
|
08e479f956
|
Merge pull request #4487 from ErnstPeng/feature-branch
Optimized zgemm kernel 8x4 LASX, 4x4 LSX and cgemm kernel 8x4 LSX for LoongArch
|
2 years ago |
Martin Kroeker
|
d4db6a9f16
|
Separate the interface for SBGEMMT from GEMMT due to differences in GEMV arguments
|
2 years ago |
pengxu
|
fe3da43b7d
|
Optimized zgemm kernel 8*4 LASX, 4*4 LSX and cgemm kernel 8*4 LSX for LoongArch
|
2 years ago |
Martin Kroeker
|
e5d2725e5a
|
Merge pull request #4185 from XiWeiGu/mips_enable_msa
MIPS: Enable MSA
|
2 years ago |
Martin Kroeker
|
479e4af089
|
Rescale input vector more often to minimize relative error (Reference-LAPACK PR 981)
|
2 years ago |
Martin Kroeker
|
a4fde2c5ac
|
Merge pull request #4451 from martin-frbg/overflow_reset
Reset "buffer management structure overflowed" state and free auxiliary struct on blas_shutdown
|
2 years ago |
Martin Kroeker
|
b537528feb
|
Merge pull request #4480 from XiWeiGu/loongarch64-fixed-{s/d}amin-lsx
LoongArch64: Fixed {s/d}amin LSX optimization
|
2 years ago |
Martin Kroeker
|
bc7154a80d
|
Merge pull request #4482 from martin-frbg/issue4476
Fix missing NO_AVX2 fallback for SapphireRapids in DYNAMIC_ARCH
|
2 years ago |
Martin Kroeker
|
6d8a273cca
|
Handle zero increment(s) in C910V ?AXPBY (#4483)
* Handle zero increment(s)
|
2 years ago |
Martin Kroeker
|
dbcf4f8b7d
|
Merge pull request #4479 from XiWeiGu/loongarch-opt-axpby
Loongarch opt axpby
|
2 years ago |
Martin Kroeker
|
dc802dd637
|
Merge pull request #4474 from ChipKerchner/sgemmIncopy_PR
Vectorize in-copy packing/copying for SGEMM - up to 4X faster.
|
2 years ago |
Martin Kroeker
|
e307675222
|
Merge pull request #4478 from martin-frbg/issue4475
Fix incompatible pointer type in BFLOAT16 GEMMT
|
2 years ago |
Martin Kroeker
|
033168cdf0
|
Merge pull request #4481 from martin-frbg/cpuid_riscv
Update lowercase cpunames for RISC-V
|
2 years ago |
Martin Kroeker
|
a29f91ae9a
|
Merge pull request #4471 from ChipKerchner/fixMakefileAIXOpenMP
Fix Makefiles to support OpenMP on AIX for xlc (clang) with xlf.
|
2 years ago |
Martin Kroeker
|
e61d96303d
|
Fix missing NO_AVX2 fallback for SapphireRapids
|
2 years ago |
Martin Kroeker
|
d02c61e82e
|
Update lowercase cpunames for RISC-V
|
2 years ago |
Martin Kroeker
|
7228c708d7
|
Merge pull request #4461 from markdryan/cpuid_riscv64_crash
Fix two issues with cpuid_riscv64.c
|
2 years ago |
gxw
|
adde725321
|
LoongArch64: Fixed {s/d}amin LSX optimization
|
2 years ago |
gxw
|
7bc93d95a1
|
LoongArch64: Opt {c/z}axpby
|
2 years ago |
gxw
|
1e1f487dc7
|
LoongArch64: Fixed {s/d}axpby
|
2 years ago |
gxw
|
3597827c93
|
utest: add axpby
|
2 years ago |
Martin Kroeker
|
68d354814f
|
Fix incompatible pointer type in BFLOAT16 mode
|
2 years ago |
Martin Kroeker
|
3848d4e9f4
|
Merge pull request #4477 from martin-frbg/c910caxpy
Temporarily disable the CAXPY/ZAXPY kernels for C910V to workaround a CI hang
|
2 years ago |
Martin Kroeker
|
4d8dee508c
|
temporarily disable the CAXPY/ZAXPY kernels
|
2 years ago |
Martin Kroeker
|
27816fa929
|
Merge pull request #4472 from sergei-lewis/dev/slewis/merge-from-riscv
Merge risc-v branch to develop
|
2 years ago |
kseniyazaytseva
|
b6949ce74c
|
add axpyc to cmake build
|
2 years ago |
kseniyazaytseva
|
441339104f
|
fix test ext cmake build
|
2 years ago |
kseniyazaytseva
|
f68e9989c4
|
Remove zero rows/columns matcopy tests
|
2 years ago |