Bine Brank
|
e8939b3d30
|
sve trsmRN and trsmRT
|
2022-01-10 20:42:20 +01:00 |
Martin Kroeker
|
5188aede5d
|
Merge pull request #3511 from martin-frbg/cmakeutils
Fix handling of ifdef/ifndef in CMAKE
|
2022-01-10 09:12:52 +01:00 |
Martin Kroeker
|
a9e297e476
|
Fix handling of ifdef/ifndef
|
2022-01-09 23:31:59 +01:00 |
Bine Brank
|
098672b51b
|
add trsm_kernel_LT_sve
|
2022-01-09 20:11:47 +01:00 |
Bine Brank
|
be7e55880c
|
sve trsm_kernel_LN
|
2022-01-09 19:40:04 +01:00 |
Martin Kroeker
|
499ae5e8f7
|
Merge pull request #3510 from martin-frbg/issue3505
Fix recent SkylakeX/DYNAMIC_ARCH DGEMM breakage
|
2022-01-09 14:50:51 +01:00 |
Martin Kroeker
|
b6b024232d
|
Merge pull request #3508 from snadampal/v1_n2
OpenBLAS: aarch64: Add neoverse-v1/n2 architecture specifics
|
2022-01-09 14:50:26 +01:00 |
Martin Kroeker
|
2573ccfb2e
|
make DYNAMIC_ARCH option available to getarch_2nd/param.h
|
2022-01-08 23:50:34 +01:00 |
Martin Kroeker
|
f1ac59f200
|
Forward DYNAMIC_ARCH option to Makefile.prebuild
|
2022-01-08 23:48:58 +01:00 |
Martin Kroeker
|
15d4b37913
|
SkylakeX: match parameters to dgemm kernels for dyn/non-dyn
|
2022-01-08 23:48:13 +01:00 |
Sunita Nadampalli
|
19c8f615dc
|
OpenBLAS: aarch64: Add neoverse-v1/n2 architecture specifics
|
2022-01-07 00:28:17 +00:00 |
Bine Brank
|
cbcea149f0
|
update contributors
|
2022-01-06 10:29:35 +01:00 |
Bine Brank
|
bb33446b40
|
fix makefile.L3
|
2022-01-06 10:26:11 +01:00 |
Bine Brank
|
f33543d029
|
combine zchemm into single file
|
2022-01-05 14:42:37 +01:00 |
Bine Brank
|
0c91d043ae
|
adapt CMake for SVE
|
2022-01-05 14:36:39 +01:00 |
Bine Brank
|
39ab219704
|
sve copy functions for cgemm chemm zsymm
|
2022-01-05 09:12:22 +01:00 |
Bine Brank
|
18102ae8c3
|
add cgemm ctrmm sve kernels
|
2022-01-05 09:09:18 +01:00 |
Bine Brank
|
87537b8c55
|
modify sve zgemmcopy kernels
|
2022-01-05 09:07:28 +01:00 |
Bine Brank
|
d30157d891
|
update configuration of kernels for A64FX and ARMV8SVE
|
2022-01-05 09:00:54 +01:00 |
Bine Brank
|
07fa6fa3b1
|
configure Makefile for sve
|
2022-01-05 08:57:51 +01:00 |
Bine Brank
|
2e2c02b762
|
fix sve ztrmm kernel
|
2022-01-04 14:42:07 +01:00 |
Bine Brank
|
68c414d3a6
|
ztrmm sve copy functions
|
2022-01-04 14:40:59 +01:00 |
Bine Brank
|
ce329ab686
|
add sve zhemm copy routines
|
2022-01-03 15:56:05 +01:00 |
Bine Brank
|
0140373802
|
add sve ztrmm
|
2022-01-02 19:15:33 +01:00 |
Martin Kroeker
|
ecf034b250
|
Merge pull request #3502 from jgillis/develop
Fix cmake crosscompilation for core2 target
|
2022-01-01 12:12:32 +01:00 |
Martin Kroeker
|
f8b1ca5039
|
Merge pull request #3504 from martin-frbg/issue3503
Guard against omp_get_num_places returning zero
|
2022-01-01 11:43:17 +01:00 |
Martin Kroeker
|
b329e45288
|
Guard against omp_get_num_places returning zero
|
2022-01-01 00:46:23 +01:00 |
Bine Brank
|
f7b6912868
|
ztrmm sve copy kernels
|
2021-12-30 21:00:16 +01:00 |
jgillis
|
ea3db69faa
|
Fix cmake crosscompilation for core2 target
Missing HAVE_SSE* cmake variables cause cc.cmake to forget about `-msse*` flags
|
2021-12-29 22:50:20 +01:00 |
Bine Brank
|
40b14e4957
|
fix zgemm kernel
|
2021-12-29 11:42:04 +01:00 |
Martin Kroeker
|
ee823b6ed9
|
Merge pull request #3500 from martin-frbg/osx_dyn_xerbla
Ensure that the right xerbla gets included in OSX DYNAMIC_ARCH builds
|
2021-12-28 22:54:27 +01:00 |
Martin Kroeker
|
6cae44d4f7
|
Ensure that the right xerbla gets included in OSX DYNAMIC_ARCH builds
|
2021-12-28 19:06:55 +01:00 |
Martin Kroeker
|
a06b4aff52
|
Merge pull request #3496 from yuanhec/develop
Fixed MSA enabled optimization on Loongson-3A4000
|
2021-12-28 18:51:56 +01:00 |
yuanhecai
|
9d455b1b09
|
Merge remote-tracking branch 'upstream/develop' into develop
|
2021-12-27 09:50:57 +08:00 |
Bine Brank
|
6ec4aab875
|
zgemm sve copy routines
|
2021-12-26 17:05:46 +01:00 |
Bine Brank
|
878064f394
|
sve zgemm kernel
|
2021-12-26 08:44:05 +01:00 |
Bine Brank
|
683a7548bf
|
added macros for sve zgemm kernels
|
2021-12-25 11:46:41 +01:00 |
Martin Kroeker
|
7b146e590c
|
fix function typecast
|
2021-12-24 20:01:52 +01:00 |
Martin Kroeker
|
e9a0e52201
|
fix function typecast
|
2021-12-24 20:00:50 +01:00 |
yuanhecai
|
2db0b2e445
|
Fixed MSA enabled optimization on Loongson-3A4000
|
2021-12-23 20:29:42 +08:00 |
Martin Kroeker
|
253670383f
|
Merge pull request #3491 from gxw-loongson/develop
loongarch64: Optimize dgemm_kernel
|
2021-12-22 08:34:12 +01:00 |
Martin Kroeker
|
9809931eb4
|
clean up unused variables and unreachable statements
|
2021-12-21 18:53:55 +01:00 |
Martin Kroeker
|
6b407a16cb
|
fix function typecasts
|
2021-12-21 18:51:28 +01:00 |
Martin Kroeker
|
aecb4a5e8d
|
fix function typecasts
|
2021-12-21 18:50:22 +01:00 |
Martin Kroeker
|
c49d46f25f
|
fix function typecast
|
2021-12-21 18:49:18 +01:00 |
Martin Kroeker
|
64365c919e
|
fix function typecasts
|
2021-12-21 18:47:35 +01:00 |
Martin Kroeker
|
d1ee6ff73f
|
fix function typecasts
|
2021-12-21 18:45:28 +01:00 |
Martin Kroeker
|
07fe5b19a4
|
typecast function pointers
|
2021-12-21 12:31:54 +01:00 |
Bine Brank
|
e3c9947c0f
|
prepare kernel for sve zgemm
|
2021-12-21 11:19:27 +01:00 |
gxw
|
8d9b9c6b2a
|
loongarch64: Optimize dgemm_kernel
|
2021-12-21 09:33:06 +08:00 |