Martin Kroeker
|
c62f8e2c01
|
Prevent compiler attempts to use k0 as mask register
|
2022-02-23 20:12:20 +01:00 |
Wangyang Guo
|
bb1c4fa5bd
|
sbgemm: cooperlake: prefetch A & B
|
2021-09-07 21:30:46 +08:00 |
Wangyang Guo
|
7a2d1601ec
|
sbgemm: cooperlake: unroll core loop by 2
|
2021-09-07 21:30:46 +08:00 |
Wangyang Guo
|
45fdf951b6
|
sbgemm: cooperlake: reorder ptr increase for performance
|
2021-09-07 21:30:46 +08:00 |
Wangyang Guo
|
cece3541ab
|
sbgemm: cooperlake: fix bug in m64n12
|
2021-09-07 21:30:46 +08:00 |
Wangyang Guo
|
9df0953cde
|
sbgemm: cooperlake: kernel works for NN
|
2021-09-07 21:30:45 +08:00 |
Wangyang Guo
|
2ec9f3a8aa
|
sbgemm: cooperlake: change kernel size to 16x4
|
2021-09-07 21:30:45 +08:00 |