index
:
debian-clblast
debian/sid
upstream/latest
Debian package for CLBlast.
gspr@nonempty.org
summary
refs
log
tree
commit
diff
log msg
author
committer
range
Age
Commit message (
Expand
)
Author
2018-01-06
Fixed a performance overhead in database creation: it is again a static varia...
Cedric Nugteren
2018-01-06
Added CUDA interface to get temporary-buffer size for GEMM routine
Cedric Nugteren
2018-01-04
Added a CUDA version of the GEMM temp-buffer optional argument
Cedric Nugteren
2018-01-04
Updated the generator script to automatically generate the temp-buffer code
Cedric Nugteren
2018-01-03
Updated the ROADMAP
Cedric Nugteren
2018-01-03
Added the temp-buffer to the GEMM testers and clients
Cedric Nugteren
2018-01-03
Added a queue argument to the get-size function when running the tests/clients
Cedric Nugteren
2017-12-30
Added optional temp-buffer argument to C++ interface of GEMM
Cedric Nugteren
2017-12-28
Added interface to compute the required temporary buffer size for GEMM
Cedric Nugteren
2017-12-28
Factored out argument processing from the GEMM routine
Cedric Nugteren
2017-12-28
Refactored GEMM code in preparation of separate temp-buffer computation
Cedric Nugteren
2017-12-27
Merge pull request #234 from CNugteren/database_compilation_split
Cedric Nugteren
2017-12-27
Split the database into multiple small compilation units
Cedric Nugteren
2017-12-26
Made the database-vector a non-static member
Cedric Nugteren
2017-12-24
Fixes for the CUDA backend of CLBlast
Cedric Nugteren
2017-12-24
Fixed linking of the preprocessor test for MSVC
Cedric Nugteren
2017-12-24
Added a note that the ArrayFire Jenkins servers are down, being switched to b...
Cedric Nugteren
2017-12-23
Fixed unused variable warnings showing up with Clang
Cedric Nugteren
2017-12-23
Updated the tuning results for the IvyBridge M GT2 GPU
Cedric Nugteren
2017-12-23
Added defines to disable OpenCL deprecation warnings
Cedric Nugteren
2017-12-23
Fixed a warning under MSVC
Cedric Nugteren
2017-12-23
Merge pull request #232 from CNugteren/feature/more_tuners
Cedric Nugteren
2017-12-23
Now calling main TRSV routine again to fix compilation in MSVC
Cedric Nugteren
2017-12-23
Split the invert kernel in two parts to prevent error C1091 in MSVC 2013
Cedric Nugteren
2017-12-23
Updated the database to use the new TRSV and Invert tuners
Cedric Nugteren
2017-12-23
Added TRSV block-size tuner
Cedric Nugteren
2017-12-21
Fixed AppVeyor issue
Cedric Nugteren
2017-12-21
Fixed AppVeyor issue
Cedric Nugteren
2017-12-21
Merge branch 'master' into feature/more_tuners
Cedric Nugteren
2017-12-20
Made plotting script more resilient to missing data
Cedric Nugteren
2017-12-20
Added tuning results for Apple AMD Radeon Pro 580
Cedric Nugteren
2017-12-20
Added try-except to database script parser to skip invalid files
Cedric Nugteren
2017-12-19
Added skeleton for a tuner for the invert kernel
Cedric Nugteren
2017-12-18
Reformatted tuning code to make compilation faster
Cedric Nugteren
2017-12-17
Fixed an issue with the tuner: it was using platform vendor rather than devic...
Cedric Nugteren
2017-12-17
Merge pull request #230 from CNugteren/kernel_preprocessor
Cedric Nugteren
2017-12-17
Removed all ARM Mali tuning results; re-added Mali-T760 and Mali-T628 results...
Cedric Nugteren
2017-12-17
Fixed an unnecessary overflow issue on 32-bit systems
Cedric Nugteren
2017-12-16
Updated the known issues
Cedric Nugteren
2017-12-10
Fixed for error C1091 in MSVC 2013
Cedric Nugteren
2017-12-10
Split GEMM kernel in 4 files instead of 3 due to MSVC 2013 string length limit
Cedric Nugteren
2017-12-10
Updated roadmap: completed pre-processor implementation
Cedric Nugteren
2017-12-10
Fixed a missing include
Cedric Nugteren
2017-12-10
Fixed an issue in the tuners to prevent error -14 from persisting (CL_EXEC_ST...
Cedric Nugteren
2017-12-10
Fixed an Android compilation issue
Cedric Nugteren
2017-12-09
Completed kernel modifications for pre-processor of all other kernels
Cedric Nugteren
2017-12-09
Made the pre-processor run by default for ARM and Qualcomm GPUs
Cedric Nugteren
2017-12-09
Modified the direct GEMM kernel to support array-to-register promotion
Cedric Nugteren
2017-12-09
Reformatted GEMM kernel to support array-to-register promotion
Cedric Nugteren
2017-12-09
Fixed defines parsing and substituting in pre-processor; fixed some variable ...
Cedric Nugteren
[next]