index
:
debian-clblast
debian/sid
upstream/latest
Debian package for CLBlast.
gspr@nonempty.org
summary
refs
log
tree
commit
diff
log msg
author
committer
range
Age
Commit message (
Expand
)
Author
2018-02-02
Fixed the XHAD documentation
Cedric Nugteren
2018-01-31
Created the API and stubs for the HAD (hadamard-product) routines
Cedric Nugteren
2018-01-29
Updated to CLBlast version 1.3.0
Cedric Nugteren
2018-01-29
Merge branch 'master' of github.com:CNugteren/CLBlast
Cedric Nugteren
2018-01-29
Fixed a compilation error of the kernel-preprocessor test under MSVC
Cedric Nugteren
2018-01-28
Updated the known issues
Cedric Nugteren
2018-01-27
Some fixes to the benchmark scripts
Cedric Nugteren
2018-01-26
Minor displaying improvements to the graph plotting scripts
Cedric Nugteren
2018-01-26
Fixed an event synchronisation issue in the batched gemm routines
Cedric Nugteren
2018-01-25
Improved the benchmark scripts; added gemmstridedbatched benchmark
Cedric Nugteren
2018-01-25
Moved some constants from global scope to a function; removed unnecessary inc...
Cedric Nugteren
2018-01-25
Changed the default number of runs for the GEMV tuner to fix issues for FP16
Cedric Nugteren
2018-01-20
Merge pull request #244 from CNugteren/kernel_selection_batched_gemm
Cedric Nugteren
2018-01-18
Made GEMM routine tuning a bit more generic in preparation of possible separa...
Cedric Nugteren
2018-01-18
Made the batched routines also chose direct/indirect kernel like the main GEM...
Cedric Nugteren
2018-01-15
Factored out the generic parts of the GEMM routine tuner
Cedric Nugteren
2018-01-14
Small improvements to benchmarking for cuBLAS
Cedric Nugteren
2018-01-11
Merge pull request #240 from CNugteren/retrieve_tuning_parameters
Cedric Nugteren
2018-01-11
Added test for the RetrieveParameters function
Cedric Nugteren
2018-01-11
Added a RetrieveParameters function to inspect tuning parameters
Cedric Nugteren
2018-01-11
Fixed bug in override parameters test
Cedric Nugteren
2018-01-11
Merge pull request #239 from CNugteren/gemm_strided_batched
Cedric Nugteren
2018-01-08
Implemented the in-direct version of the strided-batched GEMM kernel
Cedric Nugteren
2018-01-07
Implemented direct version of strided-batched GEMM kernel
Cedric Nugteren
2018-01-07
Added API and tests for new GemmStridedBatched routine
Cedric Nugteren
2018-01-06
Fixed a minor nullptr related issue in the code generator
Cedric Nugteren
2018-01-06
Prevented half-precision batched routines from failing in the tests
Cedric Nugteren
2018-01-06
Reduced duplicate code in the batched GEMM implementation
Cedric Nugteren
2018-01-06
Updated changelog and roadmap
Cedric Nugteren
2018-01-06
Fixed a vendor naming bug in the tuners and in the database
Cedric Nugteren
2018-01-06
Merge pull request #238 from CNugteren/gemm_api_with_temp_buffer
Cedric Nugteren
2018-01-06
Fixed the CUDA interface: replaced nullptr with 0
Cedric Nugteren
2018-01-06
Fixed a performance overhead in database creation: it is again a static varia...
Cedric Nugteren
2018-01-06
Added CUDA interface to get temporary-buffer size for GEMM routine
Cedric Nugteren
2018-01-04
Added a CUDA version of the GEMM temp-buffer optional argument
Cedric Nugteren
2018-01-04
Updated the generator script to automatically generate the temp-buffer code
Cedric Nugteren
2018-01-03
Updated the ROADMAP
Cedric Nugteren
2018-01-03
Added the temp-buffer to the GEMM testers and clients
Cedric Nugteren
2018-01-03
Added a queue argument to the get-size function when running the tests/clients
Cedric Nugteren
2018-01-01
Merge pull request #236 from CNugteren/trsm_compilation
Cedric Nugteren
2017-12-31
Fixed the issue with AMD's APP compiler not being able to compile the invert ...
Cedric Nugteren
2017-12-31
Revert "Added a simple test to check compilation of the invert kernels (issue...
Cedric Nugteren
2017-12-31
Revert "Added options to disable parts of the invert kernel to find out where...
Cedric Nugteren
2017-12-31
Made plotting script more flexible: extra argument to set the comparison library
Cedric Nugteren
2017-12-31
Changed the invert kernel slightly; added part1a/part1b disable-defines
Cedric Nugteren
2017-12-30
Fixed ifdef's into ifndef's
Cedric Nugteren
2017-12-30
Added options to disable parts of the invert kernel to find out where the AMD...
Cedric Nugteren
2017-12-30
Added optional temp-buffer argument to C++ interface of GEMM
Cedric Nugteren
2017-12-28
Added interface to compute the required temporary buffer size for GEMM
Cedric Nugteren
2017-12-28
Factored out argument processing from the GEMM routine
Cedric Nugteren
[prev]
[next]