index
:
debian-clblast
debian/sid
upstream/latest
Debian package for CLBlast.
gspr@nonempty.org
summary
refs
log
tree
commit
diff
log msg
author
committer
range
path:
root
/
scripts
Age
Commit message (
Collapse
)
Author
2018-09-16
Merge branch 'master' into convgemm_multi_kernel
Cedric Nugteren
2018-08-05
Added an option to compile the Netlib API with static OpenCL device and context
Cedric Nugteren
2018-07-29
Removed complex numbers support for CONVGEMM
Cedric Nugteren
2018-07-29
Merge branch 'master' into CLBlast-267-convgemm
Cedric Nugteren
2018-07-13
Added tuning results for HD Graphics 6000 Broadwell GT3
Cedric Nugteren
2018-05-09
Updated the documentation for convgemm to include data layout (NCHW)
Cedric Nugteren
2018-05-06
Added convgemm skeleton, test infrastructure, and first reference implementation
Cedric Nugteren
2018-05-05
Added interface of batched convolution as GEMM
Cedric Nugteren
2018-04-15
Updated tuning results for the Skylake ULT GT2 GPU with the new kernel
Cedric Nugteren
2018-04-10
Made it possible to add tuning parameters to the database using the script
Cedric Nugteren
2018-04-10
Fixed a bug in the compression part of the database script
Cedric Nugteren
2018-04-08
Extended the maximum number of tuning parameters from 14 to 16
Cedric Nugteren
2018-04-07
Fixed a python3 import error issue with the database script
Cedric Nugteren
2018-03-27
merged
kodonell
2018-03-27
got the generator thing working
kodonell
2018-03-11
Merge pull request #262 from CNugteren/CLBlast-237-tuning-api
Cedric Nugteren
CLBlast #237: Tuning API
2018-03-10
Made benchmarking script also work for complex numbers
Cedric Nugteren
2018-03-10
Updated the documentation for the tuner API
Cedric Nugteren
2018-03-10
Fixed a few things for the new tuning API
Cedric Nugteren
2018-03-03
Fixed some small issues regarding PR#253
Cedric Nugteren
2018-03-03
Added C API for getting GEMM temp buffer size
sivagnanamn
2018-02-25
Generated PyCLBlast docstrings
Cedric Nugteren
2018-02-25
Some style improvements in the pyclblast code generator
Cedric Nugteren
2018-02-25
Added API documentation for two missing C++ functions
Cedric Nugteren
2018-02-24
Renamed the API documentation
Cedric Nugteren
2018-02-21
Fixed duplication of parameter descriptions by the doc generator
Kirill Mavreshko
2018-02-18
Prepared PyCLBlast for release as a package on PyPi
Cedric Nugteren
2018-02-18
Added all other level 1/2/3 routines to pyclblast
Cedric Nugteren
2018-02-18
Added GEMM to the Python wrapper
Cedric Nugteren
2018-02-14
First agenerated version (clblastXswap only for now) of the pyclblast wrapper
Cedric Nugteren
2018-02-02
Fixed the XHAD documentation
Cedric Nugteren
2018-01-31
Created the API and stubs for the HAD (hadamard-product) routines
Cedric Nugteren
2018-01-27
Some fixes to the benchmark scripts
Cedric Nugteren
2018-01-26
Minor displaying improvements to the graph plotting scripts
Cedric Nugteren
2018-01-25
Improved the benchmark scripts; added gemmstridedbatched benchmark
Cedric Nugteren
2018-01-14
Small improvements to benchmarking for cuBLAS
Cedric Nugteren
2018-01-11
Added a RetrieveParameters function to inspect tuning parameters
Cedric Nugteren
2018-01-07
Added API and tests for new GemmStridedBatched routine
Cedric Nugteren
2018-01-06
Fixed a minor nullptr related issue in the code generator
Cedric Nugteren
2018-01-06
Merge pull request #238 from CNugteren/gemm_api_with_temp_buffer
Cedric Nugteren
GEMM API with optional temp buffer
2018-01-06
Added CUDA interface to get temporary-buffer size for GEMM routine
Cedric Nugteren
2018-01-04
Added a CUDA version of the GEMM temp-buffer optional argument
Cedric Nugteren
2018-01-04
Updated the generator script to automatically generate the temp-buffer code
Cedric Nugteren
2017-12-31
Made plotting script more flexible: extra argument to set the comparison library
Cedric Nugteren
2017-12-28
Added interface to compute the required temporary buffer size for GEMM
Cedric Nugteren
2017-12-27
Split the database into multiple small compilation units
Cedric Nugteren
2017-12-20
Made plotting script more resilient to missing data
Cedric Nugteren
2017-12-20
Added tuning results for Apple AMD Radeon Pro 580
Cedric Nugteren
2017-12-20
Added try-except to database script parser to skip invalid files
Cedric Nugteren
2017-11-20
Made the database script properly handle multiple entries for a single device
Cedric Nugteren
[next]