This is the release note of v6.0.0rc1. See here for the complete list of solved issues and merged PRs.
Highlights
- CUDA 10.1 and cuDNN 7.5 are now supported. CuPy also starts to compile for compute capability 7.5 for Turing GPUs.
- After this release, the master branch is switched to the development of v7 series. v6.0.0 will continue developing at the
v6branch.
New Features
- New RNN API introduced in cuDNN v7.2 (#1609)
- Add
diffandunwrap(#1933, thanks @a2kiti!) - Support fusion feature of
copytomethod (#1983) - Add
lu_factorandlu_solvetocupyx.scipy.linalg(#2051, thanks @msakai!)
Enhancements
- More support
__cuda_array_interface__(#2058) - Fix fp16 issue in batch normalization (#2060)
- Keep backward compatibility on
cupy.cudnn.batch_normalization_forward_training(#2072) - Check if gc module is still available (#2116)
- Support CUDA 10.1 + cuDNN 7.5 + Turing (#2117)
Performance Improvements
- Do exact type comparison instead of
isinstancefornumpy.dtype(#2016) - Improve
_routines_manipulation(#2038)
Bug Fixes
- Fix assigning from complex to float (#1911)
- Ensure that sparse matrix shapes are always a tuple of
int(#1943, thanks @grlee77!) - Avoid sharing handles between threads (#2053)
- Check array contiguity in
copy(#2075) - Support dltensor with strides of NULL (#2097, thanks @crcrpar!)
- Fix assertion error in
_Chunk.split(#2103, thanks @liwt31!)
Code Fixes
- Use single quote (#2049)
- Use
assertinhelper.pyinstead ofself.assertXXX(#2077) - Remove unused variable (#2105, thanks @crcrpar!)
- Reorganize import in
device.pyx(#2121)
Documentation
- Document
__array_function__(#1979) - Add NCCL v2.4 support to docs (#2065)
- Add strides to docstring of ndarray (#2096, thanks @crcrpar!)
Installation
- Use deep copy in setting up RPATH (#2073)