Paddle

Commit Graph

Author	SHA1	Message	Date
Chen Weihang	1a304e6c06	[Complex] Add support for complex grad accumulated (#29889 ) * add support for complex grad accumulated * add unittest for coverage * update test dtype * remove useless blank line	5 years ago
chentianyu03	8f45d14263	add complex64 and complex128 type; add +-/@ and slice opreator for c… (#29199 ) add complex64 and complex128 type; add +-/@ and slice opreator for complex types add test cases for complex elementwise, matmul and getitem unittest * add test cases for complex types * add test cases for complex matmul unittest	5 years ago
wanghuancoder	df43905f12	use iwyu clean include (#27267 ) * use iwyu clean include, test=develop, test=win * compilation error, test=develop * fix compilation error2, test=develop * fix compilation error3, test=develop * fix compilation error4, test=develop * fix compilation error5, test=develop * fix compilation error6, test=develop * fix compilation error7, test=develop * fix compilation error8, test=develop * fix compilation error8, test=develop * fix compilation error10, test=develop * fix compilation error11, test=develop	5 years ago
wangchaochaohu	3eacced950	[cuda11 support] add support for cublas load of same function name (parameter diff) (#26963 )	5 years ago
Chen Weihang	172d4ecb6c	remove WITH_DSO compile option (#25444 )	5 years ago
Yiqun Liu	ecfddebbef	Add the implementation of inverse (#23310 )	6 years ago
Guo Sheng	a8c0fb4e86	Add cholesky_op (#23543 ) * Add cholesky_op forward part. test=develop * Complete cholesky_op forward part. test=develop * Add cholesky_op backward part. test=develop * Complete cholesky_op backward part. test=develop * Refine cholesky_op error check and docs. test=develop * Add grad_check unit test for cholesky_op. test=develop * Fix sample code in cholesky doc. test=develop * Refine some error messages of cholesky_op. test=develop * Refine some error messages of cholesky_op. test=develop * Remove unused input in cholesky_grad. test=develop * Remove unused input in cholesky_grad. test=develop * Fix stream for cusolverDnSetStream. test=develop * Update PADDLE_ENFORCE_CUDA_SUCCESS from cholesky_op to adapt to latest code. test=develop * Add CUSOLVER ERROR in enforce.h test=develop * Fix the missing return value in cholesky. test=develop	6 years ago
littletomatodonkey	1c08a2136e	test=develop, add addmm op (#23384 ) add addmm op	6 years ago
Yiqun Liu	42b5bec6f9	Integrate NVRTC to support compiling CUDA kernel at runtime (#19422 ) * Add the dynamic load of nvrtc, and support runtime compiling of CUDA kernel using nvrtc. test=develop * Call CUDA driver api to launch the kernel compiled by nvrtc. test=develop * Disable for mac and windows. test=develop * Refine the codes to support manually specified num_threads and workload_per_thread. test=develop * Refine the CUDA kernel to support large dims. test=develop	6 years ago
chengduozh	f7847ca6a3	fix cublas warp error test=develop	7 years ago
chengduo	00b9e9a135	Refine cublas to support CUBLAS_TENSOR_OP_MATH (#13929 ) * refine cublase test=develop * code refine * refine cublas * add GEMME_EX * add enable_cublas_tensor_op_math doc and add cublasCall test=develop * fix CublasCall for cuda version test=develop * fix error test=develop * fix GEMM_EX to be compatible with gcc 4.8 test=develop * add GEMM_EX test=develop * to compatiable with gcc4.8 test=develop	7 years ago
dzhwinter	2d00e65819	namespace issue (#13543 ) * flags * "follow comment"	7 years ago
dzhwinter	e23ddf6ae4	status (#12764 )	7 years ago
yuyang18	c5115950a8	Use static for dlsym	7 years ago
Yu Yang	3d53631bad	Make dyload strictly use the same ABI in header	8 years ago
Kexin Zhao	7ed457e77a	Fix cuda 7.5 error with cublas GEMM (#9811 ) * fix gemm error for cuda 7.5 * fix version number	8 years ago
Yi Wang	e185502ebe	Fix cpplint errors with paddle/fluid/platform/dynload (#9715 ) * Update source files. * Update headers * Update * Update * Update * Update * Fix a CMake dependency	8 years ago
Kexin Zhao	d00bd9eb72	Update the cuda API and enable tensor core for GEMM (#9622 ) * change from hgemm to gemmEx * fix cpplint	8 years ago
kexinzhao	90215b7844	Add float16 GEMM math function on GPU (#8695 ) * test cpu float16 data transform * add isnan etc * small fix * fix containsNAN test error * add data_type transform GPU test * add float16 GPU example * fix error * fix GPU test error * initial commit * fix error * small fix * add more gemm fp16 tests * fix error * add utility function	8 years ago
qingqing01	24509f4af9	Fix the grammar in copyright. (#8403 )	8 years ago
Yi Wang	fc374821dd	Correct #include path	8 years ago
Yi Wang	90648f336d	Move file to fluid/; Edit CMakeLists.txt	8 years ago

22 Commits (1a304e6c069391dd543a3f95a8f9b0826c3e7b93)