- 25 2月, 2021 4 次提交
-
-
由 wangchaochaohu 提交于
cherry-pick #31068
-
由 liu zhengxi 提交于
* add get_cublas_handle() api * update format * add unittests * alter function name
-
由 qingqing01 提交于
Cherry-pick double grad for clip
-
由 tangwei12 提交于
* fix entry * fix distributed lookup table fuse case * fix entry bug at first time * move entry from paddle.fluid -> paddle.distributed * fix ut with paddle.enable_static() Co-authored-by: Nmalin10 <malin10@baidu.com> Co-authored-by: Nmalin10 <malin10@baidu.com>
-
- 24 2月, 2021 2 次提交
- 23 2月, 2021 8 次提交
-
-
由 Chen Weihang 提交于
[CustomOp] New custom operator extension mechanism in 2.0.1 Cherry-pick New custom operator basic implementation related PRs
-
由 Pei Yang 提交于
-
由 Zhong Hui 提交于
[BUG FIX] Fix softmax cross entropy overflow problem.
-
由 WangXi 提交于
* [Kunlun] Add condition_variable and notify() in BindThreadedSSAGraphExecutor (#30586) * [Kunlun] fix dead lock for exec_op_count_ (#30718) * Fix the problem that the number of ops executed by xpu is wrong (#30961) Co-authored-by: Nliuyuhui <liuyuhui@baidu.com>
-
由 Qi Li 提交于
ATT, cherry pick of #31132
-
由 Wojciech Uss 提交于
* A fix for oneDNN matmul kernel. Fixes issue #30309 (#30723) * A fix for #30309 with oneDNN 1.6
-
由 tangwei12 提交于
* test=develop, save/load, shrink Co-authored-by: NseiriosPlus <tangwei12@baidu.com> Co-authored-by: N123malin <malin10@baidu.com>
-
由 Shang Zhizhou 提交于
-
- 22 2月, 2021 2 次提交
-
-
由 Guanghua Yu 提交于
* add parameter in roi_align op * fix compatibility of ops * fix op test & cpu kernel * fix JaccardOverlap in nms
-
由 Wilber 提交于
-
- 20 2月, 2021 1 次提交
-
-
由 石晓伟 提交于
* bug fix of xpu lite engine, test=develop * xpu zero copy tensor, test=develop * revert paddle/fluid/inference/tests/api/CMakeLists.txt
-
- 19 2月, 2021 1 次提交
-
-
由 Wilber 提交于
-
- 18 2月, 2021 1 次提交
-
-
由 Jacek Czaja 提交于
-
- 10 2月, 2021 2 次提交
- 09 2月, 2021 1 次提交
-
-
由 Chengmo 提交于
* 【Paddle.Fleet】Fix brpc get hostname (#30703) * fix Brpc get hostname * fix int64 bug (#30780) fix push sparse int64 bug
-
- 07 2月, 2021 1 次提交
-
-
由 Zhou Wei 提交于
cherry-pick #29998 * Polish and Optimize the print/repr message of all layer * fix some code format
-
- 05 2月, 2021 2 次提交
-
-
由 chentianyu03 提交于
make abs support complex types cherry-pick: #30375 #30637
-
由 Shang Zhizhou 提交于
Co-authored-by: Ntianshuo78520a <707759223@qq.com>
-
- 04 2月, 2021 1 次提交
-
-
由 石晓伟 提交于
-
- 03 2月, 2021 1 次提交
-
-
由 Wilber 提交于
-
- 02 2月, 2021 2 次提交
-
-
由 alncat 提交于
* fixed compilation error on gcc 4.8.x due to the usage of isfinite (#30733) * modified conv+bn fuse pass to fix wrong mask in mask rcnn (#30704)
-
由 Shang Zhizhou 提交于
* add dla * add python api Co-authored-by: Nshangzhizhou <root@szth-rp-fanyi-opera49.szth.baidu.com> Co-authored-by: Nshangzhizhou <root@szth-rp-fanyi-opera49.szth.baidu.com>
-
- 27 1月, 2021 1 次提交
-
-
由 Wojciech Uss 提交于
Co-authored-by: NJacek Czaja <jacek.czaja@intel.com>
-
- 22 1月, 2021 1 次提交
-
-
由 Pei Yang 提交于
-
- 21 1月, 2021 1 次提交
-
-
由 QingshuChen 提交于
-
- 20 1月, 2021 3 次提交
-
-
由 AshburnLee 提交于
* Add tf32 support for A100 tensor core acceleration for cuBLAS (#28732) * Fixed an error * Fixed an error
-
由 AshburnLee 提交于
This PR is cherry-picked from PR: #29192 Function: Added TF32 switch for cuDNN. Turned on as default, turned off when users set the switch as False
-
由 Wilber 提交于
-
- 19 1月, 2021 5 次提交
-
-
由 pangyoki 提交于
Cherry pick PR #30520 . Fix error message of Inplace strategy.
-
由 Leo Chen 提交于
[cherry-pick] support layer_norm fp16 in dygraph amp (#30430)
-
由 Zhou Wei 提交于
cherry-pick #30553 fix bug of multicard grad ncclAllReduce, the gradient accumulater of parameters should be keep order, otherwsie, it will influence multicard ncclAllReduce of grad.
-
由 liym27 提交于
cherry-pick #30536
-
由 Zhen Wang 提交于
Fix the compiling error of update_loss_scaling when using cuda9.
-