- 29 1月, 2021 1 次提交
-
-
由 Jiaqi Liu 提交于
* Alias from paddle.fluid.layers.auc to paddle.static.auc (#30206) * add alias from fluid.layers.auc to static.auc * Update __init__.py * add auc into all list * alias acc, expose to users * add auc into 'all' list (#30310) * add auc into 'all' list * alias acc, expose to users * update sample code * fix paddle.static.acc and auc sample code bug, test=document_fix
-
- 27 1月, 2021 1 次提交
-
-
由 Wojciech Uss 提交于
Co-authored-by: NJacek Czaja <jacek.czaja@intel.com>
-
- 22 1月, 2021 1 次提交
-
-
由 Pei Yang 提交于
-
- 21 1月, 2021 1 次提交
-
-
由 QingshuChen 提交于
-
- 20 1月, 2021 10 次提交
-
-
由 AshburnLee 提交于
* Add tf32 support for A100 tensor core acceleration for cuBLAS (#28732) * Fixed an error * Fixed an error
-
由 guofei 提交于
动态图中Conv2D保存成预测模型时,对应的Op可能是conv2d,也可能是depthwise_conv2d,但目前的save_quantized_model接口并未考虑depthwise_conv2d情况,可能会致使out_scale的值保存错误,该PR主要是修复这个问题。
-
由 Zhen Wang 提交于
* Fix the bug in fleet amp_init. * Fix the amp_init unit test.
-
由 cnn 提交于
-
由 AshburnLee 提交于
This PR is cherry-picked from PR: #29192 Function: Added TF32 switch for cuDNN. Turned on as default, turned off when users set the switch as False
-
由 Aurelius84 提交于
[Dy2static]Fix paddle prefix in is_paddle_api (#30569) cherry-pick #30569
-
由 huangxu96 提交于
* add fleet amp.init() * add unittest for fleet_amp_init
-
由 Wilber 提交于
-
由 huangxu96 提交于
* Implemented AddQuantDequantPass in imperative quantization. * support 2.0 API such as Pool2D and ReLU
-
由 QingshuChen 提交于
-
- 19 1月, 2021 20 次提交
-
-
由 WangXi 提交于
-
由 WeiXin 提交于
原始PR:#30485,#30507
-
由 wanghuancoder 提交于
* if pybind.cc changed, generate total report
-
由 pangyoki 提交于
Cherry pick PR #30520 . Fix error message of Inplace strategy.
-
由 Leo Chen 提交于
[cherry-pick] support layer_norm fp16 in dygraph amp (#30430)
-
由 Zhou Wei 提交于
cherry-pick #30553 fix bug of multicard grad ncclAllReduce, the gradient accumulater of parameters should be keep order, otherwsie, it will influence multicard ncclAllReduce of grad.
-
由 WeiXin 提交于
完善static.load的var_list参数。 当加载的是多个小文件时,Tensor列表可以是所有加载文件中Tensor的子集。 原始PR:#30457
-
由 liym27 提交于
cherry-pick #30536
-
由 Zhen Wang 提交于
Fix the compiling error of update_loss_scaling when using cuda9.
-
由 Zhang Ting 提交于
* add 2.0 API: device_guard
-
由 hutuxian 提交于
-
由 hutuxian 提交于
-
由 hutuxian 提交于
-
由 tangwei12 提交于
* add trainers for pserver Change-Id: I99c0ab1cc427318f1f9bf8f8f5faff2b8890645d * add trainers for pserver Change-Id: I1a75793ec81ce126d07f4c47cae09b95d530bbc8
-
由 cc 提交于
-
由 taixiurong 提交于
* support transformer v2.0 * fix range op crash in dygraph xpu place
-
由 cc 提交于
-
由 liuyuhui 提交于
-
由 JZ-LIANG 提交于
-
由 LielinJiang 提交于
* update voc url
-
- 18 1月, 2021 6 次提交
-
-
由 Zhang Ting 提交于
cherry-pick #30527
-
由 lidanqing 提交于
Co-authored-by: NWojciech Uss <wojciech.uss@intel.com>
-
由 guofei 提交于
* Modify the calculation logic of LambOptimizer (#29313) * Modify the calculation logic of LambOptimizer * Modify the calculation logic of LambOptimizer * Modify the calculation logic of LambOptimizer
-
由 ceci3 提交于
* add pad and concat double grad * resolve conflict
-
由 Zhang Ting 提交于
* add fp16 support for tril_triu op (#30186) * add VecCastCUDAKernel (#30296) Co-authored-by: Nfurnace <34057289+windstamp@users.noreply.github.com>
-
由 123malin 提交于
* test=develop, fix fleet.metrics(mse, rmse, mae)
-