- 07 4月, 2021 3 次提交
-
-
由 zhang wenhui 提交于
* Ascend rc (#30483) * Fix compilcation on CANN20.1 and older (#30494) Fix compilcation on CANN20.1 and older * Add distribution supported (#30578) Add distribution supported * Build praser for Hcom* operators (#30627) Build praser for Hcom* operators * Pass device_ids info from launch to trainer. (#30632) Pass device_ids info from launch to trainer * Add Hccl program group (#30642) Add Hccl program group * Add startup bash files of test_ascend_group. (#30645) Add startup bash files of test_ascend_group * cleanup (#30646) cleanup test_ascend_group.py * [Feature] Build parser to support distributed training (#30658) [Feature] Build parser to support distributed training * fix compilation on ascend-20.1 (#30722) fix compilation on ascend-20.1 * Dev/fix ascend string (#30749) Dev/fix ascend string * code style (#30781) code style * Merge ascend_optimizer and ascend_parser. (#30776) Merge ascend_optimizer and ascend_parser. * Ascendrc add converted op : [range/equal/range/uniform_random/expand/squeeze], fix cast op bug (#30797) Ascendrc add converted op : [range/equal/range/uniform_random/expand/squeeze], fix cast op bug * Add paddle ascend distribution training supported (#30796) Add paddle ascend distribution training supported * pass cxx_flags to gloo cmake (#30857) * Destroy session first. (#30954) Destroy session first. * merge * fix, test=develop * fix, test=develop * fix style, test=develop * fix, test=develop * fix * fix log fatal, test=develop * fix enforce style, test=develop * fix, test=develop * fix, test=develop * fix rccl, test=develop * fix test, test=develop * fix, test=develop * fix, test=develop * fix, test=develop * fix node_num, test=develop * fix ids str, test=develop * fix ids str, test=develop * fix ids str, test=develop * fix, test=develop * fix, test=develop * fix, test=develop * fix, test=develop * fix, test=develop * fix, test=develop * fix, test=develop * fix, test=develop * fix style code, test=develop * fix style code, test=develop * fix style code, test=develop * fix style code, test=develop Co-authored-by: Nhutuxian <hutuxian2011@sina.cn> Co-authored-by: Ngongweibao <weibao.gong@gmail.com> Co-authored-by: NVoid Main <voidmain1313113@gmail.com> Co-authored-by: NLeo Chen <chenqiuliang@baidu.com> Co-authored-by: Ndingsiyu <18369187719@163.com> Co-authored-by: NOleNet <olenet@126.com>
-
由 Ouyang Chao 提交于
* improve performance of DepthwiseConv(NWHC)
-
由 tangwei12 提交于
* add PullSparseValue for pull sparse * fix bug for PullSparseValue * add test mode in lookuptable * revert API change * add comment for is_training
-
- 06 4月, 2021 3 次提交
-
-
由 wuhuanzhou 提交于
-
由 Kqnonrime 提交于
* fix two error message * fix two error message * fix error * fix error * fix error * fix error
-
由 ronnywang 提交于
-
- 03 4月, 2021 1 次提交
-
-
由 jiangcheng 提交于
-
- 02 4月, 2021 2 次提交
-
-
由 ronnywang 提交于
-
由 niuliling123 提交于
* add leaky_relu forward and backward in activation_op.cu
-
- 01 4月, 2021 6 次提交
-
-
由 Qi Li 提交于
-
由 hutuxian 提交于
-
由 zlsh80826 提交于
* add anchor generator op plugin * add anchor generator unit_test * remove dbg info * remove redundant line * replace assertion with paddle enforce * dynamic plugin replaces assertion with paddle enforce * anchor generator support dynamic shape on spatial axis * anchor generator test with fp16, dynamic shape * add anchor generator test all * add back main * reduce test input size to not exceed the timelimit of ci * change super to InferencePassTest for python2 compatibility * reuse paddle operator anchor generator * move creator construct to header with default * add cuda ifdef * reduce line * change super to InferencePassTest for python2 compatibility * fix anchor generator fp16 serialize setting * split unittest from test_all * restrict anchor generator input format before version 7234 * anchor generator only support greater than trt7.1 * change min_graph_size to 2 * min_graph size to 3 if dynamic shape * reduce dynamic shape size to avoid trt search tactic too long to exceed time limit * remove anchor from fetch list * anchor generator support all trt version * fix memory not allocated but if serialized
-
由 Zhang Zheng 提交于
-
由 Zhang Zheng 提交于
-
由 kuizhiqing 提交于
* new group * ci compatible fix * assert nccl
-
- 31 3月, 2021 7 次提交
-
-
由 Kqnonrime 提交于
* fix one error massage * fix a error message * new fix three error messages * new fix three error messages * new fix some error * new fix one error message
-
由 tianshuo78520a 提交于
-
由 wuhuanzhou 提交于
* update eigen version to f612df27, test=develop * fix compilation error, test=develop * remove patch command in eigen, test=develop * fix compilation error caused by call Eigen function with float16 and bfloat16, test=develop * fix unittest error, test=develop * fix unittest error caused by precision, test=develop * remove patch files used by old version eigen, test=develop
-
由 wuhuanzhou 提交于
* update compilation with C++14, test=develop * fix compilation error in eigen, test=develop
-
由 Thunderbrook 提交于
* fix split core * format
-
由 taixiurong 提交于
-
由 furnace 提交于
* bugfix for warpctc * fix warpctc commit id * fix warpctc commit id * fix warpctc commit id * fix warpctc commit id * fix warpctc commit id * fix WARPCTC_WITH_HIP invalid * Add logs to find out why can not dlopen libwarpctc.so * fix warpctc commit id * fix unit test test_warpctc_op * Optime failed log for dlopen * Optime failed log for dlopen * Delete extra changes * fix warpctc commit id * fix warpctc commit id * Add is_compiled_with_rocm for test_warpctc_op * fix warpctc commit id * Cancel optimize dlopen failed reason, move to next pr, due to it makes windows ci failed * Cancel optimize dlopen failed reason, move to next pr, due to it makes windows ci failed * Cancel optimize dlopen failed reason, move to next pr, due to it makes windows ci failed * fix code style problems
-
- 30 3月, 2021 2 次提交
-
-
由 Jiawei Wang 提交于
-
由 jakpiase 提交于
-
- 29 3月, 2021 3 次提交
-
-
由 niuliling123 提交于
-
由 tianshuo78520a 提交于
-
由 liym27 提交于
-
- 26 3月, 2021 2 次提交
-
-
由 cc 提交于
* Use layer to calculate output scale * add backward for moving_average_abs_max_scale and save output scales to op's attr
-
由 tianshuo78520a 提交于
* delete include framework.pb.h * fix error
-
- 25 3月, 2021 2 次提交
-
-
由 Chen Weihang 提交于
* polish two error messages * polish details
-
由 niuliling123 提交于
-
- 24 3月, 2021 3 次提交
-
-
由 winter-wang 提交于
-
由 Wojciech Uss 提交于
* fix cache key in concat oneDNN kernel * key simplified
-
由 ronnywang 提交于
-
- 23 3月, 2021 2 次提交
-
-
由 niuliling123 提交于
* add relu forward kernel and backward kernel
-
由 Qi Li 提交于
-
- 22 3月, 2021 1 次提交
-
-
由 arlesniak 提交于
-
- 21 3月, 2021 2 次提交
-
-
由 ronnywang 提交于
-
由 Ouyang Chao 提交于
-
- 19 3月, 2021 1 次提交
-
-
由 Jacek Czaja 提交于
-