1. 24 12月, 2021 1 次提交
    • C
      [pten] combine reduce_cuda codes (#38328) · 08941eda
      chentianyu03 提交于
      * combine reduce_cuda codes
      
      * support float16 in pten redcue_mean
      
      * replace ReduceCudaKernel impl with pten reduce impl
      
      * mv reduce funcs into reduce_cuda_impl
      
      * rm unsed codes and headers
      
      * mv GetReduceDim into reduce_cuda_impl
      
      * recover GetReduceDim in reduce_op.h
      
      * add new dispatch macro
      
      * fix pool op output not inited and cause transform to pten::denseTensor error
      
      * fix output tensor not initialized error
      
      * rename new dispatch macro and format code style
      
      * rm reduce_functor_op.h file
      08941eda
  2. 21 12月, 2021 1 次提交
  3. 17 12月, 2021 1 次提交
  4. 16 12月, 2021 1 次提交
  5. 13 12月, 2021 1 次提交
  6. 09 12月, 2021 1 次提交
  7. 29 11月, 2021 1 次提交
    • C
      [Pten] Add reduce mean kernel, replace with mean API (#37559) · f9e9fd19
      chentianyu03 提交于
      * add pten reduce kernel
      
      * add reduce_sum kernel
      
      * update attribute args and order
      
      * make out dtype undefined
      
      * fix empty input error
      
      * merge develop branch
      
      * rename sum as reduce function
      
      * rename sum as reduce function
      
      * fix reducekernelImpl args error
      
      * add reduce cuda kernel
      
      * modify dims type to const &
      
      * remove unsed log
      
      * fix reduce_all out eigen function error
      
      * remove unused codes
      
      * add the missing sum api define and testcase
      
      * merge develop branch
      
      * fix sum test axis value error
      
      * replace pten mean kernel with reduce_mean
      
      * revcover meam cuda to original implement
      f9e9fd19
  8. 08 9月, 2021 1 次提交
  9. 05 8月, 2021 1 次提交
    • H
      New executor dev (#34407) · 012d12b5
      hong 提交于
      * first test version
      
      * add test exec;
      
      * add data transfer; test=develop
      
      * add new exec head;
      
      * add memcpy; test=develop
      
      * add python fetch
      
      * add new test
      
      * add graph node; test=develop
      
      * remove useless new executor test; test=develop
      
      * remove gperf dependency; test=develop
      
      * fix compile bugs; test=develop
      
      * remove useless code; test=develop
      
      * remove useless code; test=develop
      
      * add uni test; test=develop
      
      * polish code; test=develop
      
      * polish code; test=develop
      
      * add interpreter cmakefile; test=develop
      
      * remove useless code; test=develop
      012d12b5
  10. 02 8月, 2021 1 次提交
    • F
      [NPU] add reduce_max (#34179) · de53f2bf
      furnace 提交于
      * [NPU] add reduce_max
      
      * [NPU] delete skipIf
      
      * [NPU] add atrrs support or check
      
      * [NPU] add attr out_dtype
      
      * [NPU] delete debug codes
      de53f2bf
  11. 02 7月, 2021 1 次提交
  12. 30 4月, 2021 1 次提交
  13. 21 4月, 2021 1 次提交
  14. 14 4月, 2021 1 次提交
  15. 17 9月, 2020 1 次提交
  16. 25 8月, 2020 1 次提交
  17. 21 8月, 2020 1 次提交
  18. 13 8月, 2020 1 次提交
  19. 09 6月, 2020 1 次提交
  20. 17 4月, 2020 1 次提交
  21. 05 4月, 2020 1 次提交
  22. 23 2月, 2020 1 次提交
  23. 31 10月, 2019 1 次提交
    • H
      GradMaker for dygraph (#19706) · 8c4573a3
      hong 提交于
      * refactor dygraph,test=develop
      
      * fix failed unittest,test=develop
      
      * polish code,test=develop
      
      * check windows ci error,test=develop
      try to fix windows ci error by np.allclose,test=develop
      
      * polish vlog and profiler, test=develop
      
      * try to fix preceding ops order,test=develop
      
      * test transformer in windows ci, test=develop
      
      * use python c-api to speed up tracer.trace,test=develop
      
      * test=develop, fix docker with paddle nccl problem
      
      * test=develop, add ut for debug string and gradient_accumulator
      
      * test=develop, add tests for layer/gradient_accumulator/prepared_op
      
      * test=develop, fix complie error for test_prepared_op
      
      * test=develop, add more ut for dygraph
      
      * test=develop, create API.spec for dygraph api change
      
      * optimize grad maker; test=develop
      
      * optimize grad maker
      
      * test
      
      * grad make optim; test=develop
      
      * fix unittest bugs; test=develop
      
      * add dygraph grad op maker and split_op
      
      * grad op maker refactor; test=develop
      
      * add dygraph grad maker; test=develop
      
      * fix op deformable_conv_v1_op bug; test=develop
      
      * fix deformable_conv prroi pool bugs;
      
      * fix new op grad op maker bug; test=develop
      
      * fix split by ref bug; test=develop
      
      * fix dygraph auto prune bug; test=develop
      
      * fix test_trace bug; test=develop
      
      * fix fused emb seq pool bug; test=develop
      
      * remove useless code in op_desc file; test=develop
      
      * remove useless code, StrVarBaseNode; test=develop
      
      * fix review issues; test=develop
      
      * fix rank_loss grad maker; test=develop
      
      * remove flag in VarBase; test=develop
      
      * fix distributed_notify_op compile bug ; test=develop
      
      * fix reshape op double grad; test=develop
      
      * fix expand as op; test=develop
      
      * add impertive type_defs.h for demo_train; test=develop
      
      * fix inference lib cmake; test=develop
      
      * fix inference lib; test=develop
      
      * fix infernce_lib; test=develop
      
      * fix inference cmake; test=develop
      
      * fix inference lib; test=develop
      
      * fix inference lib; test=develop
      
      * remove condition dygraph grad maker, modify local name; test=develop
      
      * fix split grad maker bug; test=develop
      
      * fix pyramid_op bug; test=develop
      
      * change travis time out limit; test=develop
      
      * restore travis; test=develop
      
      * change timeout limit; test=develop
      8c4573a3
  24. 28 10月, 2019 1 次提交
  25. 12 10月, 2019 1 次提交
  26. 10 10月, 2019 1 次提交
  27. 03 10月, 2019 1 次提交
  28. 26 9月, 2019 1 次提交
  29. 05 9月, 2019 1 次提交
  30. 14 5月, 2019 1 次提交
    • L
      Double backward reduce mean (#17372) · 5d1ac41b
      lvmengsi 提交于
      * test=develop, double backward reduce_mean
      
      * add comment. test=develop
      
      * fix format. test=develop
      
      * rename GradGrad -> DoubleGrad. test=develop
      
      * fix op_use_default_grad_op_maker.spec. test=develop
      5d1ac41b
  31. 12 4月, 2019 2 次提交
  32. 16 11月, 2018 1 次提交
    • W
      Refine operator cmake (#14413) · a2d9b344
      Wu Yi 提交于
      * wip simplify operator framework
      
      * wip
      
      * wip
      
      * done test=develop
      
      * clean test=develop
      
      * fix test=develop
      
      * fix deps test=develop
      
      * fix cpu build test=develop
      
      * fix tensorrt build test=develop
      
      * fix tests test=develop
      
      * fix test=develop
      
      * fix cpu build test=develop
      a2d9b344
  33. 22 7月, 2018 2 次提交
  34. 07 6月, 2018 1 次提交
  35. 05 6月, 2018 1 次提交
  36. 23 5月, 2018 1 次提交
    • W
      Enhance reduce op (#10708) · 8655904b
      whs 提交于
      * Enhance reduce op for multi dims.
      
      * Uncomment some unitest.
      
      * Uncomment unitest.
      
      * Remove unused code.
      
      * Fix infershape and python wrapper.
      
      * Add more examples.
      
      * Fix l2_normalize.
      
      * Fix normalization_wrapper.
      
      * Polish code.
      1. Rename unitest function.
      2. Rename const variable.
      8655904b
  37. 19 4月, 2018 1 次提交
  38. 07 3月, 2018 1 次提交