1. 05 9月, 2022 1 次提交
  2. 19 7月, 2022 1 次提交
    • J
      Added pad3d and pad2d FP32 FWD oneDNN kernels (#43990) · 2792b8de
      jakpiase 提交于
      * Piotrek's changes for pad3d
      
      * my changes
      
      * first version of pad3d, single copy, unnecessary reads
      
      * optimized pad3d kernel
      
      * test upadte
      
      * removed magic numbers
      
      * added support for pad2d
      
      * reverted two files
      
      * reverted one old change
      
      * added support for Paddings tensor
      
      * CI fix
      
      * CI fix
      
      * fixed timeout of tests
      
      * fixed typo
      
      * changes to GetKernelTypeForVar
      
      * Revert "changes to GetKernelTypeForVar"
      
      This reverts commit 469106115c49682b25038a666fd71bd4a10fb66b.
      
      * added AsExtra() to pad2d
      Co-authored-by: NPiotr Paturej <piotr.paturej@intel.com>
      2792b8de
  3. 02 7月, 2022 1 次提交
    • L
      unify cpu context, part2 (#44012) · 755438a7
      Leo Chen 提交于
      * fix init()
      
      * delete test_device_context
      
      * replace CPUDeviceContext with CPUContext
      
      * fix test_scalar
      
      * remove dot_op.cc
      
      * fix compile
      755438a7
  4. 26 6月, 2022 1 次提交
  5. 05 6月, 2022 1 次提交
  6. 20 2月, 2022 1 次提交
  7. 19 2月, 2022 1 次提交
    • A
      [Pten]Unify paddle/pten::framework::ddim into pten::ddim (#39614) · 2fe04264
      Aurelius84 提交于
      * Unify paddle/pten::framework::ddim into pten::ddim
      
      * fix paddle namespace
      
      * compile sucessfully
      
      * fix npu src file
      
      * fix conflict
      
      * fix conflict
      
      * fix tensorrt compiler error
      
      * fix conflict
      
      * fix conflict
      
      * fix tesst file conflict
      
      * fix conflict
      
      * fix mlu file conflict
      
      * fix mlu file conflict
      
      * fix cinn header file conflict
      
      * fix conflict
      
      * fix conflict
      
      * fix conflict
      
      * fix conflict
      2fe04264
  8. 11 2月, 2022 1 次提交
  9. 25 5月, 2020 1 次提交
  10. 20 4月, 2020 1 次提交
  11. 25 3月, 2020 1 次提交
  12. 09 3月, 2020 1 次提交
  13. 28 2月, 2020 1 次提交
  14. 02 2月, 2020 1 次提交
  15. 08 1月, 2020 1 次提交
  16. 26 12月, 2019 1 次提交
  17. 25 12月, 2019 1 次提交
  18. 31 10月, 2019 1 次提交
    • H
      GradMaker for dygraph (#19706) · 8c4573a3
      hong 提交于
      * refactor dygraph,test=develop
      
      * fix failed unittest,test=develop
      
      * polish code,test=develop
      
      * check windows ci error,test=develop
      try to fix windows ci error by np.allclose,test=develop
      
      * polish vlog and profiler, test=develop
      
      * try to fix preceding ops order,test=develop
      
      * test transformer in windows ci, test=develop
      
      * use python c-api to speed up tracer.trace,test=develop
      
      * test=develop, fix docker with paddle nccl problem
      
      * test=develop, add ut for debug string and gradient_accumulator
      
      * test=develop, add tests for layer/gradient_accumulator/prepared_op
      
      * test=develop, fix complie error for test_prepared_op
      
      * test=develop, add more ut for dygraph
      
      * test=develop, create API.spec for dygraph api change
      
      * optimize grad maker; test=develop
      
      * optimize grad maker
      
      * test
      
      * grad make optim; test=develop
      
      * fix unittest bugs; test=develop
      
      * add dygraph grad op maker and split_op
      
      * grad op maker refactor; test=develop
      
      * add dygraph grad maker; test=develop
      
      * fix op deformable_conv_v1_op bug; test=develop
      
      * fix deformable_conv prroi pool bugs;
      
      * fix new op grad op maker bug; test=develop
      
      * fix split by ref bug; test=develop
      
      * fix dygraph auto prune bug; test=develop
      
      * fix test_trace bug; test=develop
      
      * fix fused emb seq pool bug; test=develop
      
      * remove useless code in op_desc file; test=develop
      
      * remove useless code, StrVarBaseNode; test=develop
      
      * fix review issues; test=develop
      
      * fix rank_loss grad maker; test=develop
      
      * remove flag in VarBase; test=develop
      
      * fix distributed_notify_op compile bug ; test=develop
      
      * fix reshape op double grad; test=develop
      
      * fix expand as op; test=develop
      
      * add impertive type_defs.h for demo_train; test=develop
      
      * fix inference lib cmake; test=develop
      
      * fix inference lib; test=develop
      
      * fix infernce_lib; test=develop
      
      * fix inference cmake; test=develop
      
      * fix inference lib; test=develop
      
      * fix inference lib; test=develop
      
      * remove condition dygraph grad maker, modify local name; test=develop
      
      * fix split grad maker bug; test=develop
      
      * fix pyramid_op bug; test=develop
      
      * change travis time out limit; test=develop
      
      * restore travis; test=develop
      
      * change timeout limit; test=develop
      8c4573a3
  19. 28 10月, 2019 1 次提交
  20. 22 7月, 2019 1 次提交
  21. 15 4月, 2019 1 次提交
  22. 02 4月, 2019 1 次提交
  23. 12 12月, 2018 1 次提交
  24. 30 11月, 2018 1 次提交
  25. 31 8月, 2018 1 次提交
    • W
      Add pad2d op. (#12950) · e10aa80f
      whs 提交于
      * Add pad2d op.
      
      * Add unitest and python api.
      
      * Fix cuda op kernel.
      
      * Fix python api.
      
      * Fix python api.
      
      * Update API.spec.
      
      * Fix python api
      e10aa80f