1. 12 1月, 2022 2 次提交
  2. 11 1月, 2022 1 次提交
    • Z
      【PTen】Add dot and matmul grad kernel in pten (#38713) · be817719
      zyfncg 提交于
      * refactor matmul directory in pten
      
      * fix merge conflict
      
      * add dot_grad kernel
      
      * add dot_grad kernel in pten
      
      * add matmul_grad kernel
      
      * update the code
      
      * delete useless code in fluid
      
      * fix some bug of running matmul grad kernel
      
      * fix merge conflict
      
      * refactor some code
      
      * refactor code
      be817719
  3. 06 1月, 2022 3 次提交
  4. 05 1月, 2022 1 次提交
  5. 04 1月, 2022 2 次提交
  6. 31 12月, 2021 2 次提交
  7. 28 12月, 2021 2 次提交
  8. 27 12月, 2021 2 次提交
  9. 26 12月, 2021 1 次提交
    • C
      [PTen] Move copy kernel impl (#38421) · 73819658
      Chen Weihang 提交于
      * add register general kernel marco
      
      * move copy kernel impl
      
      * revert needless change
      
      * polish details
      
      * fix xpu compil faild
      
      * fix xpu compile failed
      
      * polish format
      73819658
  10. 23 12月, 2021 3 次提交
  11. 22 12月, 2021 2 次提交
  12. 21 12月, 2021 1 次提交
  13. 20 12月, 2021 2 次提交
  14. 17 12月, 2021 1 次提交
  15. 16 12月, 2021 3 次提交
  16. 14 12月, 2021 2 次提交
  17. 10 12月, 2021 1 次提交
  18. 09 12月, 2021 2 次提交
  19. 08 12月, 2021 1 次提交
  20. 07 12月, 2021 1 次提交
    • Y
      [Pten]Move func from kernel_context.h into kernel_context.cc (#37804) · bfa0d7f3
      YuanRisheng 提交于
      * add inplace op adaptation
      
      * optimize inplace logic and fix bugs when run kernel that has args of vector<DenseTensor>
      
      * move func in kernel_context.h into kernel_context.cc
      
      * refactor logic that transform variable to densetensor
      
      * fix bugs when compile
      
      * update func name
      
      * fix bugs when run windows-ci
      bfa0d7f3
  21. 02 12月, 2021 1 次提交
  22. 30 11月, 2021 1 次提交
  23. 29 11月, 2021 1 次提交
    • C
      [Pten] Add reduce mean kernel, replace with mean API (#37559) · f9e9fd19
      chentianyu03 提交于
      * add pten reduce kernel
      
      * add reduce_sum kernel
      
      * update attribute args and order
      
      * make out dtype undefined
      
      * fix empty input error
      
      * merge develop branch
      
      * rename sum as reduce function
      
      * rename sum as reduce function
      
      * fix reducekernelImpl args error
      
      * add reduce cuda kernel
      
      * modify dims type to const &
      
      * remove unsed log
      
      * fix reduce_all out eigen function error
      
      * remove unused codes
      
      * add the missing sum api define and testcase
      
      * merge develop branch
      
      * fix sum test axis value error
      
      * replace pten mean kernel with reduce_mean
      
      * revcover meam cuda to original implement
      f9e9fd19
  24. 25 11月, 2021 2 次提交