1. 04 2月, 2020 1 次提交
    • Y
      [ARM] 5x5dw and sgemv support fuse activation, test=develop (#2797) · 42bbd157
      yiicy 提交于
      * refactor 5x5s1 dw conv armv8, test=develop
      
      * [ARM] refactor depthwise conv 5x5s1, and support relu6, leakey relu, test=develop
      
      * [ARM] sgemv support fuse relu6 and leakey relu,test=develop
      
      * [ARM] reduce some conv ut case, test=develop
      
      * [ARM] fix 5x5dw conv pick kernel bug, test=develop
      
      * fix code style, test=develop
      
      * [ARM] fix sgemv fuse relu6 bug, test=develop
      
      * [ARM] fix fp32 5x5s1 dw bug, test=develop
      
      * [ARM] fix fp32 5x5 dw conv pick kernel bug, test=develop
      42bbd157
  2. 03 2月, 2020 2 次提交
  3. 19 1月, 2020 2 次提交
  4. 17 1月, 2020 2 次提交
  5. 16 1月, 2020 1 次提交
  6. 15 1月, 2020 1 次提交
  7. 14 1月, 2020 2 次提交
  8. 02 1月, 2020 2 次提交
  9. 31 12月, 2019 1 次提交
    • H
      [LITE][NPU][XPU] Refine the registration and implementation of op bridges (#2700) · fb668935
      hong19860320 提交于
      * Fix the compiling error which occurs when specify the ddk_root path and build for huawei NPU.
      
      * Refine the registration of op bridges and make it similar to the registration of op and kernel.
      
      * Refine the interfaces of the graph and node for op bridges, and support creating constant and data node automatically according to the attribute 'persistable' of the target tensor.
      
      * Add the unit test of the scale and softmax op bridge for NPU.
      fb668935
  10. 26 12月, 2019 1 次提交
  11. 25 12月, 2019 1 次提交
    • Y
      [X86] Polish the implementation of fc and imporve the unittest (#2656) · a0f01efa
      Yiqun Liu 提交于
      * Remove GEMM padding in fc_compute.
      test=develop
      
      * Write a common ParallelFor function to run the for loop in parallel.
      
      * Add the codes of padding GEMM back in fc.
      
      * Refine the code of fc when padding_weight is false to avoid the definition of temporary Tensor.
      
      * Refine the unit test of fc and add testing case of padding and parallel.
      test=develop
      
      * Enable more test cases in common fc unittest, including padding and parallel for x86 target.
      
      * Remove the fc test under kernels/x86.
      test=develop
      
      * Disable relu in test of fc for non-x86 target.
      test=develop
      
      * Change the eps of arm.
      test=develop
      a0f01efa
  12. 24 12月, 2019 1 次提交
  13. 23 12月, 2019 2 次提交
    • H
    • H
      [lite][arm]add conv relu6 and leaky_relu in conv_dw_3x3s2, test=develop (#2618) · e659e4ab
      HappyAngel 提交于
      * fix conv 2-pad to 4-pad
      
      * fix compute conv shape
      
      * fix pad, test=develop
      
      * change conv_depthwise_3x3s1_fp.cc name to conv3x3s1p01_depthwise_fp32.cc to distinguish between conv3x3s1_depthwise_fp32.cc
      
      * delete printf note in conv3x3s1, test=develop
      
      * delete printf note, test=develop
      
      * delete gem_sdot.h, test=develop
      
      it is coped from __gemm_sdot_meta_.h
      
      * update compute padding, test=develop
      
      * fix padding size, must be 2 or 4. test=develop
      
      * fix format in operators/conv_op.cc, test=develop
      
      * change #if 0 to #if 1, test=develop
      
      * put 2-pad to 4-pad in AttachImpl, test=develop
      
      * fix clang-format error inn tests/math/connv_compute_test, test=develop
      
      * fix x86 test result error, test=develop
      
      * add asymmetric padding test case in liite/tests/math/conv_compute.cc, test=develop
      
      * change paddings type to support dynamically modify, test=develop
      
      * fix x86 build error in connv_compute_test, test=develop
      
      * fix opencl build error, test=develop
      
      * fix oopencl build error, test=develop
      
      * fix  opencl/conv_compute build error, test=develop
      
      * fix  opencl/conv_compute build error, test=develop
      
      * fix format in kernels/opencl/conv_computte_ttest,test=develop
      
      * fix build error, test=develop
      
      fix build error in kernels/x86/conv_compute.h
      
      * fix ccompute shape error in ooperators/conv_op.h, test=develop
      
      * add conv_reelu6 and conv leaky_relu in conv_3x3s1_direct
      
      * add conv_relu6 in c1, c2, c4,test=develop
      
      * fix conflict in conv_bloock_utils.h, test=develop
      
      * add relu6 and leankyrelu in conv_3x3s1_dw
      
      * add conv_3x3s1px_dw relu6 and leaky_relu fusion, test=develop
      
      * fix conflict in tests/math/conv_compute_arm, test=develop
      
      * fix build error in winograd arm, test=develop
      
      * channge act_param as pointer in conv_block_tuils.h, test=develop
      
      * fix winograd in no equal 4-padding compute error, test=develop
      
      * add conv relu6 and leaky_relu in conv_dw_3x3s2, test=develop
      
      * fix format, test=develop
      
      * fix format in conv_block_utils, test=develop
      
      * move updatePadding from conv_op.cc to conv_op.h, test=develop
      
      * fix format conv_op.h, test=develop
      
      * fix buuilde error in conv_oop.h, test=develop
      
      * remove flag_relu parameter in conv_3x3_depthwise, test=develop
      e659e4ab
  14. 20 12月, 2019 1 次提交
  15. 19 12月, 2019 2 次提交
  16. 18 12月, 2019 1 次提交
  17. 17 12月, 2019 3 次提交
  18. 16 12月, 2019 3 次提交
  19. 15 12月, 2019 1 次提交
  20. 13 12月, 2019 1 次提交
  21. 12 12月, 2019 1 次提交
    • X
      [LITE][OPENCL] Add conv2d_1x1 opencl kernel (#2591) · 06e754c9
      xiebaiyuan 提交于
      * add opencl conv1x1 image impl and unit test pass with relu & bias,
      add layout_compute --> buffer2image float32 --> with unit test pass
      suite checked test for more situation , test=develop
      
      * add opencl conv1x1 image impl and unit test pass with relu & bias,
      add layout_compute --> buffer2image float32 --> with unit test pass
      suite checked test for more situation , test=develop
      
      * fix white space cpp lint , test=develop
      06e754c9
  22. 11 12月, 2019 1 次提交
  23. 10 12月, 2019 1 次提交
  24. 09 12月, 2019 2 次提交
  25. 07 12月, 2019 1 次提交
    • J
      Support mask_rcnn (#2484) · cf31b835
      juncaipeng 提交于
      * add arm split lod tensor, test=develop
      
      * add arm merge lod tensor, test=develop
      
      * update split merge lod tensor, test=develop
      
      * add reduce_prob op, test=develop
      
      * support mask_rcnn succeed, test=develop
      cf31b835
  26. 04 12月, 2019 1 次提交
  27. 03 12月, 2019 1 次提交
  28. 30 11月, 2019 1 次提交