1. 31 3月, 2020 1 次提交
  2. 06 3月, 2020 1 次提交
    • X
      Develop nlp patch (#3059) · 8425c924
      xiaogang 提交于
      * fix: fix nlp ops input and output type
      * fix: add elementwise x_dims>y_dims case
      8425c924
  3. 26 2月, 2020 1 次提交
  4. 14 1月, 2020 1 次提交
  5. 11 12月, 2019 1 次提交
  6. 07 12月, 2019 1 次提交
    • J
      Support mask_rcnn (#2484) · cf31b835
      juncaipeng 提交于
      * add arm split lod tensor, test=develop
      
      * add arm merge lod tensor, test=develop
      
      * update split merge lod tensor, test=develop
      
      * add reduce_prob op, test=develop
      
      * support mask_rcnn succeed, test=develop
      cf31b835
  7. 22 11月, 2019 1 次提交
    • H
      add NHWC NCHW transform, test=develop (#2381) · b1e4f4fd
      HappyAngel 提交于
      * add nhwc to nchw
      
      * add layout in funcs
      
      * change layout as extra, test=develop
      
      * change make, test=develop
      
      * use template class method to update layout NNCHHW and NHWC transform, test=develop
      
      * fix cmake error, set layout to extra, test=develop
      
      * fix test_layout_compute_arm test, its extra
      
      * layout is extra, test=develop
      
      * fix error in kernels/arm/layout_comput.cc when register kernel, DataLayout must be NCHW, test=develop
      
      * delete extra note, test=develop
      
      * delete extra test
      
      * delete layout_test, test=develop
      
      , its in tests/math/layout_comput_test
      
      * delete extrat test, test=develop
      b1e4f4fd
  8. 28 10月, 2019 1 次提交
    • H
      [LITE][XPU] initial support for XPU (#2202) · ac1b2f9f
      hong19860320 提交于
      * Initial support for XPU
      * Fix compiling errors of XPU
      * Move XPU op kernel bridges from backends to kernels to fix deps order
      * Change the namespace and directory of XPU bridges
      * Add XPU SDK
      * Fix header files and namespace of XPU SDK
      * Add unit tests for relu and conv2d ops
      * Restore the modification of paddle_api_test
      * Supports simple model which contains only a relu layer
      * Add compiling scripts for XPU
      * Fix compiling errors of XPU
      * Add comments for XPU LoadModel and BuildModel
      ac1b2f9f
  9. 27 10月, 2019 1 次提交
    • H
      model dynamic library tailoring (#2256) · b16917a4
      huzhiqiang 提交于
      * add shell file to automatically build and collect publish result test=develop
      * modify API inference of model_optimize_tool and add option for tiny&full publish test=develop
      b16917a4
  10. 11 10月, 2019 1 次提交
    • Y
      [LITE][OPENCL] support image2d type (#2158) · 74074146
      Yuan Shuai 提交于
      * [LITE][OPENCL] support image2d. test=develop
      
      * add context changed with consider image*. test=develop
      
      * add layout, relu image kernels. test=develop
      
      * replace image_data with data, mutable_image_data with mutable_data, test=develop
      
      * comment unused var. test=develop
      
      * remove unused var. test=develop
      74074146
  11. 27 9月, 2019 1 次提交
    • Z
      can run yolov3 fp32 on cuda devices (#2092) · c4b5e32c
      Zhaolong Xing 提交于
      * add conv int8 support(in condition which the input or output channel not be the times of 4)
      add add_kernel for cuda.
      
      * can run yolov3 fp32
      test=develop
      
      * 1. fix bug with yolov3 run
      test=develop
      c4b5e32c
  12. 19 9月, 2019 1 次提交
  13. 18 9月, 2019 1 次提交
  14. 11 9月, 2019 2 次提交
  15. 06 9月, 2019 1 次提交
    • Z
      add cudnn conv fp32, int8 support (#1974) · 23d83c04
      Zhaolong Xing 提交于
      * paddle lite cuda init
      can run model with leaky_relu
      
      * add the missing file.
      test=develop
      
      * add the load from memory interface.
      test=develop
      
      * refine this pr. fix comments
      fix ci error
      test=develop
      
      * conv impl
      fp32:
      conv, conv+bais, conv+bias+relu, conv+bias+leaky_relu
      
      int8:
      conv, conv+bais+relu(int8 or fp32 output), conv+bias+leaky_relu(int8 or fp32 output)
      
      can run conv+ bias+relu using cxx_api
      test=develop
      
      * move the lite/cuda/math to backends/cuda/math
      test=develop
      23d83c04
  16. 16 8月, 2019 1 次提交