提交 · 66d2ae25e399bb32959d3eb770dd5330633f889a · PaddlePaddle / Paddle-Lite

22 11月, 2019 1 次提交
- Y
  [ARM] add sgemmc4 common and small kernel, support for winograd, test=develop (#2471) · 66d2ae25
  由 yiicy 提交于 11月 22, 2019
```
* unfinish sgemmc4

* finish armv8 sgemmc4

* arm add sgemmc4 with deal with remain

* [ARM] add sgemmc4 small kernel, test=develop
```
  66d2ae25
18 9月, 2019 1 次提交

fix bias quantize error && fix clang build error (#2049) · 81dffbe8

由 Xiaoyang LI 提交于 9月 18, 2019

* fix gemm_int8, gemv-int8 and conv-int8 math function, add float bias

* change conv impl

* neon int8 kernel support float bias

* arm compute kernel support float bias

* add math_test target

* add tensor utils for testing, fix sgemm ut error

* add gemm_int8 unit test, support float bias

* fix build script

* add conv compute unit test for arm

* fix build script, test=develop

* fix fp32 dw conv3x3s1, test=develop

* add fp32 dw conv3x3s1, test=develop

* add armv7 fp32 dw conv3x3s1, test=develop

* add fp32 depthwise conv3x3s2, test=develop

* fix fp32 conv3x3 depthwise build error, test=develop

* fix gemm_like conv trans weights error, test=develop

* fix int8 depthwise conv3x3 error, test=develop

* turn on all test for arm fp32 conv, test=develop

* fix int8 conv1x1 error

* fix int8 direct conv3x3s1 error, test=develop

* fix int8 direct conv3x3s2, test=develop

* turn on all test for arm int8 conv, test=develop

* fix int8 fc error, change mobilenetv1-int8 ground-truth result to fluid, test=develop

* remove debug info, strip ut binary, test=develop

* fix conv compute error, test=develop

* change Init() to ReInitWhenNeeded(), test=develop

* fix code style, test=develop

* remote engine_test, test=develop

* fix building server tests error, test=develop

* fix sdot clang build error, test=develop

* fix sgemm ut timeout error, test=develop

* fix clang build error, test=develop

* turn off math basic test due to ci time out, test=develop

* fix conv_int8 ut error, test=develop

81dffbe8

03 9月, 2019 1 次提交
- H
  
  create backends directory and move hardware backends into it (#1954) · 31ee212a
  由 huzhiqiang 提交于 9月 03, 2019
  
  31ee212a
29 8月, 2019 3 次提交

Add yolo_box_cuda multiclass_nms_host kernel. (#1908) · de43e479

由 Wilber 提交于 8月 29, 2019

* add yolo_box_compute cuda

* move multiclass_nms(arm) to host

* add lod in scale op

* add yolo_box_cuda cmake config

* modify shuffle_channel_fuse and transpose_softmax_transpose_fuse to support run ssd model. test=develop

* reshape and transpose op don't have xshape output.

* modify yolo_box_compute_cuda, use tensor to manage cuda memory test=develop

* add yolo_box use kernel test=develop

de43e479

L

add stack op and add reduce_mean op and their unit tests (#1888) · 20001636
由 liu zhengxi 提交于 8月 29, 2019

20001636

ad ops for faster rcnn, including affine_channel, anchor_generator,... · 53b05ce8

由 juncaipeng 提交于 8月 29, 2019

ad ops for faster rcnn, including affine_channel, anchor_generator, generate_proposals and roi_align (#1895)

* add ops for faster rcnn

* disable test for generate_proposals and roi_align, test=develop

* remove .swp file

* remove log in tensor slice

* finish the unit test for roi_align, test=develop

53b05ce8

16 8月, 2019 1 次提交
- Y
  
  publish lite (#1800) · 699d6cd0
  由 Yan Chunwei 提交于 8月 16, 2019
  
  699d6cd0
12 3月, 2019 1 次提交
- H
  
  Optimize vector-matrix and matrix-vector multiply · 1d078c3d
  由 hjchen2 提交于 3月 12, 2019
  
  1d078c3d
10 3月, 2019 1 次提交
- H
  
  update · dd575b09
  由 hjchen2 提交于 3月 10, 2019
  
  dd575b09
16 12月, 2018 1 次提交
- H
  
  Fix softmax · 33e1e2dd
  由 hjchen2 提交于 12月 16, 2018
  
  33e1e2dd
15 12月, 2018 1 次提交
- H
  
  Refactor softmax to speed up and fix bug · 3793beef
  由 hjchen2 提交于 12月 15, 2018
  
  3793beef
10 10月, 2018 1 次提交
- H
  
  Fix code style for protobuf-c header file · 5e4a1781
  由 hjchen2 提交于 10月 10, 2018
  
  5e4a1781
09 10月, 2018 1 次提交
- H
  
  Refine · 244297e8
  由 hjchen2 提交于 10月 09, 2018
  
  244297e8
23 5月, 2018 1 次提交
- W
  
  fix #224 · 7e1f55c5
  由 wangliu 提交于 5月 23, 2018
  
  7e1f55c5