提交 · 80e4172e06fc47d4a4fb5611043b01e56a502125 · PaddlePaddle / Paddle-Lite

19 9月, 2019 3 次提交
- T
  Bug fix for model save and load (#1992) · 8efbdc66
  由 TianXiaogang 提交于 9月 19, 2019
```
* fix: fix model parser and save bug

* style: delete debug code

* fix: fix light_predictor program run model with subblock bug
```
  8efbdc66
- H
  modify dynamic library: libpaddle_cxx_api.so (#2057) · fcdb7081
  由 huzhiqiang 提交于 9月 19, 2019
```
(1)modify tiny publish so to make it excutable
(2)modify the bug of compiling in armlinux
(3)change 3 full publish so into 2
```
  fcdb7081
- S
  
  [Java API] add getVersion() interface to get c++ lib's version information test=develop (#2063) · 71b05c94
  由 sangoly 提交于 9月 19, 2019
  
  71b05c94
18 9月, 2019 1 次提交

fix bias quantize error && fix clang build error (#2049) · 81dffbe8

由 Xiaoyang LI 提交于 9月 18, 2019

* fix gemm_int8, gemv-int8 and conv-int8 math function, add float bias

* change conv impl

* neon int8 kernel support float bias

* arm compute kernel support float bias

* add math_test target

* add tensor utils for testing, fix sgemm ut error

* add gemm_int8 unit test, support float bias

* fix build script

* add conv compute unit test for arm

* fix build script, test=develop

* fix fp32 dw conv3x3s1, test=develop

* add fp32 dw conv3x3s1, test=develop

* add armv7 fp32 dw conv3x3s1, test=develop

* add fp32 depthwise conv3x3s2, test=develop

* fix fp32 conv3x3 depthwise build error, test=develop

* fix gemm_like conv trans weights error, test=develop

* fix int8 depthwise conv3x3 error, test=develop

* turn on all test for arm fp32 conv, test=develop

* fix int8 conv1x1 error

* fix int8 direct conv3x3s1 error, test=develop

* fix int8 direct conv3x3s2, test=develop

* turn on all test for arm int8 conv, test=develop

* fix int8 fc error, change mobilenetv1-int8 ground-truth result to fluid, test=develop

* remove debug info, strip ut binary, test=develop

* fix conv compute error, test=develop

* change Init() to ReInitWhenNeeded(), test=develop

* fix code style, test=develop

* remote engine_test, test=develop

* fix building server tests error, test=develop

* fix sdot clang build error, test=develop

* fix sgemm ut timeout error, test=develop

* fix clang build error, test=develop

* turn off math basic test due to ci time out, test=develop

* fix conv_int8 ut error, test=develop

81dffbe8

17 9月, 2019 1 次提交
- S
  [Cxx API] add build-in version info (#2047) · e7eea682
  由 sangoly 提交于 9月 17, 2019
```
* [Cxx API] add build-in version info

* update: add version.h.in template
```
  e7eea682
16 9月, 2019 1 次提交
- H
  
  add dynamic library for tiny and full publish (#2036) · 7e6a9030
  由 huzhiqiang 提交于 9月 16, 2019
  
  7e6a9030
13 9月, 2019 1 次提交
- Z
  
  Add the memory optim pass (#2018) · 89b3026c
  由 Zhaolong Xing 提交于 9月 13, 2019
  
  89b3026c
12 9月, 2019 2 次提交
- Y
  
  integrate model_optimize_tool compilation to build.sh (#2033) · 3d9028de
  由 Yan Chunwei 提交于 9月 12, 2019
  
  3d9028de
- W
  add unsqueeze and range op (x2paddle) (#1988) · 3c08f676
  由 Wilber 提交于 9月 12, 2019
```
* add unsqueeze and range op. modify concat op test=develop

* modify exception in range_test_x86
```
  3c08f676
11 9月, 2019 2 次提交
- Y
  
  make model_optimize_tool run on host (#1990) · 83d4b0e8
  由 Yan Chunwei 提交于 9月 11, 2019
  
  83d4b0e8
- L
  add slice op, reshape op, reshape2 op, squeeze op, squeeze2 op for x86 (#2005) · 13bbd2b8
  由 liu zhengxi 提交于 9月 11, 2019
```
add slice op, reshape op,  reshape2 op, squeeze op and squeeze2 op and their unittests for x86
```
  13bbd2b8
10 9月, 2019 3 次提交
- W
  
  add elementwise_sub and modify argmax (#1964) · 62ea82d0
  由 Wilber 提交于 9月 10, 2019
  
  62ea82d0
- J
  Modify detection test (#2000) · 111db475
  由 juncaipeng 提交于 9月 10, 2019
```
* add assign_value op, arm kernel and test, add fluid_type, test=develop

* add hard_sigmoid, test=develop

* use image and new imple to test detection model, delete faster_rcnn_test, test=develop
```
  111db475
- X
  fix model_optimize_tool error when using host kernel, fix reshape op build... · f25a4571
  由 Xiaoyang LI 提交于 9月 10, 2019
```
fix model_optimize_tool error when using host kernel, fix reshape op build error on ios, test=develop (#1984)
```
  f25a4571
09 9月, 2019 2 次提交

J
add assign_value and hard_sigmoid, add fluid_type (#1983) · 92eeabeb
由 juncaipeng 提交于 9月 09, 2019
```
* add assign_value op, arm kernel and test, add fluid_type, test=develop

* add hard_sigmoid, test=develop
```
92eeabeb

Add concat and elementwise_add cuda kernel (#1979) · 6d1da405

由 Pei Yang 提交于 9月 09, 2019

* add nearest_interp_cuda kernel, test=develop

* add concat op and elementwise_add op

* remove eigen dependency from nearest_interp cuda kernel, test=develop

* free cuda pointers, test=develop

6d1da405

06 9月, 2019 2 次提交

Z
add interpolate fuse pass (#1980) · c49958a2
由 zhupengyang 提交于 9月 06, 2019
```
test=develop
```
c49958a2

add cudnn conv fp32, int8 support (#1974) · f3124b30

由 Zhaolong Xing 提交于 9月 06, 2019

* paddle lite cuda init
can run model with leaky_relu

* add the missing file.
test=develop

* add the load from memory interface.
test=develop

* refine this pr. fix comments
fix ci error
test=develop

* conv impl
fp32:
conv, conv+bais, conv+bias+relu, conv+bias+leaky_relu

int8:
conv, conv+bais+relu(int8 or fp32 output), conv+bias+leaky_relu(int8 or fp32 output)

can run conv+ bias+relu using cxx_api
test=develop

* move the lite/cuda/math to backends/cuda/math
test=develop

f3124b30

03 9月, 2019 3 次提交

H

move npu into backends(directory) and move python/ into tools/python (#1958) · c5e65402
由 huzhiqiang 提交于 9月 03, 2019

c5e65402

rewrite multiclass_nms according to fluid, test=develop (#1945) · deaddf9d

由 juncaipeng 提交于 9月 03, 2019

* add ops for faster rcnn

* disable test for generate_proposals and roi_align, test=develop

* remove .swp file

* remove log in tensor slice

* finish the unit test for roi_align, test=develop

* add box_clip op and fix tensor slice bug

* remove add four op twice

* rewrite the implement for box_coder and sequence_expand, add faster_rcnn_test, test=develop

* fix test bug of box_clip in x86 server, test=develop

* rewrite multiclass_nms according to fluid, test=develop

* fix param load bug in box_coder and multiclass_nms op, test=develop

* fix value transfor error in multiclass_nms, test=develop

deaddf9d

H

create backends directory and move hardware backends into it (#1954) · 31ee212a
由 huzhiqiang 提交于 9月 03, 2019

31ee212a

02 9月, 2019 1 次提交

Add ops and fix bugs for Faster RCNN (#1942) · 635b4958

由 juncaipeng 提交于 9月 02, 2019

* add ops for faster rcnn

* disable test for generate_proposals and roi_align, test=develop

* remove .swp file

* remove log in tensor slice

* finish the unit test for roi_align, test=develop

* add box_clip op and fix tensor slice bug

* remove add four op twice

* rewrite the implement for box_coder and sequence_expand, add faster_rcnn_test, test=develop

* fix test bug of box_clip in x86 server, test=develop

635b4958

01 9月, 2019 1 次提交
- H
  
  add the method of loading model from naive buffer for LightPredictor (#1918) · 13715c57
  由 huzhiqiang 提交于 9月 01, 2019
  
  13715c57
30 8月, 2019 3 次提交

P
add nearest_interp_cuda kernel, test=develop (#1920) · 029971b4
由 Pei Yang 提交于 8月 30, 2019
```
add nearest_interp cuda kernel for Paddle-Lite
```
029971b4

[NPU] add NPU supporting for Java API (#1915) · dbabf5c4

由 hong19860320 提交于 8月 30, 2019

* [NPU] add NPU supporting for Java API
test=develop

* [NPU] refine build script for NPU compiling
test=develop

* [NPU] fix compiling script for NPU
test=develop

dbabf5c4

add precision and persistable attrs for the tensor. (#1899) · e2e07fa4

由 Zhen Wang 提交于 8月 30, 2019

* Add precision and persistable attrs for the tensor. And fix cxx light and full api demo.

* update precision2string methods. test=develop

* move the save logic to the front of the run in mobilenetv1_full_api.cc, test=develop.

* add comments for UpdateVarsOfProgram. test=develop

e2e07fa4

29 8月, 2019 7 次提交

Add yolo_box_cuda multiclass_nms_host kernel. (#1908) · de43e479

由 Wilber 提交于 8月 29, 2019

* add yolo_box_compute cuda

* move multiclass_nms(arm) to host

* add lod in scale op

* add yolo_box_cuda cmake config

* modify shuffle_channel_fuse and transpose_softmax_transpose_fuse to support run ssd model. test=develop

* reshape and transpose op don't have xshape output.

* modify yolo_box_compute_cuda, use tensor to manage cuda memory test=develop

* add yolo_box use kernel test=develop

de43e479

S

[Java API][Comment] upate java api & delete some comments (#1912) · 79714d74
由 sangoly 提交于 8月 29, 2019

79714d74
L

add stack op and add reduce_mean op and their unit tests (#1888) · 20001636
由 liu zhengxi 提交于 8月 29, 2019

20001636

Add load from memory interface (#1903) · ecce1eff

由 Zhaolong Xing 提交于 8月 29, 2019

* paddle lite cuda init
can run model with leaky_relu

* add the missing file.
test=develop

* add the load from memory interface.
test=develop

* refine this pr. fix comments
fix ci error
test=develop

ecce1eff

S

[Java API] add setThreads & setPowerMode interface (#1907) · e91eef1c
由 sangoly 提交于 8月 29, 2019

e91eef1c

ad ops for faster rcnn, including affine_channel, anchor_generator,... · 53b05ce8

由 juncaipeng 提交于 8月 29, 2019

ad ops for faster rcnn, including affine_channel, anchor_generator, generate_proposals and roi_align (#1895)

* add ops for faster rcnn

* disable test for generate_proposals and roi_align, test=develop

* remove .swp file

* remove log in tensor slice

* finish the unit test for roi_align, test=develop

53b05ce8

[NPU] refine npu subgraph and clean code (#1902) · 0c25428c

由 tensor-tang 提交于 8月 29, 2019

* add npu script and tester

* fix npu armv7 so and refine tests

test=develop

* update fix and refine log

test=develop

* refine npu generate api

* refine npu subgraph

* refine npu gen and clean code

* fix model laod

* refine node2rm in subgraph

* refine the build npu functions

test=develop

0c25428c

28 8月, 2019 4 次提交
- H
  
  add floor op,elementwise_div op and assign op test=develop (#1882) · 26450c49
  由 huzhiqiang 提交于 8月 28, 2019
  
  26450c49
- Z
  add transpose-softmax-transpose fuse pass (#1863) · 5e8b15f5
  由 zhupengyang 提交于 8月 28, 2019
```
* add transpose-softmax-transpose fuse pass

test=develop

* enable supported lite-npu ops

test=develop
```
  5e8b15f5
- S
  
  [Publish] include light api impl in full publish lib test=develop (#1885) · 1f5ce9c9
  由 sangoly 提交于 8月 27, 2019
  
  1f5ce9c9
- S
  
  [Protobuf] add combined-param model save/load supported test=develop (#1876) · 93950441
  由 sangoly 提交于 8月 27, 2019
  
  93950441
27 8月, 2019 1 次提交
- Z
  lite cuda init: can run a simple model with leaky_relu (#1860) · 05d3b19b
  由 Zhaolong Xing 提交于 8月 27, 2019
```
* paddle lite cuda init
can run model with leaky_relu

* add the missing file.
test=develop
```
  05d3b19b
26 8月, 2019 2 次提交

Add matmul op (#1837) · b35e89d6

由 Wilber 提交于 8月 26, 2019

* test=develop add matmul_op

* use lite::arm::math::sgemm func to implement matmul

* test=develop  pre-commit command to run clang-format

* Revert "test=develop  pre-commit command to run clang-format"

This reverts commit 3f56474f.

* test=develop pre-commit command to run clang-format

b35e89d6

J

fix benchmark threads, test=develop (#1870) · 118ad09e
由 juncaipeng 提交于 8月 26, 2019

118ad09e