提交 · 7014a76b0aac93cf2d463d978ba3d9a1f945f81f · PaddlePaddle / Paddle-Lite

07 9月, 2019 1 次提交

add lite x86 ops for ASR test=develop (#1981) · 7014a76b

由 lijianshe02 提交于 9月 07, 2019

* add lite x86 ops for ASR test=develop

* add lite x86 ops for ASR test=develop

* fix x86 ci run test problems test=develop

* fix mkl path for CI test=develop

7014a76b

06 9月, 2019 2 次提交

Z
add interpolate fuse pass (#1980) · c49958a2
由 zhupengyang 提交于 9月 06, 2019
```
test=develop
```
c49958a2

add cudnn conv fp32, int8 support (#1974) · f3124b30

由 Zhaolong Xing 提交于 9月 06, 2019

* paddle lite cuda init
can run model with leaky_relu

* add the missing file.
test=develop

* add the load from memory interface.
test=develop

* refine this pr. fix comments
fix ci error
test=develop

* conv impl
fp32:
conv, conv+bais, conv+bias+relu, conv+bias+leaky_relu

int8:
conv, conv+bais+relu(int8 or fp32 output), conv+bias+leaky_relu(int8 or fp32 output)

can run conv+ bias+relu using cxx_api
test=develop

* move the lite/cuda/math to backends/cuda/math
test=develop

f3124b30

03 9月, 2019 4 次提交
- T
  refine the npu graph and subgraph (#1959) · d60b8d61
  由 tensor-tang 提交于 9月 03, 2019
```
* fix attr and refine subgraph pass test=develop

* refine the npu pass functions

* fix test

test=develop
```
  d60b8d61
- H
  
  move npu into backends(directory) and move python/ into tools/python (#1958) · c5e65402
  由 huzhiqiang 提交于 9月 03, 2019
  
  c5e65402
- H
  
  create backends directory and move hardware backends into it (#1954) · 31ee212a
  由 huzhiqiang 提交于 9月 03, 2019
  
  31ee212a
- T
  refine the attr assert (#1947) · 60bbc691
  由 tensor-tang 提交于 9月 03, 2019
```
test=develop
```
  60bbc691
02 9月, 2019 1 次提交

Add ops and fix bugs for Faster RCNN (#1942) · 635b4958

由 juncaipeng 提交于 9月 02, 2019

* add ops for faster rcnn

* disable test for generate_proposals and roi_align, test=develop

* remove .swp file

* remove log in tensor slice

* finish the unit test for roi_align, test=develop

* add box_clip op and fix tensor slice bug

* remove add four op twice

* rewrite the implement for box_coder and sequence_expand, add faster_rcnn_test, test=develop

* fix test bug of box_clip in x86 server, test=develop

635b4958

01 9月, 2019 1 次提交

[ARM][CPU] Fix time counter of arm cpu profiler (#1925) · e3fb95ae

由 Yuan Shuai 提交于 9月 01, 2019

* Fix timer of arm cpu profiler. test=develop

* Fix un-added op in cmake.test=develop

* fix cmake error

* fix cmake error, test=develop

* Fix pass sequence. test=develop

* replace option with lite_option. test=develop

* disable profile mode by default. test=develop

* Fix error option name. test=develop

e3fb95ae

30 8月, 2019 1 次提交

add precision and persistable attrs for the tensor. (#1899) · e2e07fa4

由 Zhen Wang 提交于 8月 30, 2019

* Add precision and persistable attrs for the tensor. And fix cxx light and full api demo.

* update precision2string methods. test=develop

* move the save logic to the front of the run in mobilenetv1_full_api.cc, test=develop.

* add comments for UpdateVarsOfProgram. test=develop

e2e07fa4

29 8月, 2019 6 次提交

T

add conv2d transpose fuse (#1909) · 57ee8714
由 tensor-tang 提交于 8月 29, 2019

57ee8714

Add yolo_box_cuda multiclass_nms_host kernel. (#1908) · de43e479

由 Wilber 提交于 8月 29, 2019

* add yolo_box_compute cuda

* move multiclass_nms(arm) to host

* add lod in scale op

* add yolo_box_cuda cmake config

* modify shuffle_channel_fuse and transpose_softmax_transpose_fuse to support run ssd model. test=develop

* reshape and transpose op don't have xshape output.

* modify yolo_box_compute_cuda, use tensor to manage cuda memory test=develop

* add yolo_box use kernel test=develop

de43e479

Add load from memory interface (#1903) · ecce1eff

由 Zhaolong Xing 提交于 8月 29, 2019

* paddle lite cuda init
can run model with leaky_relu

* add the missing file.
test=develop

* add the load from memory interface.
test=develop

* refine this pr. fix comments
fix ci error
test=develop

ecce1eff

T
[NPU] enable npu program rollback (#1906) · 178a93b9
由 tensor-tang 提交于 8月 29, 2019
```
test=develop
```
178a93b9

ad ops for faster rcnn, including affine_channel, anchor_generator,... · 53b05ce8

由 juncaipeng 提交于 8月 29, 2019

ad ops for faster rcnn, including affine_channel, anchor_generator, generate_proposals and roi_align (#1895)

* add ops for faster rcnn

* disable test for generate_proposals and roi_align, test=develop

* remove .swp file

* remove log in tensor slice

* finish the unit test for roi_align, test=develop

53b05ce8

[NPU] refine npu subgraph and clean code (#1902) · 0c25428c

由 tensor-tang 提交于 8月 29, 2019

* add npu script and tester

* fix npu armv7 so and refine tests

test=develop

* update fix and refine log

test=develop

* refine npu generate api

* refine npu subgraph

* refine npu gen and clean code

* fix model laod

* refine node2rm in subgraph

* refine the build npu functions

test=develop

0c25428c

28 8月, 2019 2 次提交
- Z
  add transpose-softmax-transpose fuse pass (#1863) · 5e8b15f5
  由 zhupengyang 提交于 8月 28, 2019
```
* add transpose-softmax-transpose fuse pass

test=develop

* enable supported lite-npu ops

test=develop
```
  5e8b15f5
- S
  
  [Protobuf] add combined-param model save/load supported test=develop (#1876) · 93950441
  由 sangoly 提交于 8月 27, 2019
  
  93950441
27 8月, 2019 2 次提交
- Z
  lite cuda init: can run a simple model with leaky_relu (#1860) · 05d3b19b
  由 Zhaolong Xing 提交于 8月 27, 2019
```
* paddle lite cuda init
can run model with leaky_relu

* add the missing file.
test=develop
```
  05d3b19b
- T
  [NPU] add script and refine tests (#1873) · 32065859
  由 tensor-tang 提交于 8月 27, 2019
```
* add npu script and tester

* fix npu armv7 so and refine tests

test=develop

* update fix and refine log

test=develop
```
  32065859
25 8月, 2019 1 次提交
- Y
  
  leave tiny-publish out of third-party dependencies (#1853) · 05b86272
  由 Yan Chunwei 提交于 8月 25, 2019
  
  05b86272
24 8月, 2019 2 次提交
- X
  support setting cluster and threads in MobileConfig (#1848) · bc79142c
  由 Xiaoyang LI 提交于 8月 24, 2019
```
* fix building ios tiny publish lib error

* support setting cluster and threads in MobileConfig

* fix build error, test=develop

* fix building server publish error, test=develop
```
  bc79142c
- Y
  
  Refactor op kernel compile system (#1831) · 36419766
  由 Yan Chunwei 提交于 8月 24, 2019
  
  36419766
23 8月, 2019 3 次提交
- T
  feat: (#1836) · 9388dace
  由 TianXiaogang 提交于 8月 23, 2019
```
add model_run_test_image
    add range_max_quant op
    add flatten op
    add flatten2 op
fix:
    fix density_prior_box density_size type from float to int
    fix prior_box and density_prior_box some check for get_attr
test=develop
```
  9388dace
- Z
  [NPU] support generate multiple IO subgraph (#1828) · 61836c46
  由 zhupengyang 提交于 8月 23, 2019
```
test=develop
```
  61836c46
- T
  enable shuffle channel fuse (#1834) · 1ec18e53
  由 tensor-tang 提交于 8月 23, 2019
```
test=develop
```
  1ec18e53
22 8月, 2019 2 次提交
- H
  add license (#1820) · 8c3410be
  由 hong19860320 提交于 8月 22, 2019
```
test=develop
```
  8c3410be
- Y
  
  port lite code (#1819) · 30c273de
  由 Yan Chunwei 提交于 8月 22, 2019
  
  30c273de
16 8月, 2019 1 次提交
- Y
  
  publish lite (#1800) · 699d6cd0
  由 Yan Chunwei 提交于 8月 16, 2019
  
  699d6cd0