提交 · eb7c58297491d76e02fb02a885d36ab0665d825c · PaddlePaddle / Paddle-Lite

16 10月, 2019 1 次提交

[framework][place] remove prefered_place and kHost in valid_places (#2192) · 17833acb

由 sangoly 提交于 10月 16, 2019

* [framework][place] remove prefered_place, use place order in valid_place array instead test=develop

* remove kHost from valid_places test=develop

17833acb

15 10月, 2019 4 次提交

J
Fix quant dequant fuse pass (#2190) · 31ab471e
由 juncaipeng 提交于 10月 15, 2019
```
* fix bug for accessing the removed node, test=develop
```
31ab471e
石

fix pass selection, test=develop (#2187) · a1a22a69
由石晓伟提交于 10月 15, 2019

a1a22a69

[LITE][OPENCL] Fix layout, target pass for OpenCL, add macro of... · 26d05370

由 Yuan Shuai 提交于 10月 15, 2019

[LITE][OPENCL] Fix layout, target pass for OpenCL, add macro of CONVERT_TYPE_TO and READ/WRITE image, memory reuse in ResetLazyImage2D (#2170)

* add macro of CONVERT_TYPE_TO and READ/WRITE image. test=develop

* add data type control. test=develop

* fix io op as general layout and precision. test=develop

* Fix memory reuse strategy for opencl image2d. test=develop

* remove std::array, std::map in about opencl backend. test=develop

26d05370

[NPU] Fix and refine the supporting of multi NPU models (#2037) · e184d474

由 hong19860320 提交于 10月 15, 2019

* [NPU] Fix the bug of loading multi NPU models
test=develop

* [NPU] Use lite tensor to store NPU model, fix the management of multi NPU models, support loading NPU model from memory and reduce the modification of framework
test=develop

* [NPU] Remove redundant header files for NPU bridges,
test=develop

* [NPU] fix NPU deps
test=develop

* [NPU] refine the compiling script for NPU
test=develop

* [NPU] remove redundant subdirectory in lite/CMakeLists.txt
test=develop

* [NPU] Fix and refine NPU test case
test=develop

* [NPU] revoke the modification of other non-NPU modules
test=develop

* [NPU] Remove NPU bridges if target is tiny publish
test=develop

e184d474

14 10月, 2019 2 次提交
- Z
  align yolov3 cuda int8 (#2183) · ed38d79b
  由 Zhaolong Xing 提交于 10月 14, 2019
```
test=develop
```
  ed38d79b
- J
  Optimize quant_dequant_fuse_pass (#2169) · 0260d322
  由 juncaipeng 提交于 10月 14, 2019
```
* optimize quant_dequant_fuse_pass, test=develop
```
  0260d322
11 10月, 2019 2 次提交

CUDA: can run yolov3 int8 (#2172) · 29f448c6

由 Zhaolong Xing 提交于 10月 11, 2019

* add conv int8 support(in condition which the input or output channel not be the times of 4)
add add_kernel for cuda.

* can run yolov3 fp32
test=develop

* 1. fix bug with yolov3 run
test=develop

* can run yolov3 int8 test=develop

29f448c6

[LITE][OPENCL] support image2d type (#2158) · 74074146

由 Yuan Shuai 提交于 10月 11, 2019

* [LITE][OPENCL] support image2d. test=develop

* add context changed with consider image*. test=develop

* add layout, relu image kernels. test=develop

* replace image_data with data, mutable_image_data with mutable_data, test=develop

* comment unused var. test=develop

* remove unused var. test=develop

74074146

01 10月, 2019 1 次提交
- S
  
  [Profile] Add time profile unit flags, ms or us test=develop (#2144) · d0fd7265
  由 sangoly 提交于 10月 01, 2019
  
  d0fd7265
27 9月, 2019 2 次提交
- Z
  can run yolov3 fp32 on cuda devices (#2092) · c4b5e32c
  由 Zhaolong Xing 提交于 9月 27, 2019
```
* add conv int8 support(in condition which the input or output channel not be the times of 4)
add add_kernel for cuda.

* can run yolov3 fp32
test=develop

* 1. fix bug with yolov3 run
test=develop
```
  c4b5e32c
- S
  
  [Profile] add kernel runtime profile && add op runtime summary test=develop (#2136) · b82a9eec
  由 sangoly 提交于 9月 27, 2019
  
  b82a9eec
26 9月, 2019 1 次提交
- S
  
  [Fc Fusion] fix fc fusion duplicative arguments bug test=develop (#2135) · cc851a38
  由 sangoly 提交于 9月 26, 2019
  
  cc851a38
25 9月, 2019 2 次提交
- X
  
  add workspace compute funcs for direct conv, test=develop (#2132) · f03217b4
  由 Xiaoyang LI 提交于 9月 25, 2019
  
  f03217b4
- X
  
  fix arm device_info error, fix bind big core error, improve 855 performance, test=develop (#2133) · 845af6e0
  由 Xiaoyang LI 提交于 9月 25, 2019
  
  845af6e0
23 9月, 2019 1 次提交
- W
  
  model_test add host place (#2109) · 5f75d92c
  由 Wilber 提交于 9月 23, 2019
  
  5f75d92c
21 9月, 2019 1 次提交
- W
  
  add yolo_box to mem_optimize_pass's unuse op set · 21773b6c
  由 Wilber 提交于 9月 21, 2019
  
  21773b6c
20 9月, 2019 2 次提交
- H
  fix compiling and fix code style (#2088) · 214b48ff
  由 huzhiqiang 提交于 9月 20, 2019
```
* fix compiling and fix code style test=develop
```
  214b48ff
- Z
  1. the split op's bug will triger memory optimize pass failed. (#2070) · b7f5d94b
  由 Zhaolong Xing 提交于 9月 20, 2019
```
test=develop
```
  b7f5d94b
19 9月, 2019 2 次提交

石

add full_api_static target and fix building errors, test=develop (#2064) · 4a948cfc

由石晓伟提交于 9月 19, 2019

* add full_api_static target and fix building errors, test=develop

* fix build errors, test=develop

* fix code style, test=develop

* fix lite/model_parser/pb/var_desc.cc, test=develop

* fix building errors, test=develop

* modify lite/tools/debug/CMakeLists.txt, test=develop

4a948cfc

Bug fix for model save and load (#1992) · 5404c2ee

由 TianXiaogang 提交于 9月 19, 2019

* fix: fix model parser and save bug

* style: delete debug code

* fix: fix light_predictor program run model with subblock bug

5404c2ee

18 9月, 2019 2 次提交

石

modify the device binding logic of the pass, test=develop (#2060) · dbc8f893
由石晓伟提交于 9月 18, 2019

dbc8f893

fix bias quantize error && fix clang build error (#2049) · 8d6f475e

由 Xiaoyang LI 提交于 9月 18, 2019

* fix gemm_int8, gemv-int8 and conv-int8 math function, add float bias

* change conv impl

* neon int8 kernel support float bias

* arm compute kernel support float bias

* add math_test target

* add tensor utils for testing, fix sgemm ut error

* add gemm_int8 unit test, support float bias

* fix build script

* add conv compute unit test for arm

* fix build script, test=develop

* fix fp32 dw conv3x3s1, test=develop

* add fp32 dw conv3x3s1, test=develop

* add armv7 fp32 dw conv3x3s1, test=develop

* add fp32 depthwise conv3x3s2, test=develop

* fix fp32 conv3x3 depthwise build error, test=develop

* fix gemm_like conv trans weights error, test=develop

* fix int8 depthwise conv3x3 error, test=develop

* turn on all test for arm fp32 conv, test=develop

* fix int8 conv1x1 error

* fix int8 direct conv3x3s1 error, test=develop

* fix int8 direct conv3x3s2, test=develop

* turn on all test for arm int8 conv, test=develop

* fix int8 fc error, change mobilenetv1-int8 ground-truth result to fluid, test=develop

* remove debug info, strip ut binary, test=develop

* fix conv compute error, test=develop

* change Init() to ReInitWhenNeeded(), test=develop

* fix code style, test=develop

* remote engine_test, test=develop

* fix building server tests error, test=develop

* fix sdot clang build error, test=develop

* fix sgemm ut timeout error, test=develop

* fix clang build error, test=develop

* turn off math basic test due to ci time out, test=develop

* fix conv_int8 ut error, test=develop

8d6f475e

17 9月, 2019 1 次提交
- S
  [Cxx API] add build-in version info (#2047) · 9a90da46
  由 sangoly 提交于 9月 17, 2019
```
* [Cxx API] add build-in version info

* update: add version.h.in template
```
  9a90da46
16 9月, 2019 1 次提交
- X
  
  fix math dependencies error (#2023) · 79a03c2b
  由 Xiaoyang LI 提交于 9月 16, 2019
  
  79a03c2b
13 9月, 2019 2 次提交

石

checkout if passes match targets and kernels, test=develop (#2035) · 2f3d7fd6

由石晓伟提交于 9月 13, 2019

* checkout if passes match targets and kernels, test=develop

* add pass_utils, test=develop

* fix lite/core/mir/pass_registry.h, test=develop

* improve code styles, test=develop

* fix spell error, test=develop

2f3d7fd6

Z

Add the memory optim pass (#2018) · 129b689b
由 Zhaolong Xing 提交于 9月 13, 2019

129b689b

12 9月, 2019 2 次提交
- Y
  
  integrate model_optimize_tool compilation to build.sh (#2033) · 4975c600
  由 Yan Chunwei 提交于 9月 12, 2019
  
  4975c600
- G
  
  enable native compiling on raspberry pi and rk3399 (#2021) · 9fae3170
  由 guofei 提交于 9月 12, 2019
  
  9fae3170
11 9月, 2019 3 次提交
- Y
  
  make model_optimize_tool run on host (#1990) · d72dc4d2
  由 Yan Chunwei 提交于 9月 11, 2019
  
  d72dc4d2
- 石
  make passes related to the device type, test=develop (#2012) · 3c0e8a6a
  由石晓伟提交于 9月 11, 2019
```
* make passes related to the device type, test=develop

* improve tips, test=develop
```
  3c0e8a6a
- Z
  fix conv-act-fuse-pass when there is no "bias" (#2003) · 40aac462
  由 zhupengyang 提交于 9月 11, 2019
```
test=develop
```
  40aac462
09 9月, 2019 2 次提交
- J
  add assign_value and hard_sigmoid, add fluid_type (#1983) · 9796c57d
  由 juncaipeng 提交于 9月 09, 2019
```
* add assign_value op, arm kernel and test, add fluid_type, test=develop

* add hard_sigmoid, test=develop
```
  9796c57d
- Z
  add calib cuda kernel. (#1977) · 9681b642
  由 Zhen Wang 提交于 9月 09, 2019
```
* add calib cuda kernel.

* add unit test for calib cuda kernel. test=develop
```
  9681b642
07 9月, 2019 1 次提交

add lite x86 ops for ASR test=develop (#1981) · 25b775d6

由 lijianshe02 提交于 9月 07, 2019

* add lite x86 ops for ASR test=develop

* add lite x86 ops for ASR test=develop

* fix x86 ci run test problems test=develop

* fix mkl path for CI test=develop

25b775d6

06 9月, 2019 2 次提交

Z
add interpolate fuse pass (#1980) · 3c004508
由 zhupengyang 提交于 9月 06, 2019
```
test=develop
```
3c004508

add cudnn conv fp32, int8 support (#1974) · 23d83c04

由 Zhaolong Xing 提交于 9月 06, 2019

* paddle lite cuda init
can run model with leaky_relu

* add the missing file.
test=develop

* add the load from memory interface.
test=develop

* refine this pr. fix comments
fix ci error
test=develop

* conv impl
fp32:
conv, conv+bais, conv+bias+relu, conv+bias+leaky_relu

int8:
conv, conv+bais+relu(int8 or fp32 output), conv+bias+leaky_relu(int8 or fp32 output)

can run conv+ bias+relu using cxx_api
test=develop

* move the lite/cuda/math to backends/cuda/math
test=develop

23d83c04

03 9月, 2019 3 次提交
- T
  refine the npu graph and subgraph (#1959) · d3c07632
  由 tensor-tang 提交于 9月 03, 2019
```
* fix attr and refine subgraph pass test=develop

* refine the npu pass functions

* fix test

test=develop
```
  d3c07632
- H
  
  move npu into backends(directory) and move python/ into tools/python (#1958) · 0328b5c2
  由 huzhiqiang 提交于 9月 03, 2019
  
  0328b5c2
- H
  
  create backends directory and move hardware backends into it (#1954) · fede4a1c
  由 huzhiqiang 提交于 9月 03, 2019
  
  fede4a1c