提交 · e76100236c6448055d353475cdc375beded82cef · PaddlePaddle / Paddle-Lite

06 11月, 2019 2 次提交
- S
  
  [profile] fix profile count bug & exclude warmup profile monitor test=develop (#2374) · 2eb9e2dd
  由 sangoly 提交于 11月 06, 2019
  
  2eb9e2dd
- T
  fix: fix omp thread bug (#2371) · 9c29345e
  由 TianXiaogang 提交于 11月 06, 2019
```
* fix: fix omp thread bug

* fix: fix compile android level & test=develop

* fix: fix omp bug

* fix: fix omp bug
test=develop
```
  9c29345e
05 11月, 2019 1 次提交
- L
  fix StepRNN model run related bugs (#2300) · 99d4f70e
  由 lijianshe02 提交于 11月 05, 2019
```
* fix step rnn model run bugs test=develop
```
  99d4f70e
01 11月, 2019 3 次提交
- J
  
  add int8 benchmark, test=develop (#2338) · 363118e0
  由 juncaipeng 提交于 11月 01, 2019
  
  363118e0
- L
  
  enable the model tests precision checks, test=develop (#2325) · b323d1f8
  由 liu zhengxi 提交于 11月 01, 2019
  
  b323d1f8
- 石
  
  refactor: BindTargets and ExcludeTargets, test=develop (#2321) · 9749fde1
  由石晓伟提交于 11月 01, 2019
  
  9749fde1
31 10月, 2019 2 次提交

Add 'inference' folder into build.lite.x86, 'inference' folder contains head... · cb32a145

由 huzhiqiang 提交于 10月 31, 2019

Add 'inference' folder into build.lite.x86, 'inference' folder contains head files、library files and demo of x86 compiling (#2301)

Add inference folder into build.lite.x86, build.lite.x86/inference folder contains head files、library files and demo of x86 compiling

cb32a145

Alter the model unit tests api (#2297) · bd0b9667

由 liu zhengxi 提交于 10月 31, 2019

* alter the unit tests api for googlenet, resnet50, inceptionv4, mobilenetv1, mobilenetv2 and thier corresponding cmake, test=develop

* disable checks and correct some incorrect statement, test=develop

bd0b9667

29 10月, 2019 3 次提交
- H
  remove dynamic library method from ci_build.sh (#2275) · 8ed7fc46
  由 huzhiqiang 提交于 10月 29, 2019
```
remove dynamic library compiling from `ci_build.sh  build_test_server`
```
  8ed7fc46
- G
  Make armlinux complie libpaddle_*.so (#2276) · 64373fcf
  由 guofei 提交于 10月 29, 2019
```
test=develop
```
  64373fcf
- H
  [LITE][NPU] Add supporting for Huawei offical DDK (#2262) · 92f0b3c5
  由 hong19860320 提交于 10月 29, 2019
```
* Add supporting for Huawei offical DDK
* Fix the param of graph op in NPU graph computing kernel
```
  92f0b3c5
28 10月, 2019 1 次提交

[LITE][XPU] initial support for XPU (#2202) · ac1b2f9f

由 hong19860320 提交于 10月 28, 2019

* Initial support for XPU
* Fix compiling errors of XPU
* Move XPU op kernel bridges from backends to kernels to fix deps order
* Change the namespace and directory of XPU bridges
* Add XPU SDK
* Fix header files and namespace of XPU SDK
* Add unit tests for relu and conv2d ops
* Restore the modification of paddle_api_test
* Supports simple model which contains only a relu layer
* Add compiling scripts for XPU
* Fix compiling errors of XPU
* Add comments for XPU LoadModel and BuildModel

ac1b2f9f

27 10月, 2019 1 次提交

model dynamic library tailoring (#2256) · b16917a4

由 huzhiqiang 提交于 10月 27, 2019

* add shell file to automatically build and collect publish result test=develop
* modify API inference of model_optimize_tool and add option for tiny&full publish test=develop

b16917a4

25 10月, 2019 1 次提交
- H
  X86 dynamic compile (#2218) · 68003c0e
  由 huzhiqiang 提交于 10月 25, 2019
```
build x86 dynamic library in ci_build.sh
./lite/tool/ci_build.sh build_test_server
```
  68003c0e
24 10月, 2019 2 次提交

Make inceptionv4, resnet50, googlenet can run on x86 paltform (#2250) · ca7fefa1

由 liu zhengxi 提交于 10月 24, 2019

* make inceptionv4, resnet50, googlenet can run on x86 paltform and fix the compare part in x86 unittests, test=develop

* fix googlenet tests for benchmark record, test=develop

* [framework][profile] fix profile dump bug when op is feed and fetch test=develop (sangoly)

ca7fefa1

S
[python api] add armlinux supported and publish paddle-lite python (#2252) · 60709e60
由 sangoly 提交于 10月 24, 2019
```
* [python api] add armlinux supported and publish paddle-lite python lib & demo

* add cuda build
test=develop
```
60709e60

23 10月, 2019 3 次提交

remove log in reshape, fix conv error when padding size=4 (#2199) · f8ff5aa4

由 Xiaoyang LI 提交于 10月 23, 2019

* remove log in reshape, fix conv error when padding size=4, test=develop

* fix style, test=develop

* remove useless code, test=develop

* remove redundant model test file, test=develop

* change cluster to power_mode, test=develop

* fix build error, test=develop

* change cluster to power_mode, test=develop

* change opt_nb to use_optimize_nb, test=develop

* null, test=develop

f8ff5aa4

S
add python api (#2225) · bbfaacec
由 sangoly 提交于 10月 23, 2019
```
* [python api] init add python api test=develop
```
bbfaacec
G
Make armlinux compile libpaddle_light_api_shared.so (#2238) · 0ab6cf42
由 guofei 提交于 10月 23, 2019
```
* Make armlinux compile libpaddle_light_api_shared.so
test=develop
```
0ab6cf42

22 10月, 2019 2 次提交

Optimize quant_dequant (#2215) · aefb4ea3

由 juncaipeng 提交于 10月 22, 2019

* Add DeleteQuantOpFuser
* Add fake_quantize_dequantize_moving_avg_abs_max_op
* Add DeleteQuantDequantOpFuser

aefb4ea3

Transformer pr (#2214) · 330644b0

由 TianXiaogang 提交于 10月 22, 2019

* feat: add beam_search_special function for support nlp model

* fix: add beam_search_compute kernel input and output

* feat: add assign op & copy_compute kernel

* feat: add fill_const_batch_size_like op & kernel

* feat: add layer_norm op and kernel and ut

* fix: fix some bugs
    fix mul_op infer_shape bug when x_dim_idx = 2, x_dims.size()=3 & y_dim_idx = 1, y_dims.size()=2
    fix elementwise_compute bug when y axis is all 1
    fix beam_search choose math_func wrong bug
    fix layer_norm get attr bug
    fix fill_constant_batch_size_like shape_set bug

* feat: add gather op and kernel & and transform ut

* feats: add ops and fix bugs to support transformer op
       fix type_cast passes to skip `while`
       fix elementwise infer_shape bug when x.dims=3 and y.dims={1} & axis=0
       fix lookup_table compute bug
       fix read_from_array/beam_search/increment/compate/gather ops data_type problems

* fix:
    transfomer ut add word read inferface
    fix copy/gather/norm/layer_norm include path problem

* fix:debug info

* fix: fix input reshape bug

* fix: fix norm bug

* style: style fix & test=develop

* style: fix operators cmakelist

* style: fix operators cmakelist; test=develop

* fix and test=develop

* fix and test=develop

* style: style fix; test=develop

330644b0

21 10月, 2019 1 次提交
- 石
  link static library with cuda, test=develop (#2228) · 31f1e382
  由石晓伟提交于 10月 21, 2019
```
* add static libraries of cuda, test=develop

* update cuda make
```
  31f1e382
18 10月, 2019 1 次提交

Fix codestyle of GetInputName&GetOutputName (#2185) · 27a40b8f

由 huzhiqiang 提交于 10月 18, 2019

* add shell file to automatically build and collect publish result test=develop

* modify codestyle of getInputNames test=develop

* test=develop

* rm publish.sh

* remove copy of func param

* test=develop

* test=devcelop

* test=develop

* test=develop

* const & test=develop

* modify variable defination test=develop

* test=develop

* test=develop

* test=develop

* test=develop

27a40b8f

16 10月, 2019 2 次提交

Z
Ban feed and fetch op during inference (#2198) · ad541652
由 Zhaolong Xing 提交于 10月 16, 2019
```
* init: delete feed and fetch op, using zero copy
test=develop

* delete the unused test
test=develop
```
ad541652

[framework][place] remove prefered_place and kHost in valid_places (#2192) · 17833acb

由 sangoly 提交于 10月 16, 2019

* [framework][place] remove prefered_place, use place order in valid_place array instead test=develop

* remove kHost from valid_places test=develop

17833acb

15 10月, 2019 2 次提交

J
fix benchmark, test=develop (#2188) · dbb660ee
由 juncaipeng 提交于 10月 15, 2019
```
* fix benchmark, test=develop
```
dbb660ee

[NPU] Fix and refine the supporting of multi NPU models (#2037) · e184d474

由 hong19860320 提交于 10月 15, 2019

* [NPU] Fix the bug of loading multi NPU models
test=develop

* [NPU] Use lite tensor to store NPU model, fix the management of multi NPU models, support loading NPU model from memory and reduce the modification of framework
test=develop

* [NPU] Remove redundant header files for NPU bridges,
test=develop

* [NPU] fix NPU deps
test=develop

* [NPU] refine the compiling script for NPU
test=develop

* [NPU] remove redundant subdirectory in lite/CMakeLists.txt
test=develop

* [NPU] Fix and refine NPU test case
test=develop

* [NPU] revoke the modification of other non-NPU modules
test=develop

* [NPU] Remove NPU bridges if target is tiny publish
test=develop

e184d474

14 10月, 2019 2 次提交
- H
  add GetInputNames 、 GetOutPutNames 、 GetInputByName and GetTensor method (#2154) · 1cd077dc
  由 huzhiqiang 提交于 10月 14, 2019
```
* add GetInputNames and GetOutPutNames and GetInputByName method test=develop
```
  1cd077dc
- J
  Optimize quant_dequant_fuse_pass (#2169) · 0260d322
  由 juncaipeng 提交于 10月 14, 2019
```
* optimize quant_dequant_fuse_pass, test=develop
```
  0260d322
11 10月, 2019 2 次提交
- J
  
  add rsqrt op, test=develop (#2176) · 78ddd64d
  由 juncaipeng 提交于 10月 11, 2019
  
  78ddd64d
- H
  move the method of SetThread and SetPowerMode from MobileConfig into ConfigBase (#2147) · 6aba3b8b
  由 huzhiqiang 提交于 10月 11, 2019
```
* move the method of SetThread and SetPowerMode from MobileConfig into ConfigBase 
* cxxPredictor will support SetThread and SetPowerMode method
```
  6aba3b8b
27 9月, 2019 1 次提交

can run yolov3 fp32 on cuda devices (#2092) · c4b5e32c

由 Zhaolong Xing 提交于 9月 27, 2019

* add conv int8 support(in condition which the input or output channel not be the times of 4)
add add_kernel for cuda.

* can run yolov3 fp32
test=develop

* 1. fix bug with yolov3 run
test=develop

c4b5e32c

25 9月, 2019 2 次提交
- X
  
  fix MobileConfig get mode and threads error (#2134) · 5c98e002
  由 Xiaoyang LI 提交于 9月 25, 2019
  
  5c98e002
- X
  
  fix arm device_info error, fix bind big core error, improve 855 performance, test=develop (#2133) · 845af6e0
  由 Xiaoyang LI 提交于 9月 25, 2019
  
  845af6e0
23 9月, 2019 1 次提交
- W
  
  model_test add host place (#2109) · 5f75d92c
  由 Wilber 提交于 9月 23, 2019
  
  5f75d92c
19 9月, 2019 4 次提交

石

add full_api_static target and fix building errors, test=develop (#2064) · 4a948cfc

由石晓伟提交于 9月 19, 2019

* add full_api_static target and fix building errors, test=develop

* fix build errors, test=develop

* fix code style, test=develop

* fix lite/model_parser/pb/var_desc.cc, test=develop

* fix building errors, test=develop

* modify lite/tools/debug/CMakeLists.txt, test=develop

4a948cfc

Bug fix for model save and load (#1992) · 5404c2ee

由 TianXiaogang 提交于 9月 19, 2019

* fix: fix model parser and save bug

* style: delete debug code

* fix: fix light_predictor program run model with subblock bug

5404c2ee

modify dynamic library: libpaddle_cxx_api.so (#2057) · 2ad127dd

由 huzhiqiang 提交于 9月 19, 2019

(1)modify tiny publish so to make it excutable
(2)modify the bug of compiling in armlinux
(3)change 3 full publish so into 2

2ad127dd

S

[Java API] add getVersion() interface to get c++ lib's version information test=develop (#2063) · 4656442f
由 sangoly 提交于 9月 19, 2019

4656442f

18 9月, 2019 1 次提交

fix bias quantize error && fix clang build error (#2049) · 8d6f475e

由 Xiaoyang LI 提交于 9月 18, 2019

* fix gemm_int8, gemv-int8 and conv-int8 math function, add float bias

* change conv impl

* neon int8 kernel support float bias

* arm compute kernel support float bias

* add math_test target

* add tensor utils for testing, fix sgemm ut error

* add gemm_int8 unit test, support float bias

* fix build script

* add conv compute unit test for arm

* fix build script, test=develop

* fix fp32 dw conv3x3s1, test=develop

* add fp32 dw conv3x3s1, test=develop

* add armv7 fp32 dw conv3x3s1, test=develop

* add fp32 depthwise conv3x3s2, test=develop

* fix fp32 conv3x3 depthwise build error, test=develop

* fix gemm_like conv trans weights error, test=develop

* fix int8 depthwise conv3x3 error, test=develop

* turn on all test for arm fp32 conv, test=develop

* fix int8 conv1x1 error

* fix int8 direct conv3x3s1 error, test=develop

* fix int8 direct conv3x3s2, test=develop

* turn on all test for arm int8 conv, test=develop

* fix int8 fc error, change mobilenetv1-int8 ground-truth result to fluid, test=develop

* remove debug info, strip ut binary, test=develop

* fix conv compute error, test=develop

* change Init() to ReInitWhenNeeded(), test=develop

* fix code style, test=develop

* remote engine_test, test=develop

* fix building server tests error, test=develop

* fix sdot clang build error, test=develop

* fix sgemm ut timeout error, test=develop

* fix clang build error, test=develop

* turn off math basic test due to ci time out, test=develop

* fix conv_int8 ut error, test=develop

8d6f475e