提交 · 21138eb12a69d3b7c7310b993ce05787de9f0af0 · 机器未来 / Paddle

24 5月, 2019 2 次提交

Conv concat relu quantization (#17466) · 5b2a3c4b

由 Sylwester Fraczek 提交于 5月 24, 2019

* add conv_concat_relu fuse

test=develop

* add test code

test=develop

* added missing include with unordered_map

test=develop

* review fixes for wojtuss

test=develop

* remove 'should (not) be fused' comment statements

one of them was invalid anyway

test=develop

5b2a3c4b

fix quantize_squash_pass segfault when no tensor linked to Bias (#17292) · bccb0ba4

由 Sylwester Fraczek 提交于 5月 24, 2019

* fix quantize_squash_pass segfault when there is no tensor linked do Bias input

test=develop

* add googlenet test

test=develop

* fix concat CreateKey not using input format

test=develop

bccb0ba4

23 5月, 2019 1 次提交
- Z
  fix trt ci bug temporary. (#17565) · 38da1030
  由 Zhaolong Xing 提交于 5月 23, 2019
```
ban all trt ut. will fix it later.

test=develop
```
  38da1030
22 5月, 2019 2 次提交

L
fix bug that saved optimal model path in test_analyzer_save_model con… (#17555) · daf88968
由 lijianshe02 提交于 5月 22, 2019
```
* modify saved model path in analyzer_save_model.cc test=develop
```
daf88968

Enable the convolution/relu6(bounded_relu) fusion for FP32 on Intel platform. (#17130) · 2281ebf0

由 guomingz 提交于 5月 22, 2019

* Relu6 is the bottleneck op for Mobilenet-v2. As the mkldnn supports the conv/relu6 fusion, we implement it fusion via cpass way. Due to the int8 enabling for this fusion will be supported in MKLDNN v0.20, so this PR is focused on the fp32 optimization.

Below table shows the benchmark(FPS) which measured on skx-8180(28 cores)
Batch size | with fusion | without fusion
-- | -- | --
1 | 214.7 | 53.4
50 | 1219.727 | 137.280

test=develop

* Fix the format issue

test=develop

* Add the missing nolint comments.

test=develop

* Fix the typos.

test=develop

* Register the conv_brelu_mkldnn_fuse_pass for the MKLDNN engine.

test=develop

* Adjust the indentation.

test=develop

* Add the test_conv_brelu_mkldnn_fuse_pass case.

test=develop

* Slightly update the code per Baidu comments.
Let the parameter definition embedded into the code.
That's will make the code easy to understand.

test=develop

2281ebf0

21 5月, 2019 3 次提交

T
remove unused SERIAL compiler option (#17500) · 3d19f44a
由 Tao Luo 提交于 5月 21, 2019
```
test=develop
```
3d19f44a

Enabling resnet101, vgg16, vgg19 INT8v2 model tests (#17468) · 36757ed2

由 lidanqing 提交于 5月 21, 2019

* Add 6 models tests support in CMake

* enabling resnet101, vgg16, vgg19 INT8v2 model tests
test=develop

* remove SERIAL
test=develop

36757ed2

fix security bugs : (#17464) · ba70cc49

由 liuwei1031 提交于 5月 21, 2019

http://newicafe.baidu.com:80/issue/PaddleSec-33/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-28/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-25/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-24/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-21/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-20/show?from=page

test=develop

ba70cc49

20 5月, 2019 2 次提交
- T
  remove unused expected_kernel_cache_pass (#17486) · 32da5e9c
  由 Tao Luo 提交于 5月 20, 2019
```
test=develop
```
  32da5e9c
- W
  fix the random compilation failure on windows test=develop (#17475) · ca3ba378
  由 wopeizl 提交于 5月 20, 2019
```
* fix the random compilation failure on windows 
```
  ca3ba378
16 5月, 2019 1 次提交

Add setting Scope function for the graph class (#17417) · 4a1b7fec

由 Zhen Wang 提交于 5月 16, 2019

* add set_not_owned function for graph

* add scope set. test=develop

* add scope_ptr enforce not null before setting.test=develop

4a1b7fec

15 5月, 2019 1 次提交
- F
  bug fix (#17392) · e48dd92f
  由 flame 提交于 5月 15, 2019
```
fix secure bug
```
  e48dd92f
09 5月, 2019 1 次提交

fix: (#17279) · 7a3bb061

由 Zhaolong Xing 提交于 5月 09, 2019

1. infernce multi card occupy
2. facebox model inference occupy too much
test=develop

7a3bb061

08 5月, 2019 1 次提交
- W
  improved unit test output (#17266) · 984aa905
  由 Wojciech Uss 提交于 5月 08, 2019
```
added printing data type to differentiate int8 and fp32 latency results

test=develop
```
  984aa905
07 5月, 2019 2 次提交

石

Cherry-pick benchmark related changes from release/1.4 (#17156) · a72dbe9a

由石晓伟提交于 5月 07, 2019

* cherry-pick commit from 88770542

* cherry-pick commit from 3f0b97df

* cherry-pick from 16691:Anakin subgraph support yolo_v3 and faster-rcnn

(cherry picked from commit 8643dbc2)

* Cherry-Pick from 16662 : Anakin subgraph cpu support

(cherry picked from commit 7ad182e1)

* Cherry-pick from 1662, 16797.. : add anakin int8 support

(cherry picked from commit e14ab180)

* Cherry-pick from 16813 : change singleton to graph RegistBlock
test=release/1.4

(cherry picked from commit 4b9fa423)

* Cherry Pick : 16837 Support ShuffleNet and MobileNet-v2

Support ShuffleNet and MobileNet-v2, test=release/1.4

(cherry picked from commit a6fb066f)

* Cherry-pick : anakin subgraph add opt config layout argument #16846
test=release/1.4

(cherry picked from commit 8121b3ec)

* 1. add shuffle_channel_detect

(cherry picked from commit 6efdea89)

* update shuffle_channel op convert, test=release/1.4

(cherry picked from commit e4726a06)

* Modify symbol export rules

test=develop

a72dbe9a

call SetNumThreads everytime to avoid missing omp thread setting (#17224) · 54636a19

由 Leo Zhao 提交于 5月 07, 2019

* call SetNumThreads everytime to avoid missing omp thread setting

resolve #17153
test=develop

* add paddle_num_threads into config for test_analyzer_pyramid_dnn

resolve #17153
test=develop

54636a19

05 5月, 2019 1 次提交
- W
  
  use two GPUs to run the exclusive test test=develop (#17187) · 83c4f772
  由 wopeizl 提交于 5月 05, 2019
  
  83c4f772
30 4月, 2019 1 次提交

fix bn fuse vardesc and add model saver (#17143) · 79ed1c76

由 tensor-tang 提交于 4月 30, 2019

* fix bn fuse vardesc and add model saver

test=develop

* unify save model in test helper

test=develop

* fix mkdir on windows

test=develop

* remove magic number use bn bias var desc

test=develop

79ed1c76

23 4月, 2019 2 次提交
- T
  
  load persistables with selected rows, test=develop (#17047) · 13295d90
  由 tangwei12 提交于 4月 23, 2019
  
  13295d90
- L
  fix runtime_context_cache bug when gpu model has an op runs only on cpu · 490e7462
  由 luotao1 提交于 4月 23, 2019
```
test=develop
```
  490e7462
22 4月, 2019 1 次提交

add parallel build script to ci … (#16901) · d9991dcc

由 wopeizl 提交于 4月 22, 2019

* add parallel build script to ci test=develop
* 1. classify the test case as single card/two cards/multiple cards type
   2. run test case according to the run type

d9991dcc

19 4月, 2019 1 次提交
- T
  disable runtime_context_cache pass by default · aa7b975b
  由 Tao Luo 提交于 4月 19, 2019
```
test=develop
```
  aa7b975b
18 4月, 2019 2 次提交
- N
  fix trt anakin subgraph compile rely · bc6b0ca1
  由 nhzlx 提交于 4月 18, 2019
```
test=develop
```
  bc6b0ca1
- G
  
  Polish DGC code (#16818) · cbdb8a17
  由 gongweibao 提交于 4月 18, 2019
  
  cbdb8a17
17 4月, 2019 1 次提交
- T
  use multi-thread to speedup CI tests · bc037c13
  由 Tao Luo 提交于 4月 17, 2019
```
test=develop
```
  bc037c13
15 4月, 2019 3 次提交
- R
  minus trt ci times. · 1965a224
  由 root 提交于 4月 15, 2019
```
test=develop
```
  1965a224
- L
  add SaveOptimModel interface in analysis_predictor.h and test it in a… (#16441) · de26df44
  由 lijianshe02 提交于 4月 15, 2019
```
* add SaveOptimModel interface in analysis_predictor.h and test it in analyzer_dam_tester and analyzer_resnet50_tester test=develop
```
  de26df44
- L
  improve preprocess script and read from tar · de02d40e
  由 lidanqing 提交于 4月 15, 2019
```
test=develop
```
  de02d40e
12 4月, 2019 2 次提交
- S
  fix memory optim temporarily · f58c3ec1
  由 superjomn 提交于 4月 12, 2019
```
test=develop
```
  f58c3ec1
- Y
  Fix the order while sorting the operators (#16756) · 93cedfdb
  由 Yihua Xu 提交于 4月 12, 2019
```
* Fix the order when sorting operators.

test=develop

* Enable transfomer compare test item.

test=develop

* Use set to replace vector.

test=develop
```
  93cedfdb
11 4月, 2019 1 次提交

Security issue (#16774) · 85363848

由 liuwei1031 提交于 4月 11, 2019

* disable memory_optimize and inpalce strategy by default, test=develop

* fix security issue
http://newicafe.baidu.com:80/issue/PaddleSec-3/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-8/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-12/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-32/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-35/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-37/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-40/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-43/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-44/show?from=page
http://newicafe.baidu.com:80/issue/PaddleSec-45/show?from=page

test=develop

* revert piece.cc, test=develop

* adjust api.cc,test=develop

85363848

09 4月, 2019 1 次提交
- T
  disable seqpool concat pass by default saving CI time · d6c1b5a7
  由 tensor-tang 提交于 4月 09, 2019
```
test=develop
```
  d6c1b5a7
04 4月, 2019 2 次提交
- T
  reduce all analyzer_test ci elasped time · d5c8d4ac
  由 Tao Luo 提交于 4月 04, 2019
```
test=develop
```
  d5c8d4ac
- B
  
  MKLDNN INT8 v2 readme.md (#16515) · 88ceda51
  由 bingyanghuang 提交于 4月 04, 2019
  
  88ceda51
03 4月, 2019 3 次提交
- L
  test_analyzer_int8 tests use default pass order · bd636a9e
  由 luotao1 提交于 4月 03, 2019
```
test=develop
```
  bd636a9e
- L
  
  merge confict, test=develop · b236091e
  由 lujun 提交于 4月 03, 2019
  
  b236091e
- Y
  
  fix identity temporarily (#15942) · 044ae249
  由 Yan Chunwei 提交于 4月 03, 2019
  
  044ae249
02 4月, 2019 3 次提交
- W
  
  fix repeating passes (#16606) · ec2750b3
  由 Wojciech Uss 提交于 4月 02, 2019
  
  ec2750b3
- W
  
  fix dataset reading and add support for full dataset (#16559) · 9b6a0296
  由 Wojciech Uss 提交于 4月 02, 2019
  
  9b6a0296
- L
  fix preprocess script with processbar, integrity check and logs (#16608) · 2ca0de3c
  由 lidanqing 提交于 4月 02, 2019
```
* fix preprocess script with processbar, integrity check and logs
test=develop

* delete unnecessary empty lines, change function name
test=develop
```
  2ca0de3c

机器未来 / Paddle 与 Fork 源项目一致

机器未来 / Paddle
与 Fork 源项目一致