提交 · 3fe6bf5ee6297f719d4194325a6c9f7ecf010ac4 · PaddlePaddle / Paddle

05 7月, 2019 1 次提交
- B
  
  fix command line bug in int8v2 readme (#18507) · 3fe6bf5e
  由 bingyanghuang 提交于 7月 05, 2019
  
  3fe6bf5e
03 7月, 2019 2 次提交
- 石
  Remove the obsolete cmake options (#18481) · 047bba85
  由石晓伟提交于 7月 03, 2019
```
* remove the obsolete cmake options, test=develop

* remove unittests, test=develop
```
  047bba85
- T
  add transfer_scope_cache unit-test (#18467) · d234aa02
  由 Tao Luo 提交于 7月 03, 2019
```
test=develop
```
  d234aa02
02 7月, 2019 1 次提交
- T
  remove unused AnalysisPredictor::SetMkldnnThreadID() (#18444) · 3123d187
  由 Tao Luo 提交于 7月 02, 2019
```
test=develop
```
  3123d187
01 7月, 2019 1 次提交

Fix Pooling output scale (#18186) · 7023a86c

由 Michał Gallus 提交于 7月 01, 2019

* Int8: Fix Pooling output scale

test=develop

* Update scales quantization for certain operators

These include: concat, transpose, pool and reshape. test=develop

* Move concat minimum scale finding to quantizer

test=develop

7023a86c

27 6月, 2019 3 次提交

M
Reset DeviceContext after quantization warmup (#18182) · 84096932
由 Michał Gallus 提交于 6月 27, 2019
```
test=develop
```
84096932
S
add int8 mkldnn prior_box (#17242) · 9252e8fa
由 Sylwester Fraczek 提交于 6月 27, 2019
```
add prior_box quantization code

add scale algo rules for prior box

test=develop
```
9252e8fa

some fixes for int8 mobilenet_ssd tester (#18112) · 5fd68ac1

由 lidanqing 提交于 6月 27, 2019

* some fixes for int8 mobilenet_ssd tester
test=develop

* change wrong data file name
test=develop

* change test images bin file from 200 images to 100 images

* change directory existence to file existence during downloading
test=develop

* reuse download_data
test=develop

* run full dataset when iterations=0
test=develop

5fd68ac1

21 6月, 2019 1 次提交
- W
  
  fix package generation for inference test=develop (#18220) · daa32d53
  由 wopeizl 提交于 6月 21, 2019
  
  daa32d53
19 6月, 2019 3 次提交
- 翟
  Change int8v2 CAPI unit test name and add log in the prediction stage (#18200) · de42fe8f
  由翟飞跃提交于 6月 19, 2019
```
* fix issue 18111;test=develop

* fix timer;test=develop

* refine code;test=develop
```
  de42fe8f
- 翟
  fix spelling errors (#17941) · 802ea509
  由翟飞跃提交于 6月 19, 2019
```
* fix spelling errors; test=develop

* Update API.spec

update md5

* Update API.spec

* change the order of api;test=develop
```
  802ea509
- 翟
  
  add mkldnn Int8v2 slim doc (#17909) · 78441c54
  由翟飞跃提交于 6月 19, 2019
  
  78441c54
16 6月, 2019 2 次提交
- W
  unify FP32 vs. INT8 comparison tests output (#18111) · ca5642c8
  由 Wojciech Uss 提交于 6月 16, 2019
```
test=develop
```
  ca5642c8
- W
  reuse C-API INT8 unit test application (#18077) · c26130f3
  由 Wojciech Uss 提交于 6月 16, 2019
```
* reuse C-API INT8 unit test application

test=develop

* updates after review

test=develop
```
  c26130f3
14 6月, 2019 1 次提交

add Mobilienet ssd int8 analyzer tester (#18075) · 46625415

由 lidanqing 提交于 6月 14, 2019

* add pascalvoc preprocess script and mobilenet-ssd analyzer_tester, wait 17737

* change converting local dataset to downloading and converting tarfile
test=develop

* change the test data_path
test=develop

* change copyright (c) 2016 to copyright (c) 2019
test=develop

46625415

13 6月, 2019 2 次提交
- 石
  
  fix ci test cmake test=develop (#18060) · 42f12a4a
  由石晓伟提交于 6月 13, 2019
  
  42f12a4a
- M
  
  Disable MKLDNN FC in Resnet50 test (#18030) · 8462e2b8
  由 Michał Gallus 提交于 6月 13, 2019
  
  8462e2b8
12 6月, 2019 1 次提交
- 石
  modify the access level of anakin engine (#18015) · 04ea7cb0
  由石晓伟提交于 6月 12, 2019
```
test=develop
```
  04ea7cb0
11 6月, 2019 2 次提交

石

Update the Anakin interfaces for content-dnn and MLU (#17890) · bce259e5

由石晓伟提交于 6月 11, 2019

* update anakin-engine interfaces for content-dnn

test=develop

* support only-gpu mode of Anakin

modify eltwise parse

test=develop

* modification for thread-safe

test=develop

* Integrated template instance

test=develop

* increase template parameters

test=develop

* support MLU predictor

test=develop

* update anakin cmake files

test=develop

* update TargetWrapper::set_device

* update the initialization of anakin subgraph

test=develop

* use the default constructor of base class

test=develop

bce259e5

Light mem reuse strategy for inference. (#17925) · 4e8d5a03

由 Zhaolong Xing 提交于 6月 11, 2019

* fix: when use the load model from memory mode, the RAM occupy is high

test=develop

* ligth mem reuse
test=develop

* fix cpplint
test=develop

4e8d5a03

06 6月, 2019 4 次提交

M

[NGraph] Bert model for a capi, ngraph's support test=develop (#17844) · c1379bf2
由 mozga-intel 提交于 6月 06, 2019

c1379bf2
石
update the initialization of anakin subgraph (#17880) · d008260f
由石晓伟提交于 6月 06, 2019
```
test=develop
```
d008260f
Z
fix: when use the load model from memory mode, the RAM occupy is high (#17788) · ae576f3c
由 Zhaolong Xing 提交于 6月 06, 2019
```
test=develop
```
ae576f3c

翟

INT8 MKL-DNN v2 integrate to slim (#17634) · 993c703b

由翟飞跃提交于 6月 06, 2019

* refactor PR 16865

* delete mergetool files

* test=develop

* test=develop

* test=develop

* test=develop

* create dir for int8 model before call SaveOptimModel

* test=develop

* mkldnn int8 only support linux; test=develop

* refine code; test=develop

* remove comment; test=develop

* refine code; test=develop

* fix bug; test=develop

* add exception for mkldnn_post_training_strategy

* reuse int8v2 CAPI dataset; test=develop

* fix accuracy check bug; test=develop

* remove tab

* convert files to unix format

* test=develop

* reduce CI time;test=develop

* reduce CI time and refine code;test=develop

* refine comment; test=develop

* add cmake FLAGS;test=develop

* remove predict_num;test=develop

993c703b

03 6月, 2019 1 次提交
- T
  make omp thread num default 1 after inference run (#17801) · e089e454
  由 Tao Luo 提交于 6月 03, 2019
```
test=develop
```
  e089e454
29 5月, 2019 3 次提交
- T
  add fc_mkldnn_pass in compare_mkldnn (#17712) · b4b16946
  由 Tao Luo 提交于 5月 29, 2019
```
test=develop
```
  b4b16946
- Z
  fix trt ci timeout error (#17701) · 4337009b
  由 Zhaolong Xing 提交于 5月 29, 2019
```
test=develop
```
  4337009b
- M
  
  Capi for a ngraph engine (#17037) · 5eb81fe5
  由 mozga-intel 提交于 5月 28, 2019
  
  5eb81fe5
28 5月, 2019 2 次提交

Improve mobilenetv2 INT8 performance by using INT8 relu as post-op (#17570) · 04b6c29e

由 lidanqing 提交于 5月 28, 2019

* add INT8 conv+relu6 fuse and enbale mobilentv2 INT8 test
test=develop

* change fasle and 0.0 to fuse_brelu and brelu_threshold
test=develop

change the "fuse_relu||fuse_brelu" to "unsigned_output"
test=develop

* Use relu instead of brelu as INT8 post-op because INT8 brelu is not enabled in mkldnn v0.18
test=develop

* continuous-integration fix
test=develop

04b6c29e

[MKL-DNN] conv_transpose mkldnn bias pass (#17644) · 6d8075ec

由 Jacek Czaja 提交于 5月 28, 2019

* - changes to graph detector

- Changes to pass

- Added ut for new pass

- use_pass

- Added pass to mkldnn passes

- fix to registration

- improved verbose messaging for conv bias passes

- Lint fixes

test=develop

* - Lint fixes

test=develop

6d8075ec

27 5月, 2019 3 次提交

add Concat quantization (#17448) · 96845d21

由 Sylwester Fraczek 提交于 5月 27, 2019

* add Concat quantization
add unit test for quantizing concat
fix for wrong value when the input is not in map of calculated scales
add use_quantizer to concat_op.cc
add scale_algo rules for concat

test=develop

* missing fix for multiple inputs quantize-squash

* wojtuss review fix: adding comment

test=develop

96845d21

Fix the bug in the AnalysisPredictor and add more directions about io APIs. (#17639) · 8bd651b7

由 Zhen Wang 提交于 5月 27, 2019

* fix the bug that sub_scope_ may be null in AnalysisPredictor::Run.

* add more directions about io APIs' docs.

* update the API.spec. test=develop test=document_preview

8bd651b7

Code clean of Allocator (#17602) · 4aa931dd

由 Zeng Jinle 提交于 5月 27, 2019

* Revert "Revert "Fix allocator bug""

This reverts commit 174d0d0b.

* Revert "fix travis ci"

This reverts commit 5656fa9f.

test=develop

* add inlined_vector.h, test=develop

* add inlined_vector_test,test=develop

* clean code of allocator,test=develop

* delete zero_size_allocator.h,test=develop

* fix failed unittest,test=develop

4aa931dd

25 5月, 2019 1 次提交

TRT: Support set dynamic range in int8 mode. (#17524) · 61221ebc

由 Zhaolong Xing 提交于 5月 25, 2019

* fluid int8 train and trt int8 predict align.
trt int8 predict init
op converter

* 2. align fluid int8 train and trt int8 inference.
enhance quant dequant fuse pass
enhance op converter, trt engine, trt engine op, trt subgraph pass.

* 3. add delete_quant_dequant_pass for trt

test=develop

* 4. add the missing file
test=develop

* 5. i modify the c++ interface, but forget to modify the pybind code
fix the IS_TRT_VERSION_GE bug, and fix elementwise op converter
test=develop

61221ebc

24 5月, 2019 3 次提交

[MKL-DNN] Add Fully Connected Op for inference only(#15226) · 0c39b97b

由 Michał Gallus 提交于 5月 24, 2019

* fuse mul and elementwise add to fc

* Reimplement the FC forward operator

* Fix FC MKLDNN integration by transposing weights

* Add FC MKLDNN Pass

test=develop

* FC MKLDNN Pass: change memcpy to std::copy

* Fix MKLDNN FC handling of mismatch input and weights dims

* Lower tolerance for MKL-DNN in resnet50 test

test=develop

* Adjust FC to support MKLDNN Op placement

test=develop

* Adjust Placement Op to set use_mkldnn attribute for graph

test=develop

* MKLDNN FC: fix weights format so that gemm version is called

test=develop

* FC MKLDNN: Remove tolerance decrease from tester_helper

* FC MKL-DNN: Refactor the code, change input reorder to weight reorder

* MKL-DNN FC: Introduce operator caching

test=develop

* FC MKL-DNN: Fix the tensor type in ExpectedKernelType

test=develop

* FC MKL-DNN: fix style changes

test=develop

* FC MKL-DNN: fallback to native on non-supported dim sizes

test=develop

* FC MKLDNN: fix CMake paths

test=develop

* FC MKLDNN: Refine placement pass graph mkldnn attribute

test=develop

* Fix Transpiler error for fuse_conv_eltwise

test=develop

* Fix missing STL includes in files

test=develop

* FC MKL-DNN: Enable new output size computation

Also, refine pass to comply with newest interface.
test=develop

* FC MKL-DNN: enable only when fc_mkldnn_pass is enabled

* FC MKL-DNN: Allow Weights to use oi or io format

* FC MKL-DNN: Adjust UT to work with correct dims

test=develop

* Enable MKL DEBUG for resnet50 analyzer

test=develop

* FC MKL-DNN: Improve Hashing function

test=develop

* FC MKL-DNN: Fix shape for fc weights in transpiler

* FC MKL-DNN: Update input pointer in re-used fc primitive

* Add log for not handling fc fuse for unsupported dims

test=develop

* FC MKL-DNN: Move transpose from pass to Op Kernel

test=develop

* FC MKL-DNN: Disable transpose in unit test

test=develop

* FC MKL-DNN: Remove fc_mkldnn_pass from default list

* Correct Flag for fake data analyzer tests

test=develop

* FC MKL-DNN: Add comment about fc mkldnn pass disablement

test=develop

* FC MKL-DNN: Disable fc in int8 tests

test=develop

0c39b97b

Conv concat relu quantization (#17466) · 5b2a3c4b

由 Sylwester Fraczek 提交于 5月 24, 2019

* add conv_concat_relu fuse

test=develop

* add test code

test=develop

* added missing include with unordered_map

test=develop

* review fixes for wojtuss

test=develop

* remove 'should (not) be fused' comment statements

one of them was invalid anyway

test=develop

5b2a3c4b

fix quantize_squash_pass segfault when no tensor linked to Bias (#17292) · bccb0ba4

由 Sylwester Fraczek 提交于 5月 24, 2019

* fix quantize_squash_pass segfault when there is no tensor linked do Bias input

test=develop

* add googlenet test

test=develop

* fix concat CreateKey not using input format

test=develop

bccb0ba4

23 5月, 2019 1 次提交
- Z
  fix trt ci bug temporary. (#17565) · 38da1030
  由 Zhaolong Xing 提交于 5月 23, 2019
```
ban all trt ut. will fix it later.

test=develop
```
  38da1030
22 5月, 2019 2 次提交

L
fix bug that saved optimal model path in test_analyzer_save_model con… (#17555) · daf88968
由 lijianshe02 提交于 5月 22, 2019
```
* modify saved model path in analyzer_save_model.cc test=develop
```
daf88968

Enable the convolution/relu6(bounded_relu) fusion for FP32 on Intel platform. (#17130) · 2281ebf0

由 guomingz 提交于 5月 22, 2019

* Relu6 is the bottleneck op for Mobilenet-v2. As the mkldnn supports the conv/relu6 fusion, we implement it fusion via cpass way. Due to the int8 enabling for this fusion will be supported in MKLDNN v0.20, so this PR is focused on the fp32 optimization.

Below table shows the benchmark(FPS) which measured on skx-8180(28 cores)
Batch size | with fusion | without fusion
-- | -- | --
1 | 214.7 | 53.4
50 | 1219.727 | 137.280

test=develop

* Fix the format issue

test=develop

* Add the missing nolint comments.

test=develop

* Fix the typos.

test=develop

* Register the conv_brelu_mkldnn_fuse_pass for the MKLDNN engine.

test=develop

* Adjust the indentation.

test=develop

* Add the test_conv_brelu_mkldnn_fuse_pass case.

test=develop

* Slightly update the code per Baidu comments.
Let the parameter definition embedded into the code.
That's will make the code easy to understand.

test=develop

2281ebf0

PaddlePaddle / Paddle 11 个月 前同步成功

PaddlePaddle / Paddle
11 个月前同步成功