提交 · 21138eb12a69d3b7c7310b993ce05787de9f0af0 · 机器未来 / Paddle

24 5月, 2019 1 次提交

Conv concat relu quantization (#17466) · 5b2a3c4b

由 Sylwester Fraczek 提交于 5月 24, 2019

* add conv_concat_relu fuse

test=develop

* add test code

test=develop

* added missing include with unordered_map

test=develop

* review fixes for wojtuss

test=develop

* remove 'should (not) be fused' comment statements

one of them was invalid anyway

test=develop

5b2a3c4b

22 5月, 2019 1 次提交

Enable the convolution/relu6(bounded_relu) fusion for FP32 on Intel platform. (#17130) · 2281ebf0

由 guomingz 提交于 5月 22, 2019

* Relu6 is the bottleneck op for Mobilenet-v2. As the mkldnn supports the conv/relu6 fusion, we implement it fusion via cpass way. Due to the int8 enabling for this fusion will be supported in MKLDNN v0.20, so this PR is focused on the fp32 optimization.

Below table shows the benchmark(FPS) which measured on skx-8180(28 cores)
Batch size | with fusion | without fusion
-- | -- | --
1 | 214.7 | 53.4
50 | 1219.727 | 137.280

test=develop

* Fix the format issue

test=develop

* Add the missing nolint comments.

test=develop

* Fix the typos.

test=develop

* Register the conv_brelu_mkldnn_fuse_pass for the MKLDNN engine.

test=develop

* Adjust the indentation.

test=develop

* Add the test_conv_brelu_mkldnn_fuse_pass case.

test=develop

* Slightly update the code per Baidu comments.
Let the parameter definition embedded into the code.
That's will make the code easy to understand.

test=develop

2281ebf0

20 5月, 2019 1 次提交
- T
  remove unused expected_kernel_cache_pass (#17486) · 32da5e9c
  由 Tao Luo 提交于 5月 20, 2019
```
test=develop
```
  32da5e9c
07 5月, 2019 1 次提交

石

Cherry-pick benchmark related changes from release/1.4 (#17156) · a72dbe9a

由石晓伟提交于 5月 07, 2019

* cherry-pick commit from 88770542

* cherry-pick commit from 3f0b97df

* cherry-pick from 16691:Anakin subgraph support yolo_v3 and faster-rcnn

(cherry picked from commit 8643dbc2)

* Cherry-Pick from 16662 : Anakin subgraph cpu support

(cherry picked from commit 7ad182e1)

* Cherry-pick from 1662, 16797.. : add anakin int8 support

(cherry picked from commit e14ab180)

* Cherry-pick from 16813 : change singleton to graph RegistBlock
test=release/1.4

(cherry picked from commit 4b9fa423)

* Cherry Pick : 16837 Support ShuffleNet and MobileNet-v2

Support ShuffleNet and MobileNet-v2, test=release/1.4

(cherry picked from commit a6fb066f)

* Cherry-pick : anakin subgraph add opt config layout argument #16846
test=release/1.4

(cherry picked from commit 8121b3ec)

* 1. add shuffle_channel_detect

(cherry picked from commit 6efdea89)

* update shuffle_channel op convert, test=release/1.4

(cherry picked from commit e4726a06)

* Modify symbol export rules

test=develop

a72dbe9a

23 4月, 2019 1 次提交
- L
  fix runtime_context_cache bug when gpu model has an op runs only on cpu · 490e7462
  由 luotao1 提交于 4月 23, 2019
```
test=develop
```
  490e7462
19 4月, 2019 1 次提交
- T
  disable runtime_context_cache pass by default · aa7b975b
  由 Tao Luo 提交于 4月 19, 2019
```
test=develop
```
  aa7b975b
09 4月, 2019 1 次提交
- T
  disable seqpool concat pass by default saving CI time · d6c1b5a7
  由 tensor-tang 提交于 4月 09, 2019
```
test=develop
```
  d6c1b5a7
03 4月, 2019 3 次提交
- L
  test_analyzer_int8 tests use default pass order · bd636a9e
  由 luotao1 提交于 4月 03, 2019
```
test=develop
```
  bd636a9e
- L
  
  merge confict, test=develop · b236091e
  由 lujun 提交于 4月 03, 2019
  
  b236091e
- Y
  
  fix identity temporarily (#15942) · 044ae249
  由 Yan Chunwei 提交于 4月 03, 2019
  
  044ae249
02 4月, 2019 1 次提交
- W
  
  fix repeating passes (#16606) · ec2750b3
  由 Wojciech Uss 提交于 4月 02, 2019
  
  ec2750b3
28 3月, 2019 2 次提交

Anakin ssd support · d065b5bf

由 nhzlx 提交于 3月 28, 2019

refine trt first run
add quant dequant fuse pass
omit simplify_anakin_priorbox_detection template
omit transpose_flatten_concat_fuse template
test=develop

d065b5bf

C-API quantization core 2 (#16396) · 09dfc7a2

由 Wojciech Uss 提交于 3月 27, 2019

* C-API quantization core

test=develop
Co-authored-by: NSylwester Fraczek <sylwester.fraczek@intel.com>

* Decouple Quantizer from AnalysisPredictor

test=develop

* fixes after review

test=develop

* renamed mkldnn quantize stuff

test=develop

* remove ifdef from header file

test=develop

09dfc7a2

21 3月, 2019 2 次提交
- L
  add expected_kernel_cache_pass · 056599a7
  由 luotao1 提交于 3月 21, 2019
```
test=develop
```
  056599a7
- W
  Add enabling quantization (#16326) · cbe2dbf0
  由 Wojciech Uss 提交于 3月 21, 2019
```
* Add enabling quantization

test=develop

* remove unused (here) function
```
  cbe2dbf0
20 3月, 2019 4 次提交
- N
  
  git cherry-pick from feature/anakin-engine: update anakin subgraph #16278 · 07dcf285
  由 nhzlx 提交于 3月 20, 2019
  
  07dcf285
- N
  
  cherry-pick from feature/anakin-engine: refine paddle-anakin to new interface. #16276 · c407dfa3
  由 nhzlx 提交于 3月 20, 2019
  
  c407dfa3
- N
  
  cherry-pick from feature/anakin-engine: deal the changing shape when using anakin #16189 · a25331bc
  由 nhzlx 提交于 3月 20, 2019
  
  a25331bc
- N
  cherry-pick from feature/anakin-engine: refine anakin subgraph. #16157 · 69d37f81
  由 nhzlx 提交于 3月 20, 2019
```
support change input size
```
  69d37f81
19 3月, 2019 1 次提交
- L
  add runtime_context_cache_pass · 82af8031
  由 luotao1 提交于 3月 19, 2019
```
test=develop
```
  82af8031
04 3月, 2019 1 次提交
- Y
  Add the include of cudnn.h to enable the use of CUDNN_VERSION. (#15961) · 2bdf4464
  由 Yiqun Liu 提交于 2月 28, 2019
```
test=develop
```
  2bdf4464
28 2月, 2019 1 次提交
- Y
  Add the include of cudnn.h to enable the use of CUDNN_VERSION. (#15961) · 1616c32a
  由 Yiqun Liu 提交于 2月 28, 2019
```
test=develop
```
  1616c32a
14 2月, 2019 1 次提交
- Y
  
  move passes to src to avoid different behavior in deployment (#15705) · 3a5d6e5e
  由 Yan Chunwei 提交于 2月 14, 2019
  
  3a5d6e5e
21 1月, 2019 1 次提交
- Y
  
  fea/infer memory optim2 (#14953) · 885c4e57
  由 Yan Chunwei 提交于 1月 21, 2019
  
  885c4e57
14 11月, 2018 1 次提交
- Y
  
  Combine Inference Analysis with IR (#13914) · 9f252e00
  由 Yan Chunwei 提交于 11月 14, 2018
  
  9f252e00

机器未来 / Paddle 与 Fork 源项目一致

机器未来 / Paddle
与 Fork 源项目一致