提交 · 50326563d5c9c6925f7ecb12e3c5a346d26f0fa0 · Crayon鑫 / Paddle

04 6月, 2019 1 次提交
- L
  enable mkldnn primitive reuse for platform reorder (#17826) · 50326563
  由 Leo Zhao 提交于 6月 04, 2019
```
test=develop
```
  50326563
03 6月, 2019 2 次提交
- W
  revise the cudnn conv choose algorithm to improve the performance(mask rcnn benchmark) (#17753) · c10157a5
  由 wangchaochaohu 提交于 6月 03, 2019
```
* revise conv layer cudnn algo choose test=develop

* update for code style test=develop

* update for code style test=develop
```
  c10157a5
- C
  polish error doc (#17772) · 863c7516
  由 chengduo 提交于 6月 03, 2019
```
test=develop
```
  863c7516
29 5月, 2019 1 次提交
- G
  
  fix 2dconn test=develop (#17681) · 0d561ef4
  由 gongweibao 提交于 5月 29, 2019
  
  0d561ef4
27 5月, 2019 1 次提交
- G
  
  Add multi-ncclcomm and 2D ncclallreduce support. (#17263) · 65bbf950
  由 gongweibao 提交于 5月 27, 2019
  
  65bbf950
24 5月, 2019 2 次提交
- W
  add __str__ method for tensor and lodtensor to support print test=dev… (#17588) · 6724a652
  由 wopeizl 提交于 5月 24, 2019
```
* add __str__ method for tensor and lodtensor to support print test=develop
```
  6724a652
- M
  [NGraph] Enable assign operator for a ngraph, test=develop (#17437) · f2694e12
  由 mozga-intel 提交于 5月 23, 2019
```
*  Enable assign operator for a ngraph, test=develop

* Cross_entropy operators needs to be updated
```
  f2694e12
23 5月, 2019 2 次提交

由 Zeng Jinle 提交于 5月 23, 2019

* Revert "Revert "Fix allocator bug""

This reverts commit 174d0d0b.

* Revert "fix travis ci"

This reverts commit 5656fa9f.

test=develop

* add inlined_vector.h, test=develop

* add inlined_vector_test,test=develop

c6189637

M

[NGraph] Enable reshape operator test=develop (#17512) · 109b5aed
由 mozga-intel 提交于 5月 22, 2019

109b5aed

22 5月, 2019 1 次提交

Enable the convolution/relu6(bounded_relu) fusion for FP32 on Intel platform. (#17130) · 2281ebf0

由 guomingz 提交于 5月 22, 2019

* Relu6 is the bottleneck op for Mobilenet-v2. As the mkldnn supports the conv/relu6 fusion, we implement it fusion via cpass way. Due to the int8 enabling for this fusion will be supported in MKLDNN v0.20, so this PR is focused on the fp32 optimization.

Below table shows the benchmark(FPS) which measured on skx-8180(28 cores)
Batch size | with fusion | without fusion
-- | -- | --
1 | 214.7 | 53.4
50 | 1219.727 | 137.280

test=develop

* Fix the format issue

test=develop

* Add the missing nolint comments.

test=develop

* Fix the typos.

test=develop

* Register the conv_brelu_mkldnn_fuse_pass for the MKLDNN engine.

test=develop

* Adjust the indentation.

test=develop

* Add the test_conv_brelu_mkldnn_fuse_pass case.

test=develop

* Slightly update the code per Baidu comments.
Let the parameter definition embedded into the code.
That's will make the code easy to understand.

test=develop

2281ebf0

20 5月, 2019 1 次提交
- Q
  Fix compiling error with cuDNN 5.1 (#17458) · 97f0ec23
  由 qingqing01 提交于 5月 20, 2019
```
test=develop
```
  97f0ec23
15 5月, 2019 1 次提交
- Z
  
  fix_dygraph_mem_leak, test=develop (#17396) · eab34b2d
  由 Zeng Jinle 提交于 5月 15, 2019
  
  eab34b2d
10 5月, 2019 1 次提交

Double backward of conv2d. (#17211) · e32c9888

由 qingqing01 提交于 5月 10, 2019

* Add conv2d_grad_grad_op
* Extracte the cuDNN conv algo searching code in conv_cudnn_helper.h.
    - Now use it in conv2d_grad_grad.
    - Will simply the searching code in conv2d and conv2d_grad in next PR.
* Enhance and fix bug in unit testing of gradient_checker.
* Support to fetch empty variables，return None in Python.

e32c9888

08 5月, 2019 3 次提交

Refine elementwise kernel. (#16952) · 792443ef

由 zhaoyuchen2018 提交于 5月 08, 2019

* Refine elementwise kernel.

Add a simple cuda kernel if grad x and y both exist
Use 2D block cuda kernel to do broadcast.

test=develop
Signed-off-by: Nzhaoyuchen <zhaoyuchen01@baidu.com>

* refine code.

test=develop
Signed-off-by: Nzhaoyuchen <zhaoyuchen01@baidu.com>

* refine code.

test=develop
Signed-off-by: Nzhaoyuchen <zhaoyuchen01@baidu.com>

792443ef

C
update assert (#17282) · db5e74ab
由 chengduo 提交于 5月 08, 2019
```
test=develop
```
db5e74ab

Adding lrn op for ngraph engine (#17189) · 7bd1d03e

由 baojun 提交于 5月 07, 2019

* added lrn op test=develop

* Added CreateConstant method test=develop

* avoid duplicates test=develop

7bd1d03e

07 5月, 2019 1 次提交
- T
  remove unused FLAGS_warpctc_dir (#17162) · ff1661f1
  由 Tao Luo 提交于 5月 07, 2019
```
* remove unused FLAGS_warpctc_dir

test=develop

* remove FLAGS_warpctc_dir

test=develop
```
  ff1661f1
30 4月, 2019 1 次提交
- H
  Fix a typo in gpu_info.cc (#17175) · e4a53324
  由 Huihuang Zheng 提交于 4月 30, 2019
```
test=develop
```
  e4a53324
28 4月, 2019 1 次提交

Use CudnnWorkspaceHandle in exhaustive search (#17082) · b9494058

由 Huihuang Zheng 提交于 4月 28, 2019

1. Use CudnnWorkspaceHandle in exhaustive search of conv_cudnn.
2. For Ops using CudnnWorkspaceHandle in exhaustive search, release their GPU memory after exhaustive search.

test=develop

b9494058

23 4月, 2019 1 次提交
- Z
  Make conv cudnn workspace size configurable (#17036) · 0c335dcd
  由 Zeng Jinle 提交于 4月 23, 2019
```
* make_conv_cudnn_ws_size_configurable, test=develop

* change std::max to std::min
test=develop
```
  0c335dcd
21 4月, 2019 1 次提交

Refine model gpu memory (#16993) · 1202d3fc

由 Zeng Jinle 提交于 4月 21, 2019

* speedup gc and inplace softmax_with_cross_entropy_grad
test=develop

* refine models gpu mem
Merge skip vars and warning messages of mem opt
remove relu mem opt
test=develop

* follow comments
test=develop

1202d3fc

18 4月, 2019 1 次提交
- G
  
  Polish DGC code (#16818) · cbdb8a17
  由 gongweibao 提交于 4月 18, 2019
  
  cbdb8a17
16 4月, 2019 2 次提交

X
fix infershape bug · 5663fbfb
由 xuezhong 提交于 4月 16, 2019
```
test=develop
```
5663fbfb

[MKL-DNN] Added reusing of primitive descriptors (fp32) (#16667) · 87a44b11

由 Jacek Czaja 提交于 4月 15, 2019

* - Reuse of conv PD

- conv transpose pd reused

- Added PD reusing of softmax and Batch Norm

- Refactoring and removal of not needed routines of mkl-dnn ops

test=develop

- Fix to reusing conv

test=develop

- Lint fixes

test=develop

- Further lint fixes

test=develop

- Lint  fixes

test=develop

- lint fixes

test=develop

- Lint workaround

test=develop

* - Fix after review on including boost as third party header

test=develop

* - Fix after review. Name change to something more descriptive

test=develop

87a44b11

11 4月, 2019 1 次提交
- D
  make lodtensor_printer usable in gpu setting · a659b37a
  由 dongdaxiang 提交于 4月 11, 2019
```
test=develop
```
  a659b37a
03 4月, 2019 1 次提交
- C
  Revert "Model data cryption link all lib (#16555)" · 0b2aec14
  由 Chen Weihang 提交于 4月 03, 2019
```
test=develop
This reverts commit c38c7c56.
```
  0b2aec14
02 4月, 2019 1 次提交

Model data cryption link all lib (#16555) · c38c7c56

由 Chen Weihang 提交于 4月 02, 2019

* link the libwbaes.so into paddle

* polish detail, test=develop

* try fix mac_pr_ci error, test=develop

* add compile option, test=develop

* fix ci error, test=develop

* ignore failed to find mac lib, test=develop

* change cdn to bj, cdn can't get the latest version

* trigger ci, test=develop

* temporary delete win32 lib linking, test=develop

* change https to http, test=develop

* turn compile option on to off

* turn compile option off to on, test=develop

* try lib compiled by gcc4.8, test=develop

* update lib version, test=develop

* link other lib, test=develop

* add setup config

* delete false, test=develop

* delete no_soname, test=develop

* recover so name set

* fix, test=develop

* adjust make config, test=develop

* remove link to wbaes, test=develop

* remove useless define, test=develop

c38c7c56

30 3月, 2019 1 次提交
- G
  Fix windows compilation error! (#16546) · fea91164
  由 gongweibao 提交于 3月 30, 2019
```
* fix compiled
test=develop

* follow comments test=develop
```
  fea91164
29 3月, 2019 12 次提交
- D
  refine API spec · 3a79be6e
  由 dongdaxiang 提交于 3月 29, 2019
```
test=develop
```
  3a79be6e
- D
  fix pull sparse slow problem · 98dda08a
  由 dongdaxiang 提交于 3月 29, 2019
```
test=develop
```
  98dda08a
- D
  fix dataset testcase problem · 93c3c7f9
  由 dongdaxiang 提交于 3月 28, 2019
```
test=develop
```
  93c3c7f9
- D
  fix async_executor problem and remove some unnecessary testcase, fix trainer_desc import problem · d739bab8
  由 dongdaxiang 提交于 3月 28, 2019
```
test=develop
```
  d739bab8
- D
  fix windows compile problem · e3107a6a
  由 dongdaxiang 提交于 3月 26, 2019
```
test=develop
```
  e3107a6a
- D
  disable sys/wait.h to fix windows compile problem, include scope in lodtensor_printer · 398004ec
  由 dongdaxiang 提交于 3月 26, 2019
```
test=develop
```
  398004ec
- D
  move root_scope->DropKids() into Finalize() so that we do not have to drop all the kids · 39362a84
  由 dongdaxiang 提交于 3月 24, 2019
```
test=develop
```
  39362a84
- D
  
  fix code style · a0b59773
  由 dongdaxiang 提交于 3月 23, 2019
  
  a0b59773
- D
  support win32 flag in io.cc shell.cc, fix code style problem in fleet_wrapper,... · 365be5d5
  由 dongdaxiang 提交于 3月 23, 2019
```
support win32 flag in io.cc shell.cc, fix code style problem in fleet_wrapper, fix lodtensor_printer_test problem
test=develop
```
  365be5d5
- D
  add more example on datagenerator · dc8cf36e
  由 dongdaxiang 提交于 3月 23, 2019
```
test=develop
```
  dc8cf36e
- D
  
  refine print fetch list · 6bf796df
  由 dongdaxiang 提交于 3月 21, 2019
  
  6bf796df
- D
  
  add printer for fetch variable · cf136064
  由 dongdaxiang 提交于 2月 18, 2019
  
  cf136064

Crayon鑫 / Paddle 与 Fork 源项目一致

Crayon鑫 / Paddle
与 Fork 源项目一致