提交 · fa9d3fa5bf8ccab35a499ace24f08a32cc9c03c0 · Crayon鑫 / Paddle

16 10月, 2020 5 次提交

Incorporate cudnn_lstm into LSTM api (#27217) · fa9d3fa5

由 Guo Sheng 提交于 10月 16, 2020

* Incorporate cudnn_lstm into LSTM api.
test=develop

* Make coalesce_tensor support alignment optionally.
test=develop

* Reorganize RNN apis. test=develop

* Fix cudnn rnn layout conversion.
test=develop

* Add sequence_length support for RNN cudnn implement.
Add optional init_h and init_c gradient for cudnn_lstm_op.
test=develop

* Use create_parameter for rnn cudnn impl.
test=develop

* Move `self._flat_weight = self.create_parameter()` in RNNBase to main_program.
test=develop

* Update RNN api unittest to use set_device.
test=develop

* Fix set_place for unit tests of RNN apis.
test=develop

* Fix use_align in coalesce_tensor_op.
test=develop

* Adjust RNN apis arguments according to comments.
test=develop

* Polish documents for SimpleRNN apis.
test=develop

* Refine random seed in cudnn_lstm_op.
Expose rnn params from sublayers to RNN.
test=develop

* Fix RNN saving for jit.save.
Refine cudnn_lstm dropout behavior.
test=develop

* Fix doc of GRU. test=develop

* Use ShareDataWith to avoid copying for cudnn_lstm_op test.
test=develop

* Remove updates on cudnn_lstm temporarily.
test=develop

* Use ShareDataWith to avoid copying for cudnn_lstm_op test.
test=develop

* Refine random seed in cudnn_lstm_op.
test=develop

* Fix test_lstm by adjust ConcreteProgram buffer getter.
test=develop

* Use create_parameter instead of create_var for rnn._flat_weight for static graph usage.
test=develop

* Remove W input for cudnn_lstm to pass unused_var_check.
test=develop

* Add test_predict for RNN unit tests coverage.
test=develop

* Fix code style of rnn.
test=develop

* Fix F.rnn usage in rnn.py.
test=develop

fa9d3fa5

G

error message optimization in mean_xpu,softmax_with_cross_entropy_op_xpu,test=kunlun (#27967) · f94d0537
由 Guanghua Yu 提交于 10月 16, 2020

f94d0537

Fix xpu enforce (#27978) · d330cf66

由 Jack Zhou 提交于 10月 16, 2020

* test=kunlun;

Add elementwise XPU OP kernel for KUNLUN core, including (but still cannot process common broadcast):

    * elementwise_div op
    * elementwise_max op
    * elementwise_mul op (with grad op)
    * elementwise_sub op (with grad op)

* 0.05->0.01

* add xpu error message description;test=kunlun

d330cf66

[oneDNN] Conv dilation support (#27914) · 7cb4a8b8

由 lidanqing 提交于 10月 16, 2020

* conv dilated mkldnn support: forward and backward pass

* add mkldnn conv_transpose dilation UT
test=develop

* remove unnecessary PADDLE_ENFORCE

* add int8 and bf16 dilated conv UT

* update according to reviews

7cb4a8b8

M

fix kunlun kernel of reshape op (#27988) · 64c26349
由 mapingshuo 提交于 10月 16, 2020

64c26349

15 10月, 2020 6 次提交

T
Feature/large scale kv save base/delta (#27470) · 202bfab1
由 tangwei12 提交于 10月 15, 2020
```
* add size method for large scale

* add large scale UT

* add ut for checkpoint
```
202bfab1

【paddle.fleet】geo send sparse optimize (#27719) · aa3b4ed7

由 123malin 提交于 10月 15, 2020

* test=develop, fix geo sgd communicator

* test=develop, gloo_init_method

* test=develop, bug fix for gloo http_init

aa3b4ed7

M

reshape support bool, test=develop (#27944) · 5ccaaab8
由 mapingshuo 提交于 10月 15, 2020

5ccaaab8

Add reduce sum and reduce mean xpu op (#27939) · 4a4f7736

由 Qinghe JING 提交于 10月 15, 2020

* add reduce xpu op test=develop;test=kunlun

* add reduce xpu op test=develop;test=kunlun

* add reduce xpu op test=develop;test=kunlun

* add reduce xpu op test=develop;test=kunlun

* add reduce xpu op test=develop;test=kunlun

4a4f7736

Z
add tensor clone (#27953) · bf412f46
由 Zhou Wei 提交于 10月 15, 2020
```
* add tensor clone

* fix unittest test_var_base
```
bf412f46

support channel last in BatchNorm*d · 2e845182

由 Feiyu Chan 提交于 10月 15, 2020

1. support channel last in BatchNorm*d (#27875)
2. fix a bug in batch_norm_op cuda kernel by extracting ResizeToChannelFist(Last), TransToChannelFirst(Last) to operators/layer_utils.h

2e845182

14 10月, 2020 17 次提交
- L
  Support setting xpu place in dygraph mode (#27909) · 9a2a4b5f
  由 Leo Chen 提交于 10月 14, 2020
```
* support setting xpu place

* add ut, test=kunlun
```
  9a2a4b5f
- M
  Fix adam (#27778) · 263a9e97
  由 MRXLT 提交于 10月 14, 2020
```
* fix adam

* fix gpu adam

* fix code style

* fix ut

* update ut add cuda code
```
  263a9e97
- D
  kunlun add op (#27890) · b0edda4d
  由 Double_V 提交于 10月 14, 2020
```
* add stack pool2d roi_align xpu op,test=kunlun

* error message opt, test=kunlun

* add xpu unittest,test=kunlun

* skip check grad,test=kunlun

* fix boostget , test=kunlun
```
  b0edda4d
- J
  Add elementwise XPU OP kernel for KUNLUN core, including (but still cannot process common broadcast · c791df09
  由 Jack Zhou 提交于 10月 14, 2020
```
Add elementwise XPU OP kernel for KUNLUN core, including (but still cannot process common broadcast
```
  c791df09
- W
  
  xpu support for fill_constant Op (#27675) · c5fcc96d
  由 wangchaochaohu 提交于 10月 14, 2020
  
  c5fcc96d
- C
  【paddle.fleet】fix sparse load (#27680) · 328cb289
  由 Chengmo 提交于 10月 14, 2020
```
* add sparse tensor load method
```
  328cb289
- T
  
  fix paddle error informations (#27889) · cf70d5b3
  由 tangwei12 提交于 10月 14, 2020
  
  cf70d5b3
- W
  update the code for the topk message optimize · 95aa5342
  由 wawltor 提交于 10月 14, 2020
```
update the code for the topk message optimize 
```
  95aa5342
- C
  Polish some error message in opeators (#27876) · 4ba977c7
  由 Chen Weihang 提交于 10月 14, 2020
```
* polish some error message

* add white list

* revert shell script change
```
  4ba977c7
- 1
  【paddle.fleet】bug fix for parameter_recv (#27838) · a4f85074
  由 123malin 提交于 10月 14, 2020
```
* test=develop, bug fix for parameter_recv
* test=develop, for unittest, test_fleet_rolemaker_new
```
  a4f85074
- Q
  support kunlun matmul_v2 (#27910) · 2712d076
  由 QingshuChen 提交于 10月 14, 2020
```
*test=kunlun
```
  2712d076
- Z
  Multi task (#26002) · 5a83496c
  由 zhang wenhui 提交于 10月 14, 2020
```
* add multitask

* add multitask, test=develop

* fix code style, test=develop

* add partail push dense, test=develop

* fix has_kay in py3, test=develop

* fix, test=develop

* fix, test=develop

* fix, test=develop
```
  5a83496c
- Z
  fix norm api doc, test=develop (#27652) · 7a58431c
  由 zhang wenhui 提交于 10月 14, 2020
```
* fix norm api doc, test=develop

* fix error message, test=develop

* fix api norm, test=develop

* add adagrad, test=develop

* fix bug, test=develop

* fix bug, test=develop

* add spetral_norm, test=develop

* fix adagrad, test=develop

* merge , test=develop
```
  7a58431c
- Y
  Lookup table v2 xpu (#27888) · 3eb106da
  由 yinhaofeng 提交于 10月 14, 2020
```
* add lookup_table_v2_op_xpu, test=kunlun

* add lookup_table_v2_op_xpu, test=kunlun

* change some Tips ,test=kunlun
```
  3eb106da
- Z
  tune backward filter algorithm for float16 (#27529) · d5cc144c
  由 Zhang Ting 提交于 10月 14, 2020
```
* use exhaustive_search for float16

* tune algo only when dtype is float16
```
  d5cc144c
- H
  
  fix error msg (#27887) · 3f2a6ab6
  由 hutuxian 提交于 10月 14, 2020
  
  3f2a6ab6
- X
  Add dropout and log_loss for kunlun (#27790) · ae01801f
  由 xiaoting 提交于 10月 14, 2020
```
* add dropout,log_loss, test=kunlun
* fix dropout, test=kunlun
* polish error message, test=kunlun
* change boost::get to BOOST_GET_CONST, test=kunlun
* fix copyright, test=kunlun
```
  ae01801f
13 10月, 2020 8 次提交
- G
  support mean,softmax_with_cross_entropy on Baidu Kunlun (#27792) · 70c8c313
  由 Guanghua Yu 提交于 10月 13, 2020
```
* support mean,softmax_with_cross_entropy on Baidu Kunlun,test=kunlun

* fix unittests error,test=kunlun

* delete boost::get,test=kunlun
```
  70c8c313
- C
  add xpu sgd & momentum (#27728) · 1607e87c
  由 Chengmo 提交于 10月 13, 2020
```
* add xpu sgd & momentum
```
  1607e87c
- H
  
  Add batch_norm and layer_norm XPU kernels (#27818) · c90d3556
  由 hong19860320 提交于 10月 13, 2020
  
  c90d3556
- X
  add conv for xpu, test=kunlun (#27809) · 6da7a745
  由 xiaoting 提交于 10月 13, 2020
```
* add conv for xpu, test=kunlun

* polish error_message, test=kunlun

* polish error_message, test=kunlun

* fix copyrigth, test=kunlun
```
  6da7a745
- T
  add xpu slice op (#27349) · 04be37c5
  由 Thunderbrook 提交于 10月 13, 2020
```
* add xpu slice op
test=xpu

* add slice xpu op
test=xpu

* code style
test=kunlun

* style
test=kunlun

* format
test=kunlun
```
  04be37c5
- T
  op error info (#27856) · 8c25dfaa
  由 Thunderbrook 提交于 10月 13, 2020
```
* op error info

* style

* code format
```
  8c25dfaa
- S
  add gather_op xpu, test=kunlun (#27822) · 6d63cd2b
  由 ShenLiang 提交于 10月 13, 2020
```
* add gather_op xpu, test=develop, test=kunlun

* fix ut, test=develop, test=kunlun

* fix the ut,test=develop, test=kunlun
```
  6d63cd2b
- F
  
  fix error message for nce_op (#27863) · 1d95a0fb
  由 Feiyu Chan 提交于 10月 13, 2020
  
  1d95a0fb
12 10月, 2020 4 次提交
- G
  Refine the gradient calculation errors caused by renaming in while_grad (#27814) · 2e1bca99
  由 guofei 提交于 10月 12, 2020
```
test=develop
```
  2e1bca99
- W
  add load_op_xpu for Baidu Kunlun (#27817) · 8fa4c098
  由 wanghuancoder 提交于 10月 12, 2020
```
* add load_op_xpu for Baidu Kunlun, test=kunlun

* add is_compiled_with_xpu for unit test, test=kunlun

* add is_compiled_with_xpu for unit test, test=kunlun
```
  8fa4c098
- J
  
  [oneDNN] adaptive pool support (#27747) · 55e63763
  由 Jacek Czaja 提交于 10月 12, 2020
  
  55e63763
- Z
  use IndexList to improve performance of instance_norm op (#25132) · 16999ae4
  由 Zhang Ting 提交于 10月 12, 2020
```
* use IndexList to improve performance, test=develop

* remove EIGEN_HAS_INDEX_LIST, test=develop

* use IndexList only when EIGEN_HAS_INDEX_LIST is true
```
  16999ae4

Crayon鑫 / Paddle 与 Fork 源项目一致

Crayon鑫 / Paddle
与 Fork 源项目一致