提交 · 436808c6981be3fb808bb22794ee2885d7cd257e · Crayon鑫 / Paddle

23 11月, 2021 7 次提交

Z

elu support alpha < 0 (#37316) (#37437) · 436808c6
由 zhupengyang 提交于 11月 23, 2021

436808c6

cherry pick save/load in the_one_ps (#37461) · 58a51130

由 wangguanqun 提交于 11月 23, 2021

* save/load in ps runtime(the_one_ps) (#36097)

* add trainer desc config to distributed strategy

* code style modified

* data_feed set lod

* fix bug

* code style

* fix bug

* save load

* save load

* save unittest

* add unittest of the_one_ps

* unittest

* add todo in communicator sendsparse

* fix bug in save_inference_model (#37362)

58a51130

S
[Cherry-pick 2.2]Enhance error message of scatter op (#37431) · d5e73f07
由 sneaxiy 提交于 11月 23, 2021
```
* enhance scatter err msg check

* fix ci error
```
d5e73f07

[Dy2stat]Allow users to switch eval/train mode when using @to_static to... · eed736dc

由 0x45f 提交于 11月 23, 2021

[Dy2stat]Allow users to switch eval/train mode when using @to_static to decorate a function (#37383) (#37432)

本PR之前使用@to_static装饰一个单独的function时，对于生成的Program无法切换train/eval模式，只能运行在train模式下。这也就导致动转静后用户多次调用function显存会一直增长。
本PR之后，使用@to_static装饰一个单独的function时，可以通过function.train()或者function.eval()的方式来切换train/eval模式。

eed736dc

W

fix shape api (#37412) · 2778fcd9
由 Wilber 提交于 11月 23, 2021

2778fcd9
J

fix memeory_optimize_pass bug (#37324) (#37413) · 0fa96e91
由 JingZhuangzhuang 提交于 11月 23, 2021

0fa96e91
J

cherry_pick conv2d omission (#37419) · 6b3ffe99
由 JingZhuangzhuang 提交于 11月 23, 2021

6b3ffe99

22 11月, 2021 3 次提交
- C
  Fix a bug of quantization (#36982) (#37381) · 9ffb43be
  由 ceci3 提交于 11月 22, 2021
```
* fix a quantization bug
Co-authored-by: NXGZhang <46363693+XGZhang11@users.noreply.github.com>
```
  9ffb43be
- S
  [cherry-pick] Add paddle.incubate.graph_send_recv API(#37205) (#37343) · 109f8a8e
  由 Siming Dai 提交于 11月 22, 2021
```
* Add paddle.incubate.graph_send_recv API

* fix bug in CudaAtomicMin and CudaAtomicMax

* add empty line
```
  109f8a8e
- L
  fix bug to support dropout eval grad computing. (#37305) (#37331) · 604b6fc0
  由 Li Min 提交于 11月 22, 2021
```
fix bug to support dropout eval grad computing. cherry-pick #37305.
```
  604b6fc0
19 11月, 2021 3 次提交

0
[Dy2stat]Support `for i in [1,2,3]` statements in dy2stat (#37259) (#37356) · 44db219a
由 0x45f 提交于 11月 19, 2021
```
该PR使得动转静模块能够正确转换如下的for i in [1, 2, 3]语句。
```
44db219a
0
set net.forward to original forward function in flops (#36852) (#37357) · b5594759
由 0x45f 提交于 11月 19, 2021
```
set net.forward to original forward function in flops when net is a dy2stat model.
```
b5594759

[cherry-pick]Add sparse attention doc warning (#37189) · 5fd8312d

由 Liu-xiandong 提交于 11月 19, 2021

* fix cusparse compile bug in CUDA11.2, test=develop

* modify sparse_attention docs, test=document_fix (#36554)

* modify sparse_attention docs, test=develop

* add warning

* add warning ,test=document_fix

5fd8312d

17 11月, 2021 4 次提交
- W
  [Paddle-Inference] fix_qkv_plugin: fix half scale (#37096) (#37264) · 027664e8
  由 Wangzheee 提交于 11月 17, 2021
```
* fix_qkv_plugin: half_scale

* [Paddle-Inference] fix_qkv_plugin: fix half scale
```
  027664e8
- W
  
  cherrypick fix op_teller (#37266) · 3fdbab25
  由 Wangzheee 提交于 11月 17, 2021
  
  3fdbab25
- W
  
  add_fc_convert_layers_name (#37157) (#37267) · 71b04f61
  由 Wangzheee 提交于 11月 17, 2021
  
  71b04f61
- J
  
  add fetch to black list (#37123) (#37261) · 8cb370f7
  由 JingZhuangzhuang 提交于 11月 17, 2021
  
  8cb370f7
16 11月, 2021 3 次提交

[cherry-pick-2.2.1]fix fused_transformer_encoder_layer bug (#37229) · 36dd295e

由 zhangkaihuo 提交于 11月 16, 2021

修复了fused_transformer_encoder_layer fine-tune过程发现的一些问题：

    fused_attention_op添加attn_mask=None的支持：PR
    pre_layer_norm处理问题：PR
    参数处理，计算错误的问题：PR
    add_bias计算错误问题：PR
    添加pure fp16的支持：PR

36dd295e

Z
fix bug of indexing with ellipsis (#37192) · 79b9f47e
由 zyfncg 提交于 11月 16, 2021
```
修复了一维Tensor在使用省略号(...)索引时维度检测异常的问题。
```
79b9f47e
石
clean inference logs when config.DisableGlogInfo is triggered (#36356) (#37212) · dc873eba
由石晓伟提交于 11月 16, 2021
```
Co-authored-by: NPei Yang <peiyang@baidu.com>
```
dc873eba

15 11月, 2021 1 次提交
- Z
  MLPerf Optimization for Release/2.2 (#37109) · 287ca7d5
  由 Zeng Jinle 提交于 11月 15, 2021
```
* add mlperf optimization PRs

* update
```
  287ca7d5
10 11月, 2021 1 次提交
- J
  Fix rnn grad bug in cpu when dropout is zero (#37080) (#37086) · 70cb0a54
  由 Jack Zhou 提交于 11月 10, 2021
```
* fix rnn grad bug when num_layers is set 2 and dropout_prob is set 0

* add more test for rnn
```
  70cb0a54
08 11月, 2021 2 次提交
- W
  Optimized the solve op code:renamed var and removed template func (#36981) (#37011) · a787b278
  由 Weilong Wu 提交于 11月 08, 2021
```
    Renamed the variable and function
    Removed the original template function
    Removed the tests_properties in CMakeLists.txt
```
  a787b278
- Z
  setitem support passing stop_gradient from value to tensor (#37028) · 76cab751
  由 zyfncg 提交于 11月 08, 2021
```
att,Fix issue:36902
```
  76cab751
01 11月, 2021 2 次提交
- L
  [cherry-pick]fix cusparse compile bug in CUDA11.2, test=release/2.2 (#36913) · ab2004bb
  由 Liu-xiandong 提交于 11月 01, 2021
```
* fix cusparse compile bug in CUDA11.2, test=develop

* fix bug
```
  ab2004bb
- F
  
  negative label in softmax cross entropy (#36907) · dcadc256
  由 Feng Xing 提交于 11月 01, 2021
  
  dcadc256
30 10月, 2021 1 次提交
- Y
  Move the ASP training API to paddle.static.sparsity. (#36525) (#36860) · 09bc9c06
  由 Yiqun Liu 提交于 10月 30, 2021
```
Cherry-pick #36525
```
  09bc9c06
29 10月, 2021 2 次提交
- W
  
  tmp disable windows demo_ci ut (#36847) · f2daef50
  由 Wilber 提交于 10月 29, 2021
  
  f2daef50
- F
  1. fix ifftshift(missing negative sign before shifts); (#36835) · fa7aa6b8
  由 Feiyu Chan 提交于 10月 29, 2021
```
2. add complex data type support for paddle.shape at graph assembly.
```
  fa7aa6b8
28 10月, 2021 11 次提交
- 0
  
  polish _remove_no_value_return_var() function (#36826) (#36830) · c716cf35
  由 0x45f 提交于 10月 28, 2021
  
  c716cf35
- P
  【Cherry-pick PR 36511】fix out_of_range bug of multinomial op's cuda kernel (#36511) (#36808) · d8ffb261
  由 pangyoki 提交于 10月 28, 2021
```
Cherry-pick PR #36511
```
  d8ffb261
- Z
  
  fix dygraph adamw (#36745) (#36794) · e3db65d5
  由 zhaoyingli 提交于 10月 28, 2021
  
  e3db65d5
- L
  fix device docs;test=document_fix (#36784) (#36827) · 0b7f43ec
  由 Ligoml 提交于 10月 28, 2021
```
* fix device docs;test=document_fix

* update __init__.py
```
  0b7f43ec
- P
  Cherry-pick-36556: add paddle.version.cuda and paddle.version.cudnn API (#36556) (#36795) · 05b8630f
  由 pangyoki 提交于 10月 28, 2021
```
* add paddle.version.cuda and paddle.version.cudnn API

* fix little bug

* fix bug

* add doc string

* fix mkdir error

* fix windows path

* fix new paddle/version path

* fix unittest

* fix format
```
  05b8630f
- X
  
  Update quant_conv2d_dequant_fuse_pass.cc (#36821) · 7647d402
  由 XGZhang 提交于 10月 28, 2021
  
  7647d402
- X
  [cherry-pick 2.2]support quantization of bert (#36820) · f20c5c9c
  由 XGZhang 提交于 10月 28, 2021
```
* [cherry-pick 2.2]support quantization of bert

support quantization for maumul_v2

* Update quantization_pass.py
```
  f20c5c9c
- L
  [fix-doc-bug] Fix fused_attention_op english doc test=document_fix (#36803) (#36829) · 9a964901
  由 Li Min 提交于 10月 28, 2021
```
* Fix fused_attention english doc test=document_fix
```
  9a964901
- F
  change api to support trt8 in pool3d_op_convert (#36783) (#36812) · 5fb28500
  由 feng_shuai 提交于 10月 28, 2021
```
* change api for support trt8
```
  5fb28500
- H
  [Cherry-pick] Enable CTC grad compute on GPU (#36780) · 8ede9e6f
  由 Hui Zhang 提交于 10月 28, 2021
```
* Revert "Align CTC grad scale same with ESPNet (#34729)"

This reverts commit 10f9644c.

* ctc grad compute on gpu
```
  8ede9e6f
- L
  Fix fused_attention_op and fused_feedforward_op bug when pre_layer_norm is false. (#36793) (#36816) · ae592233
  由 Li Min 提交于 10月 28, 2021
```
* Fix bug when pre_layer_norm is false.
```
  ae592233

Crayon鑫 / Paddle 与 Fork 源项目一致

Crayon鑫 / Paddle
与 Fork 源项目一致