提交 · 0fff9306676ca2256de8cdd60eb0d30878521b95 · BaiXuePrincess / Paddle

04 3月, 2021 3 次提交

L

Fix bug for set_value op when input dtype is not float32 (#31411) · 0fff9306
由 liym27 提交于 3月 04, 2021

0fff9306

[Dy2stat] Fix Read-Only Attribute as while_loop Output (#31415) · 6bf02a12

由 Huihuang Zheng 提交于 3月 04, 2021

Fix Read-Only Attribute as while_loop Output:

Usually, our convert_while_loop will be like:
```
    [a, b, c] = paddle.jit.dy2static.convert_while_loop(
            condition_name, body_name, [a, b, c])
```
where a, b, c are in loop_var_names.

However, if loop_var_names contains property such as foo.x, we cannot
assign the attribute as output of convert_while_loop because Python
property is a kind of read-only attribute. To handle the case, we replace
the attributes which are output of convert_while_loop with generated
variables, then if we know the attribute is not read-only at runtime, we
assign the attribute. The created statements are like:
```
    [a, b, __attribute_variable_1] = paddle.jit.dy2static.convert_while_loop(
            condition_name, body_name, [a, b, foo.x])
    if not isinstance(getattr(type(foo), x, None), property): foo.x = __attribute_variable_1
```

6bf02a12

J

Added LSTM BF16 and fixed GRU BF16 (#31234) · 5b4f8aac
由 jakpiase 提交于 3月 04, 2021

5b4f8aac

03 3月, 2021 5 次提交
- Q
  
  [ROCM] fix softmax with loss and update python scripts, test=develop (#31373) · db50fb67
  由 Qi Li 提交于 3月 03, 2021
  
  db50fb67
- P
  
  TRT conv2d converter support SAME padding (#31379) · 32211fe9
  由 Pei Yang 提交于 3月 03, 2021
  
  32211fe9
- Z
  
  [Custom OP]polish doc of custom OP (#31369) · 13e4280f
  由 Zhou Wei 提交于 3月 03, 2021
  
  13e4280f
- Q
  
  [ROCM] update fluid operators for rocm (part6), test=develop (#31301) · 946dbdae
  由 Qi Li 提交于 3月 03, 2021
  
  946dbdae
- W
  Add attrs `deformable_groups` for deformable_conv API (#31335) · 1cbccfa5
  由 wangna11BD 提交于 3月 03, 2021
```
* add attrs deformable_groups
```
  1cbccfa5
02 3月, 2021 2 次提交

P
add n-d input support for trt scale converter (#31316) · 2e9e3fad
由 Pei Yang 提交于 3月 02, 2021
```
* add n-d input support for trt scale converter

* add flatten for ut

* fix dims
```
2e9e3fad

lamb_op_xpu;test=kunlun (#31012) · d79fdc3d

由 Gradie 提交于 3月 02, 2021

* lamb_op_xpu;test=kunlun

* modify lamb_op_xpu.cc;test=kunlun

* delete atol lamb_op_xpu; test=kunlun

* update xpu.cmake;test=kunlun

* test_error 1e-5,lamb_op_xpu;test=kunlun

* error1e-5,lamb_op_xpu,test=kunlun

* delete atol lamb_xpu;test=kunlun

* modify atol,lamb_op_xpy;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu, XPUOptest;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu,modify xpu_cmake; test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu,modify xpucmake;test=kunlun

d79fdc3d

28 2月, 2021 2 次提交
- Z
  
  fix test_check_abi (#31288) · aebf2234
  由 Zhou Wei 提交于 2月 28, 2021
  
  aebf2234
- Z
  
  [Custom OP]add MSVC compile check on Windows (#31265) · cc89120a
  由 Zhou Wei 提交于 2月 28, 2021
  
  cc89120a
27 2月, 2021 1 次提交

[Custom OP]add PD_THROW and PD_CHECK for User Error message (#31253) · af9066e8

由 Zhou Wei 提交于 2月 27, 2021

* [Custom OP]add PD_THROW and PD_CHECK for User error message

* PD_THROW and PD_CHECK, fix comment

* fix Windows error message

* fix Windows error message

* fix CI

af9066e8

26 2月, 2021 6 次提交
- J
  
  [Custom OP] Support stream set on Custom Op (#31257) · 038ce70d
  由 Jiabin Yang 提交于 2月 26, 2021
  
  038ce70d
- A
  [Dy2Stat] Fix eval_if_exist_else_none bug (#31261) · 1dd40870
  由 Aurelius84 提交于 2月 26, 2021
```
* fix eval_if_exist_else_none bug

* fix typo

* fix typo

* fix test_op_num unittest
```
  1dd40870
- J
  [Custom Op] Remove unsupport dtypes (#31232) · 0c38708a
  由 Jiabin Yang 提交于 2月 26, 2021
```
* remove remove_unsupport_dtype

* remove remove_unsupport_dtype

* remove test dtype

* add more include

* change dtype.h's enum as enum class to avoid conflict with inference lib

* make enum as enum class

* remove additional test

* merge develop

* polish code
```
  0c38708a
- W
  
  xpu support fuse allreduce (#31104) · b8bce682
  由 WangXi 提交于 2月 26, 2021
  
  b8bce682
- C
  [CustomOp] Split build op marco & polish details (#31229) · 126633c5
  由 Chen Weihang 提交于 2月 26, 2021
```
* split build op marco & polish details

* revert register api del

* fix other unittest
```
  126633c5
- A
  [CustomOp] Add Modeling with Custom op unittest (#31218) · e8d24b54
  由 Aurelius84 提交于 2月 26, 2021
```
* add unittest for static/dygraph/dy2stat

* add PE unittet

* remove usless code

* add unittest in CMakeList.txt
```
  e8d24b54
25 2月, 2021 4 次提交
- L
  add int pad support for Pad1D/2D/3D (#31209) · ad50fa71
  由 littletomatodonkey 提交于 2月 25, 2021
```
* add int pad support for Pad1D/2D/3D

* fix type

* fix format
```
  ad50fa71
- J
  
  OneDNN hardswish integration (#30211) · 2f116534
  由 jakpiase 提交于 2月 25, 2021
  
  2f116534
- A
  [CustomOp]Add cpp_extension en doc (#31187) · 912022fa
  由 Aurelius84 提交于 2月 25, 2021
```
* add cpp_extension en doc

* remove cuda_cflags and add optional in doc

* refine style

* fix indent problem

* add default None
```
  912022fa
- C
  [CustomOp] Support attributes as func input in custom op (#31128) · e8cdb49a
  由 Chen Weihang 提交于 2月 25, 2021
```
* add simple attr support and test

* add int, float attr support

* support other attribute

* add custom attrs test in cmake

* polish details

* fix test failed

* add backward test

* update test flags
```
  e8cdb49a
24 2月, 2021 8 次提交

[CustomOp] Support to specific extra_cflags and exctra_cuda_flags independently (#31059) · 406f4a75

由 Aurelius84 提交于 2月 24, 2021

* split cxx/nvcc compile flags

* enhance input argument check

* rename extra_cflags into extrac_cxx_flags

* add name checking in setup

* fix test_dispatch failed

* fix word typo and rm usless import statement

* refine import statement

* fix unittest failed

* fix cuda flags error

406f4a75

[Paddle-TRT] support group_norm (#31040) · 00b09e86

由 Pei Yang 提交于 2月 24, 2021

* add group norm plugin

* fix compile problems

* move concat axis check to trt op teller

* add nbDims for scale and bias nv dims

* add group norm unit test

* fix unittest

* add trt version restriction for group norm op teller

* fix unittest

00b09e86

C

change test_multiprocess_reader_exception cmake (#31174) · c209751c
由 Chen Weihang 提交于 2月 24, 2021

c209751c
Y

fix ut timeout (#31061) · 15312145
由 YUNSHEN XIE 提交于 2月 24, 2021

15312145
C
[CustomOp] Add new paddle custom op so (#31141) · 1ce96fa1
由 Chen Weihang 提交于 2月 23, 2021
```
* add new custom op so

* fix use new method error

* fix test failed
```
1ce96fa1

fix entry (#31079) · ebbdf525

由 tangwei12 提交于 2月 24, 2021

* fix entry

* fix distributed lookup table fuse case

* fix entry bug at first time

* move entry from paddle.fluid -> paddle.distributed

* fix ut with paddle.enable_static()
Co-authored-by: Nmalin10 <malin10@baidu.com>

ebbdf525

[Custom OP]Fix problem of custom op unitests on Windows CI (#31114) · 4b220550

由 Zhou Wei 提交于 2月 24, 2021

* fix some problem of Windows custom op

* fix some problem of Windows custom op

* fix some problem of Windows custom op

4b220550

add warning message when dtypes of operator are not same (#31136) · 70131b47

由 chentianyu03 提交于 2月 24, 2021

* add error msg when dtypes of operator are not same

* add error msg when dtypes of operator are not same

* change error msg to warning msg when dtypes of operator are not same

* modify test case to fit for python2

70131b47

23 2月, 2021 4 次提交

[CustomOp] Split test and add inference test (#31078) · e60fd1f6

由 Chen Weihang 提交于 2月 23, 2021

* split test & add inference test

* add timeout config

* change to setup install

* change to jit compile

* add verbose for test

* fix load setup name repeat

* polish details

* resolve conflict

* fix code format error

e60fd1f6

Optimization of Transformer API (#30957) · edacb629

由 xiemoyuan 提交于 2月 23, 2021

* Support 'bool' and 'int' for attention mask.

* Update docs.

* Add unittest for Transformer.

* fix bugs.

edacb629

Save load/save pickle protocol (#31044) · ee1801c1

由 WeiXin 提交于 2月 23, 2021

* add default argument  for paddle.save/static.save

* edit documentation of

* Add comments for special processing for protocol=2 and protocol=3.

* Update python/paddle/fluid/io.py
Co-authored-by: Nlanxianghit <47554610+lanxianghit@users.noreply.github.com>
Co-authored-by: Nlanxianghit <47554610+lanxianghit@users.noreply.github.com>

ee1801c1

Z

fix UNIX cmake problem (#31113) · 44ee251f
由 Zhou Wei 提交于 2月 23, 2021

44ee251f

22 2月, 2021 3 次提交

[Dy2stat] Refactoring tensor_shape_transformer.py to Fix Change after Assign Bug (#31082) · cf43a321

由 Huihuang Zheng 提交于 2月 22, 2021

**Problem**
In our old shape transformer logic, if user write:
```
s = tensor.shape
...
y = paddle.some_api(s)
```
Dy2stat will change it to
```
...
y = paddle.some_api(convert_var_shape(tensor))
```
However it will cause fatal bug if user changes the shape of `x` after assign. For example:
```
s = tensor.shape
...
tensor = paddle.some_change_shape_api(tensor)
...
y = paddle.some_api(s)
```
Then the Dy2stat will get wrong result because the code is translated into:
```
tensor = paddle.some_change_shape_api(tensor)
...
y = paddle.some_api(convert_var_shape(tensor)) # tensor shape has been changed, not origin `s` value
```

**Solution Logic**

It can not be solved in the old logic, so I refactoring tensor_shape_transformer logic. Now we will use `s` to store shape attribute and generate a var `s__STATIC_CONVERT_VAR_SHAPE_SUFFIX` to store static shape API `shape(tensor)`
```
s = tensor.shape
...
y = paddle.some_api(s)
```
Dy2stat will change it to
```
s = tensor.shape
s__STATIC_CONVERT_VAR_SHAPE_SUFFIX = shape(tensor)
...
y = paddle.some_api(choose_shape_attr_or_api(s, s__STATIC_CONVERT_VAR_SHAPE_SUFFIX ))
```
In this case, the code is consistent with origin dygraph meaning and it fixed the change after assign bug.

**Code Key Note**

To help reviewers, the key change of this PR is changing `self.name_to_var_shape` from "mapping name to shape node" to "mapping name to its STATIC_CONVERT_VAR_SHAPE_SUFFIX name", then if a variable name has the SUFFIX, we can choose to use attribute shape or shape api. Other changes go with the key change.

**Consideration**
The issue of this PR is that we store extra static `shape` API result, will it harms the speed of Dy2stat? In some cases it will, but we argue that the benefit would be greater than the cost.

1. The extra calling to static `shape` API will happen when coder assign among shape variables. Take the following dygraph code as an instance:
```
s1 = tensor.shape
s2 = s1
s3 = s2
...
```
Then we called extra static `shape` APIs again and again, however users seldom write code like this.

2. If the shape variable is used a lot, for example:
```
s = tensor.shape
y1 = paddle.some_api1(s)
y2 = paddle.some_api2(s)
y3 = paddle.some_api3(s)
```
Our old logic will create 3 shape APIs but now just 1. This is more common user code pattern. In fact, if reviewers take a look at the current unit test in this PR, you could see the op numbers decrease after this PR. So we argue that this PR can also improve speed in this code pattern.

cf43a321

fix dist fleet ctr ut (#31087) · 0e4b1542

由 tangwei12 提交于 2月 22, 2021

* fix dist fleet ctr ut

Change-Id: I59bf5123c7bd47bd0e8f1ca2a26295257597c0f5

* fix dist fleet ctr ut

Change-Id: Iafcdd172364be47fe67b753774ce09af050bcbce

* Update CMakeLists.txt

0e4b1542

[2.0Custom OP]Support New Custom OP on Windows (#31063) · adaec007

由 Zhou Wei 提交于 2月 22, 2021

* [2.0.1]Support New Custom OP on windows

* fix CI

* fix code style

* fix CI

* fix CI

* fix coverage

* fix CI

* fix CI

adaec007

20 2月, 2021 2 次提交

[CustomOp] Add more dispatch marco for users (#31058) · 6beeafe7

由 Chen Weihang 提交于 2月 20, 2021

* add more dispatch marco

* add more dispatch marco

* add more tests

* revert unneeded change

* add timeout for test dispatch

* add float and complex test

* remove and marco

6beeafe7

add squeeze_op/unsqueeze_op on kunlun;fix conv op and parallel... · d5323dab

由 TTerror 提交于 2月 20, 2021

add squeeze_op/unsqueeze_op on kunlun;fix conv op and parallel executor;optimize lookup_table op (#31056)

* add squeeze_op/unsqueeze_op on kunlun; fix conv op and parallel executor on kunlun; optimize lookup_table op on kunlun

* update squeeze/unsqueeze op

d5323dab

BaiXuePrincess / Paddle 与 Fork 源项目一致

BaiXuePrincess / Paddle
与 Fork 源项目一致