提交 · c5ae43a2676f275b32543a281a555004182a98d6 · PaddlePaddle / Paddle

24 2月, 2022 25 次提交

R

fix paddle.where torch diff (#39859) · c5ae43a2
由 ronnywang 提交于 2月 24, 2022

c5ae43a2
A
[phi]migrate increment addmm multinomial cholesky kernels to phi (#39858) · b695fd95
由 Aganlengzi 提交于 2月 24, 2022
```
* migrate increment addmm multinomial cholesky kernels to phi

* test pr39869

* test pr39869

* fix style and ci
```
b695fd95
L
[phi] move randint to phi (#39872) · 127440c3
由 Leo Chen 提交于 2月 24, 2022
```
* move randint to phi

* use host generator
```
127440c3

build a Paddle Graph from CINN compiled program for execution with PE (#39724) · 4d042a83

由 TeFeng Chen 提交于 2月 24, 2022

* build a Paddle Graph from CINN compiled program for execution with PE

* update names of some variables

* fix random fail in build_cinn_pass_test and update some comments

* fix compiler error by merging phi pr

4d042a83

Optimize nearest_interp backward (#39067) · df0b4434

由 Lijunhui 提交于 2月 24, 2022

* nearest_interp_bw init

* optimize kernel config

* optimize kernel config

* fix struct init

* optimize code

* rm duplicated struct

df0b4434

C
Fix unittests for eigh op (#39568) · 539fb0d7
由 crystal 提交于 2月 24, 2022
```
* fix eigh test

* modify atol and rtol
```
539fb0d7

[Phi]Move cross OP to phi (#39829) · 6c358a7c

由 0x45f 提交于 2月 24, 2022

* move cross forward OP

* move cross grad op to phi

* move infershape

* refine infershape

* rename ctx

* set dtype and layout in InferMeta

* refine code

6c358a7c

L
[phi] move bce_loss to phi (#39868) · 6fc5d88a
由 Linjie Chen 提交于 2月 24, 2022
```
* move bce_loss to phi

* refine PADDLE_ENFORCE

* revert PADDLE_ENFORCE

* fix ci
```
6fc5d88a

[doc]Fix maxunpool2d example (#39862) · eb4ad509

由 xiaoting 提交于 2月 24, 2022

* fix maxunpool2d example, test=document_fix

* fix maxunpool2d example, test=document_fix

eb4ad509

【Phi】Migrate poisson op into phi (#39814) · bbe441fc
由 zhouweiwei2014 提交于 2月 24, 2022
```
* Migrate poisson op into phi

* fix CI

* fix comment
```
bbe441fc
Z

config fleet optimize. test=develop (#39849) · 23bbd912
由 zmxdream 提交于 2月 24, 2022

23bbd912

Added nearest interp v2 BF16 FWD kernel (#39490) · 2ec943a7

由 jakpiase 提交于 2月 24, 2022

* added nearest interp v2 bf16

* disabled bilinear interp nhwc test

* added skipping UT for gpu

* added NHWC support

* removed unnecessary statements

* minor change

* CI fix

* added appropriate changes to interpolate_v1

* fix after review

* minor change

* minor change

* revert unwanted deletions

* CI fix

2ec943a7

Refactored GradNodeAccumulation data structure and behaviour (#39526) · 1abfc8dd

由 Zhanlue Yang 提交于 2月 24, 2022

* Refactored GradNodeAccumulation data structure and behaviour

* Fixed CI issues

* Fix compilation issues

* Fixed minor issues

* Reverted changes for intermediate and OverwriteOutput

* fixed minor issue

* Fixed code format issues

* Fixed CI-Coverage issue

* Fixed CI issues

1abfc8dd

L
fix 'invalid escape sequence' (#39842) · 4e26fa57
由 Leo Chen 提交于 2月 24, 2022
```
* fix 'invalid escape sequence'

* fix assert error
```
4e26fa57
A
[Phi] Fix XPU OP segmentation Fault problem (#39827) · 7a7a7cad
由 Aurelius84 提交于 2月 24, 2022
```
* [Phi] Fix XPU OP segmentation Fault problem

* fix cast_op_xpu in kunlun1

* fix cast_op_xpu in kunlun1
```
7a7a7cad

[pten] add optional type for infermeta (#39848) · 94b31f90

由 chentianyu03 提交于 2月 24, 2022

* modify infershape by args_def

* add optional type for infermate

* add optional type for infermate

* add optional type for infermate

* support scalar type

* change OptionalInputAt function to none template

* support phi::DataType

94b31f90

X

[Phi] Fix comilation dependecy in selected_rows with memory (#39834) · dd2c997d
由 xiongkun 提交于 2月 24, 2022

dd2c997d
J
Fix for split op in BF16 inference (#39548) · 75f91ce4
由 jakpiase 提交于 2月 24, 2022
```
* Fix for split bf16 inference

* added test for pass

* changes after review
```
75f91ce4
H
Optimize where_op and abs_grad_op by the elementwise interface (#39609) · c9699556
由 huangxu96 提交于 2月 24, 2022
```
* Optimize the where_op by the elementwise_op funtion

* Modified where_op & abs_grad_op by elementwise interface
```
c9699556
H
Add Note for Place of Executor in Parallel Environment (#39063) · 867224b2
由 Huihuang Zheng 提交于 2月 24, 2022
```
Add note for Place of Executor in parallel environment
```
867224b2
J

fix bug for block state (#39854) · 5fd7b5c3
由 JZ-LIANG 提交于 2月 24, 2022

5fd7b5c3

[Eager] save load testcase (#39571) · 6b5749eb

由 wanghuancoder 提交于 2月 24, 2022

* eager, test=develop

* fix bug, test=develop

* eager, test=develop

* merge legacy to fluid

* eager, test=develop

* eager, test=develop

* Refactor TensorAdd func by template and remove gradient_accumulation in eager

* Remove needless target name

* eager, test=develop

* eager, test=develop

* Use overload instead of template

* Remove legacy code

* Remove legacy code

* selectedrows, test=develop

* Remove DataType test

* eager, test=develop

* eager, test=develop

* support gan, test=develop

* Using Tensor directly instead of using EagerTensor

* support gradient_accumulation

* make test_imperative_lod_tensor_to_selected_rows longer

* make test_imperative_lod_tensor_to_selected_rows longer

* refine code

* ptb, test=develop

* Rename all EagerTensor to Tensor

* Rename some EagerTensor to Tensor

* rename EagerTensor to EagerVariable

* eager, test=develop

* eager, test=develop

* eager, test=develop

* eager, test=develop

* add more test

* eager, test=develop

* Support copiable selected rows and merge develop

* save load, eager, test=develop

* save load, eager, test=develop

* refine, test=develop

* refine, test=develop

* refine, test=develop

* revert static_runner, test=develop

* EagerTensor to Tensor, test=develop

* refine, test=develop

* refine, test=develop

* clear grad, test=develop

* merge, develop

* merge, develop

* merge, test=develop

* merge, test=develop
Co-authored-by: NJiabinYang <360788950@qq.com>
Co-authored-by: NWeilong Wu <veyron_wu@163.com>

6b5749eb

N

Fix a bug in IndexKernel out-of-memory (#39867) · 2136bd42
由 niuliling123 提交于 2月 24, 2022

2136bd42
L
optimize performance of lookup_table_v2_op (#39856) · d6038c22
由 Li Min 提交于 2月 24, 2022
```
* optimize block config  and fp16 atomicAdd perf for lookup_table_v2_grad.
```
d6038c22
C
[PHi] Skip kernel declare for cuda only kernel on rocm (#39869) · 76a6b88d
由 Chen Weihang 提交于 2月 24, 2022
```
* skip kernel declare for cuda only kernel on rocm

* fix error
```
76a6b88d

23 2月, 2022 15 次提交
- J
  
  added paddle_bfloat to requirements (#39740) · 2457a7d1
  由 jakpiase 提交于 2月 23, 2022
  
  2457a7d1
- S
  Add ProcessGroupNCCL for distributed training (#39737) · 0b205817
  由 ShenLiang 提交于 2月 23, 2022
```
* add processgroup_nccl
```
  0b205817
- 石
  infrt runtime supports phi, test=develop (#39836) · 058e1d85
  由石晓伟提交于 2月 23, 2022
```
* runtime supports pten kernels, test=develop

* fixes a bug, test=develop
```
  058e1d85
- Z
  
  Support dispensable inputs for eager final state codegen (#39743) · ca11a0e5
  由 Zhanlue Yang 提交于 2月 23, 2022
  
  ca11a0e5
- C
  
  move array_ref_test and small_vector_test into paddle/utils and format header macro define (#39831) · 96d530c1
  由 chentianyu03 提交于 2月 23, 2022
  
  96d530c1
- S
  move trunc_op's infere shape to phi (#39772) · 95280a36
  由 Sing_chan 提交于 2月 23, 2022
```
* move trunc_op's infere shape

* modify according to risheng's comment
```
  95280a36
- L
  [phi] move randperm to phi (#39816) · 30992ea0
  由 Leo Chen 提交于 2月 23, 2022
```
* move randperm to phi

* fix npu

* fix memory::Copy
```
  30992ea0
- Y
  
  [Phi] move flip op to phi kernel (#39822) · ad294a81
  由 Yang 提交于 2月 23, 2022
  
  ad294a81
- C
  [Phi] Polish default signature attr and output select impl (#39810) · 64ed92bd
  由 Chen Weihang 提交于 2月 23, 2022
```
* polish default sig impl

* revert dispenable out
```
  64ed92bd
- [MLU] add cncl parallel context and mlu resource pool (#39803) · 6241913b
  由 mhhhh1 提交于 2月 23, 2022
```
* [MLU] add cncl parallel context and mlu resource pool

* [MLU] fix the cncl_context_test
```
  6241913b
- change CUDA implementaion of bernoulli OP (#39732) · b9675acc
  由 zhouweiwei2014 提交于 2月 23, 2022
```
* change CUDA implementaion of bernoulli OP

* fix CI
```
  b9675acc
- Z
  refactor range unittest for kunlun (#39800) · 69a04209
  由 zhangxiaoci 提交于 2月 23, 2022
```
*test=kunlun
```
  69a04209
- R
  
  [phi] migrate atan2_op into phi (#39806) · b089e7cd
  由 ronnywang 提交于 2月 23, 2022
  
  b089e7cd
- L
  [phi] move unbind to phi (#39789) · dba694f4
  由 Leo Chen 提交于 2月 23, 2022
```
* move unbind to phi

* revert infer shape

* add header file

* move concat_and_split to phi
```
  dba694f4
- L
  [KP] Add elementwise add xpu after phi, test=develop (#39787) · 1a1a2ce8
  由 Liu-xiandong 提交于 2月 23, 2022
```
* [KP] Add elementwise add xpu, test=develop

* modify the File Permissions

* modify the copyright time

* modify code style

* modify code style
```
  1a1a2ce8

PaddlePaddle / Paddle 接近 2 年 前同步成功

PaddlePaddle / Paddle
接近 2 年前同步成功