提交 · 63b6a11b7564caf8bef1883cede1677c84b468c1 · BaiXuePrincess / Paddle

07 9月, 2022 11 次提交
- Z
  Clear extra attrs of reduce op in OpMaker (#45786) · 63b6a11b
  由 zyfncg 提交于 9月 07, 2022
```
* clear extra attrs of reduce op in opmaker

* fix reduce_mean
```
  63b6a11b
- H
  
  [XPU] move rnn op to phi. (#45822) · 91631492
  由 houj04 提交于 9月 07, 2022
  
  91631492
- W
  Layernorm shift partition (#45736) · 960109af
  由 wenbin 提交于 9月 07, 2022
```
* first commit

* conver done

* correct format

* layernorm_shift_partition

* correct convert

* redefine plugin

* runable

* bug fix

* modify ShiftPartitionPattern

* correct

* add UT

* modify ut

* compile

* modify enforce

* modify UT
```
  960109af
- C
  [Auto Parallel] Support Iterable dataset for auto parallel (#45518) · b77fa1d9
  由 caozhou 提交于 9月 07, 2022
```
* support iterable dataset for auto parallel

* add split_data proto

* fix unittest bug

* fix recompute bug

* update cmake
```
  b77fa1d9
- Q
  [MLU] fix sync_bn of mlu and add unittests (#45707) · 500f070d
  由 qipengh 提交于 9月 07, 2022
```
* [MLU] fix sync_bn of mlu and add unittests

* [MLU] remove redunant code of pytest
```
  500f070d
- L
  
  add device context getter (#45790) · b7d219be
  由 LiYuRio 提交于 9月 07, 2022
  
  b7d219be
- L
  Performance fix for broadcast kernel [Part2] (#40051) · 87cba48b
  由 limingshu 提交于 9月 07, 2022
```
* first commit

* merged with develop

* merged with develop

* fix merge sequential one dims bugs
```
  87cba48b
- S
  [PHI] Migrate scale kernel (#45537) · 429b5b5b
  由 Sławomir Siwek 提交于 9月 07, 2022
```
* scale kernel

* endline

* add inplace

* fix merge conflicts

* Merge conflicts
```
  429b5b5b
- X
  [InferMeta] add compile-time infermeta logic for stack infermeta. (#45528) · 5a4ceb32
  由 xiongkun 提交于 9月 07, 2022
```
* add compile-time infermeta logic for stack infermeta.

* add unittest for stack infermeta where -1 exists in shapes.

* remove backward changes.
```
  5a4ceb32
- Z
  
  [Sparse]Rename sparse kernel (#45730) · 36739748
  由 zhangkaihuo 提交于 9月 07, 2022
  
  36739748
- S
  Fix UpdateLossScalingKernel to prevent data transform error (#45809) · c084a7b1
  由 sneaxiy 提交于 9月 07, 2022
```
* fix amp kernel

* update to remove PADDLE_WITH_XPU macro
```
  c084a7b1
06 9月, 2022 26 次提交
- Y
  [PHI]Add TensorArray for PHI (#45479) · 68f99b78
  由 YuanRisheng 提交于 9月 06, 2022
```
* add tensor array

* fix ci bugs

* fix ci bugs

* fix ci bugs

* fix ci bugs

* update by comment

* update code
```
  68f99b78
- D
  
  fix cmake download program (#45800) · 3f3f923b
  由 danleifeng 提交于 9月 06, 2022
  
  3f3f923b
- H
  [jit] pe engine with mkldnn (#45728) · 0a7e6f90
  由 Hui Zhang 提交于 9月 06, 2022
```
* using mkldnn

* using with mkldnn macro

* fix use mkldnn
```
  0a7e6f90
- W
  
  enable memory optimize when fp16. (#45792) · 1967c6a6
  由 Wilber 提交于 9月 06, 2022
  
  1967c6a6
- J
  Added concat workaround for vivo model (#45091) · 8f37c66f
  由 jakpiase 提交于 9月 06, 2022
```
* concat workaround

* CI rerun
```
  8f37c66f
- Y
  
  migrate deformable_conv and merged momentum kernels to phi, test=kunlun (#45691) · 7f3c7aeb
  由 ykkk2333 提交于 9月 06, 2022
  
  7f3c7aeb
- R
  Enable startup program for standalone executor (#45314) · 6df93364
  由 Ruibiao Chen 提交于 9月 06, 2022
```
* Enable startup program for standalone executor

* Disable test_py_reader_using_executor

* Fix test_parallel_executor_mnist

* Fix CI errors

* Fix CI errors
```
  6df93364
- C
  Update protobuf output format for profiler (#45724) · 23bc0e3c
  由 chenjian 提交于 9月 06, 2022
```
* update protobuf format

* fix protobuf content

* fix file mode

* fix compiling error when gpu not exists

* fix compiling error when gpu not exists

* fix compiling error when gpu not exists

* fix compiling error when gpu not exists

* support rocm
```
  23bc0e3c
- Z
  [Paddle Inference] fix bugs in quant_conv2d_dequant_fuse_pass when weight is... · ddc244d3
  由 zhoutianzi666 提交于 9月 06, 2022
```
[Paddle Inference] fix bugs in quant_conv2d_dequant_fuse_pass when weight is shared  between ops (#45719)

* fix_old_format

* fix bug in quant_conv2d_dequant

* fix bug in quant_conv2d_dequant
```
  ddc244d3
- Y
  
  migrate unsqueeze kernels to phi, test=kunlun (#45673) · 4acf1ef7
  由 ykkk2333 提交于 9月 06, 2022
  
  4acf1ef7
- O
  
  take some notes about sparse API (#45720) · 5c95e5c8
  由 OccupyMars2025 提交于 9月 06, 2022
  
  5c95e5c8
- Y
  
  fix mkldnn bugs (#45770) · 23def396
  由 YuanRisheng 提交于 9月 06, 2022
  
  23def396
- W
  [Eager, Performance optimization] reduce_all interface move reduce_all flag... · 192b3033
  由 Weilong Wu 提交于 9月 06, 2022
```
[Eager, Performance optimization] reduce_all interface move reduce_all flag from python to C++ (#45744)

* [Eager, Performance optimization] move reduce_all flag from python to c++

* polish reduce_all

* fix ci error

* fix errors
```
  192b3033
- N
  
  Fix layout autotune in windows ci (#45751) · cd84e1bf
  由 niuliling123 提交于 9月 06, 2022
  
  cd84e1bf
- W
  
  Fix DequantizeTwoScale kernel (#45632) · 98a5af1a
  由 whs 提交于 9月 06, 2022
  
  98a5af1a
- W
  [Eager, Performance optimization] Reduce min/max kernel polish (#45755) · a6476418
  由 Weilong Wu 提交于 9月 06, 2022
```
* [Eager, Performance optimization] reduce_max / min polish

* polish reduce_max / min

* update min/max kernel reduce_all logic

* fix a mistake

* fix ci errors

* fix errors
```
  a6476418
- X
  
  elementwise op support fp16 (#45496) · f6d9ec27
  由 xiaohemaikoo 提交于 9月 06, 2022
  
  f6d9ec27
- Z
  Clear extra attributes of matmul_v2 in OpMaker (#45708) · d4c4c53d
  由 zyfncg 提交于 9月 06, 2022
```
* set use_cudnn=true for conv2d

* clear opmaker of matmul_v2

* fix bug of set_attr

* add extra attr checker in infer_shape
```
  d4c4c53d
- Z
  
  clear extra attrs of some op in opmaker (#45758) · 22f042ba
  由 zyfncg 提交于 9月 06, 2022
  
  22f042ba
- L
  [TRT] Add silu converter (#45588) · dd0f9b96
  由 LielinJiang 提交于 9月 06, 2022
```
* add silu converter
```
  dd0f9b96
- L
  Fix grad error of groupnorm op when cuda version==11.7 (#45738) · b0a3638f
  由 LielinJiang 提交于 9月 06, 2022
```
* fix grad error of grounorm op when cuda version==11.7
```
  b0a3638f
- W
  [Paddle-Inference] remove int8 fallback (#45762) · 31efe00a
  由 Wangzheee 提交于 9月 06, 2022
```
* remove int8 fallback
```
  31efe00a
- C
  
  polish xpu enforce msg, test=kunlun (#45749) · b1f1dd05
  由 Chen Weihang 提交于 9月 06, 2022
  
  b1f1dd05
- C
  
  add op count by lib method (#45680) · 8d4f2613
  由 Chen Weihang 提交于 9月 06, 2022
  
  8d4f2613
- W
  
  Completes basic dtypes for collective api in eager mode (#45574) · 7a92e74b
  由 Wen Sun 提交于 9月 06, 2022
  
  7a92e74b
- H
  
  [XPU] rmsprop to phi. (#45734) · 1137677a
  由 houj04 提交于 9月 06, 2022
  
  1137677a
05 9月, 2022 3 次提交

[PHI] Move oneDNN helper classes to new location (#45626) · 269bd1fe

由 piotrekobi 提交于 9月 05, 2022

* gaussian random

* mkldnn to onednn renaming

* fix merge conflicts

* remove fluid code

* onednn renaming

* Move classes from mkldnn_reuse.h to onednn_reuse.h

* Move more functions from mkldnn_helper.h to onednn_helpper.h

* Change MKLDNN to OneDNN in VLOG message
Co-authored-by: NSilv3S <slawomir.siwek@intel.com>

269bd1fe

New format quant model support for MKLDNN (#45416) · 4e4f4586

由 yeliang2258 提交于 9月 05, 2022

* support onnx format quantized model

* update code

* add test

* add test

* fix

* fix test

* fix cmake

* update code

* change scale file path to calibration file path

* update code

* update code

* fix build bug

* fix build bugs

* fix

* fix

4e4f4586

K
[Bug Fix] fix compile error in gcc540 (#45702) · fd56f08e
由 kangguangli 提交于 9月 05, 2022
```
* fix compile error in gcc540
```
fd56f08e

BaiXuePrincess / Paddle 与 Fork 源项目一致

BaiXuePrincess / Paddle
与 Fork 源项目一致