提交 · 42c7bb4794832e5eb5f6f2e789cadc0dedac2345 · OPTHREE / Paddle

15 3月, 2022 4 次提交
- Q
  
  [MLU] add check_finite_and_unscale op for amp (#40458) · 42c7bb47
  由 qipengh 提交于 3月 15, 2022
  
  42c7bb47
- Y
  
  add yaml (#40533) · 5cb506b0
  由 YuanRisheng 提交于 3月 15, 2022
  
  5cb506b0
- A
  [IPU] add IPU related CI configures (#40354) · 8852591f
  由 Allen Guo 提交于 3月 15, 2022
```
* add ci

* rm retry tests

* format

* restore retry tests

* update timeout for ipu uts
```
  8852591f
- H
  [Dygraph] Refactoring of reducer in DataParallel (#40389) · 1a32391c
  由 Haohongxiang 提交于 3月 15, 2022
```
* refactor reducer

* modify cmakelists

* solve conflicts

* rename group and update process_group

* fix bugs of ProcessGroupNCCL

* modify for CIs

* refactoring reducer
```
  1a32391c
14 3月, 2022 12 次提交

[Phi]Add diag_v2 grad kernel (#40447) · e157f2af

由 Siming Dai 提交于 3月 14, 2022

* Add diag grad kernel

* fix unittest case

* add float16, remove const &

* delete diag_grad in op_utils.h

e157f2af

Add an elementwise + activation fusion pass. (#36541) · 3f219160

由 Tomasz Socha 提交于 3月 14, 2022

* Add elementwise add and activation fuse pass

* Fix copy ellision

* More flexible pattern detector

* More flexible fusion pass

* Update lists for pass

* Add support for Pow operator

* Add support for more activation types

* Style

* Rename fusion pass

* First version of tests

* Dirty version of pass

* Polished version

* Update pbtxt

* Style

* Update names

* Style

* Use PADDLE_ENFORCE_EQ

* Save error message to variable

* WO for error checks

* CR

* Static style check

* Add missing 'activation_scale' attribute

* Add relu6 and sigmoid activations

* Style

* Fix fuse list formating

* Sync filenames for fuse pass files

* Fix cmake after move

* Fix registration

* Fix pass name in tests

* Add missing activations to checker

* WIPS

* Working mul op

* Working sub

* Working Add

* Remove pten includes

* Remove some forward declarations

* Remove Includes

* Fixes

* Remove default kernels

* Add check if post_ops attributes are avaliable

* Style

* Code adjustment

* Register default kernels

* We have year 2022 not 2021...
Co-authored-by: Njakpiase <jakpia21@gmail.com>
Co-authored-by: NSylwester Fraczek <sylwester.fraczek@intel.com>

* Fast review fixes
Co-authored-by: Njakpiase <jakpia21@gmail.com>
Co-authored-by: NSylwester Fraczek <sylwester.fraczek@intel.com>

* Review Fix

* Rename one_dnn -> onednn

* Style after review

* Fast and dirty fix for quantization

* Update tests

* Style

* Fix mkldnn_quantizer config

* Add Joanna's suggestion.

* Check if operator is explicitly disables on OneDNN

* Try to use unregistered attributes

* Style

* Test new framework

* FXI

* FXII

* Update test

* Style
Co-authored-by: Njakpiase <jakpia21@gmail.com>
Co-authored-by: NSylwester Fraczek <sylwester.fraczek@intel.com>

3f219160

F

[MLU] add merged_momentum mlu kernel (#40406) · 1f7b2516
由 fwenguang 提交于 3月 14, 2022

1f7b2516

Support custom op and paddle.autograd.bacward in eager (#40423) · 227fa408

由 Jiabin Yang 提交于 3月 14, 2022

* eager, test=develop

* fix bug, test=develop

* eager, test=develop

* merge legacy to fluid

* eager, test=develop

* eager, test=develop

* Refactor TensorAdd func by template and remove gradient_accumulation in eager

* Remove needless target name

* eager, test=develop

* eager, test=develop

* Use overload instead of template

* Remove legacy code

* Remove legacy code

* selectedrows, test=develop

* Remove DataType test

* eager, test=develop

* eager, test=develop

* support gan, test=develop

* Using Tensor directly instead of using EagerTensor

* support gradient_accumulation

* make test_imperative_lod_tensor_to_selected_rows longer

* make test_imperative_lod_tensor_to_selected_rows longer

* refine code

* ptb, test=develop

* Rename all EagerTensor to Tensor

* Rename some EagerTensor to Tensor

* rename EagerTensor to EagerVariable

* eager, test=develop

* eager, test=develop

* eager, test=develop

* eager, test=develop

* add more test

* eager, test=develop

* Support copiable selected rows and merge develop

* save load, eager, test=develop

* save load, eager, test=develop

* refine, test=develop

* remove useless _set_value method

* refine, test=develop

* refine, test=develop

* revert static_runner, test=develop

* EagerTensor to Tensor, test=develop

* refine, test=develop

* refine, test=develop

* clear grad, test=develop

* merge, develop

* merge, develop

* merge, test=develop

* merge, test=develop

* Support quant and part of slice

* support legacy static save

* extend slim tests time

* remove imperative on inference

* remove imperative on inference

* merge develop

* fix typo

* fix typo

* split slice related code into 2 part for imperative and eager

* split slice from inference

* split slice from inference

* fix test_tensor_register_hook

* support custom op in eager mode

* fix inference deps error

* split eager utils from custom operator

* fix type match

* fix typo
Co-authored-by: NWang Huan <wanghuan29@baidu.com>
Co-authored-by: NWeilong Wu <veyron_wu@163.com>
Co-authored-by: Nwanghuancoder <wanghuancoder@163.com>

227fa408

W
[Eager] [Bug Fix] fix eager trace op bug (#40402) · 65adfecf
由 wanghuancoder 提交于 3月 14, 2022
```
* fix some slice bug, test=develop

* refine, test=develop
```
65adfecf
0

adjust params order for eager.Tensor._copy_to (#40449) · c6ec8b9f
由 0x45f 提交于 3月 14, 2022

c6ec8b9f

[KP] Add unittests for... · f269ca3f

由 Lijunhui 提交于 3月 14, 2022

[KP] Add unittests for brelu,ceil,celu,elu,floor,hard_shrink,hard_sigmoid,log1p,logsigmoid,relu6,silu,soft_relu,softsign,swish (#40448)

* solve unexecuted UT

* add 24 activation op UT

* append swish&thresholded_relu to kpfirst_list

* rm thresholded_relu

f269ca3f

Z
[AutoParallel] Converter (#40434) · 3881b6cb
由 zhaoyingli 提交于 3月 14, 2022
```
* [AutoParallel] Converter
Converter API
```
3881b6cb
W

[hybrid fix] fix sharding save offload (#40477) · edd97f94
由 WangXi 提交于 3月 14, 2022

edd97f94
B

fix_group_sharded_note (#40488) · e14a6ec9
由 Baibaifan 提交于 3月 14, 2022

e14a6ec9

[multiprocessing] Add paddle.incubate.multiprocessing for sharing tensors ... · e553f758

由 Zhong Hui 提交于 3月 14, 2022

[multiprocessing] Add paddle.incubate.multiprocessing for sharing tensors  between python processes. (#37302)

* Add support for paddle.multiprocessing
* move multiprocessing to incubate.

e553f758

Refine partial_program for new run_program OP (#40355) · afafb1c3

由 0x45f 提交于 3月 14, 2022

* refine partial_program

* fix code for test_mnist.py train

* support quantify UT

* make __fake_vars and _double_grads to lazy

* fix comments

afafb1c3

12 3月, 2022 1 次提交
- A
  [custom kernel] fix static object de-initialize bug (#40414) · 573ca984
  由 Aganlengzi 提交于 3月 12, 2022
```
* [custom kernel] fix static object de-initialize bug

* fix text

* fix text

* refine log info
```
  573ca984
11 3月, 2022 8 次提交
- Z
  Submanifold convolution (#40363) · 0d78e491
  由 zhangkaihuo 提交于 3月 11, 2022
```
submanifold convolution
```
  0d78e491
- C
  [Phi] Remove needless deps in unittests (#40256) · 89ed57e2
  由 Chen Weihang 提交于 3月 11, 2022
```
* remove needless deps in unittests

* add gpu marco

* fix other unittests

* fix kernel name error

* fix test_prepare_op

* fix failed dygraph unittests

* fix gpu failed tests

* fix cinn test failed

* fix cinn test failed

* fix dropout tests
```
  89ed57e2
- Y
  
  [hybrid] Support tensor parallel and cache structure for fused attention op. (#40101) · 1882c496
  由 Yuang Liu 提交于 3月 11, 2022
  
  1882c496
- Z
  
  [MLU]add allgather_op mlu kernel (#40356) · dc773828
  由 zn 提交于 3月 11, 2022
  
  dc773828
- update square & sigmoid unittest (#40404) · 807bff4a
  由 z8hanghuan 提交于 3月 11, 2022
  
  807bff4a
- H
  
  minor fix matmul and onehot xpu. test=kunlun (#40419) · 594e412d
  由 houj04 提交于 3月 11, 2022
  
  594e412d
- G
  
  add EMD method of post_quant (#40421) · 82c30f71
  由 Guanghua Yu 提交于 3月 11, 2022
  
  82c30f71
- B
  
  fix_import_distribute_bugs (#40396) · bd2d4fd0
  由 Baibaifan 提交于 3月 11, 2022
  
  bd2d4fd0
10 3月, 2022 5 次提交

Inference add ONNXRuntime back-end (#39988) · 431afc39

由 heliqi 提交于 3月 10, 2022

* add onnxruntime predictor

* Add code comments

* support link paddle2onnx onnxruntime

* support onnxruntime with python

* support onnxruntime with python

* support onnxruntime with windows

* paddle2onnx compile with windows

* supoort windows compile

* supoort windows compile with onnxruntime

* supoort windows compile with paddle2onnx

* supoort mac compile

* compile with mac

* compile with mac

* add code comments

* fix remind word

* code optimization

* add test case

* add test case

* add inference demo_ci test case

* fix compile paddle2onnx with no python

* add inference demo_ci test case

* add inference demo_ci test case

* add inference infer_ut test case

* support c go api and test cases

* add converage test case

* add converage test case

* add capi test case

* add capi test case

431afc39

C
[Auto Parallel]Update reshard for while sub block (#40366) · 2747de2b
由 caozhou 提交于 3月 10, 2022
```
* update reshard for while sub block

* fix code format error
```
2747de2b
Z

Supported auto code gen for sparse kernels (#40276) · 2b6da4de
由 Zhanlue Yang 提交于 3月 10, 2022

2b6da4de

add tril_triu for xpu, *test=kunlun (#40246) · 1128db30

由 z8hanghuan 提交于 3月 10, 2022

* add tril_triu for xpu, *test=kunlun

* add tril_triu for xpu, *test=kunlun

* add tril_triu for xpu, *test=kunlun

* add tril_triu for xpu, *test=kunlun

* add tril_triu for xpu, *test=kunlun

1128db30

Move dropout to phi (#40148) · 99fc1b08

由 hong 提交于 3月 10, 2022

* move dropout to phi; test=develop

* fix xpu, npu compile error; test=develop

99fc1b08

09 3月, 2022 10 次提交
- 0
  
  [Dy2st]Fix Exception in utils.py function "is_paddle_module" (#40243) · 0604df9e
  由 0x45f 提交于 3月 09, 2022
  
  0604df9e
- Z
  [PHI] Fix some bug of code auto-gen in C++ API (#40262) · 63fb0347
  由 zyfncg 提交于 3月 09, 2022
```
* support code auto-gene for sparse backward api

* fix bug of intermediate api and name of return var
```
  63fb0347
- B
  
  add_sharding_api (#40129) · f40ed5f4
  由 Baibaifan 提交于 3月 09, 2022
  
  f40ed5f4
- F
  
  change timeout for pool (#40341) · 1defc8f3
  由 feng_shuai 提交于 3月 09, 2022
  
  1defc8f3
- W
  fix the full_like with fill the value of inf (#40232) · ec582895
  由 wawltor 提交于 3月 09, 2022
```
* fix the full_like with fill the value of inf

* update the test case for the fill_any_like

* updae the comments for the full_like
```
  ec582895
- 0
  adapt run_program OP for eager (#40198) · 3e9601ba
  由 0x45f 提交于 3月 09, 2022
```
* adapt run_program OP for eager

* fix program_id

* refine code

* fix test
```
  3e9601ba
- W
  
  fix the document of ones_like, zeros_like (#40233) · 7b18c55b
  由 wawltor 提交于 3月 09, 2022
  
  7b18c55b
- N
  add MobileNetV3 (#38653) · 68af310b
  由 Nyakku Shigure 提交于 3月 09, 2022
```
* add mobilenetv3
```
  68af310b
- S
  Fix time of utest in distributed (#40163) · 7ea9235c
  由 ShenLiang 提交于 3月 09, 2022
```
* fix time of utest
```
  7ea9235c
- W
  
  [hybrid] fused_feedforward op support tensor model parallel (#40160) · e0866dc6
  由 WangXi 提交于 3月 09, 2022
  
  e0866dc6

OPTHREE / Paddle 与 Fork 源项目一致

OPTHREE / Paddle
与 Fork 源项目一致