提交 · 9075a0fd2e792edf5c3841c0053e980a922b04de · 机器未来 / Paddle

17 12月, 2021 2 次提交
- J
  ipu add dockerfile (#37792) · 9075a0fd
  由 jianghaicheng 提交于 12月 17, 2021
```
* ipu add dockerfile

* resolve comments
```
  9075a0fd
- W
  
  fix bind failed with Address already in use (#38174) · 446a62e8
  由 WangXi 提交于 12月 17, 2021
  
  446a62e8
16 12月, 2021 30 次提交

S

modify according to zhouwei's comment (#38166) · a37be82f
由 Sing_chan 提交于 12月 16, 2021

a37be82f
L
[new-exec] skip add_dependency when cc_test skipped because of CI_SKIP_CPP_TEST=ON (#38191) · 30973183
由 Leo Chen 提交于 12月 16, 2021
```
* fix cmake

* not check execution time
```
30973183
S

block warning: overriding D9025 (#38034) · 672dba1b
由 Sing_chan 提交于 12月 16, 2021

672dba1b
C

add grad maker debug log (#38183) · a43d8e59
由 chentianyu03 提交于 12月 16, 2021

a43d8e59

Faster implementation of CPU kernel for ROI Align operator (#37848) · 023ff4f5

由 Tomasz Socha 提交于 12月 16, 2021

* Faster implementation of CPU kernel for ROI_ALIGN Operator

* Add missing variable to CUDA roi_align_op

* Style

* Fix boundaries

* Rename variables for indexes calculation

* Remove unnecessary emplace

* Revert "Remove unnecessary emplace"

This reverts commit c10e87f7fb812f1a672fde32f2690a97d47e2f5a.

* Style

023ff4f5

C

pylayer support HIP (#38184) · 2e76d5ad
由 chentianyu03 提交于 12月 16, 2021

2e76d5ad

Fixed LD_LIBRARY_PATH for eager_code_generator (#38160) · af30f545

由 Zhanlue Yang 提交于 12月 16, 2021

* Rearranged Eager AutoCodeGen directory structure

* Removed USE_OP in Eager AutoCodeGen

* Enabled generation for Operators without Grad/Inputs/Outputs

* Resolved operators without input

* Fixed merge conflicts

* Enabled Eager AutoCodeGen for 10+ more operators

* Refactored Eager AutoCodeGen with more organized helper objects

* Enabled Eager AutoCodeGen for operators with multiple OpBases

* Adjusted Eager AutoCodeGen to Enable Passing Output Tensor as Input Argument

* Handled Dispensable Inputs/Outputs in Eager AutoCodeGen

* Enabled Eager AutoCodeGen for All Existing Operators & Possible Future Operators

* Fixed CI issues

* Fixed LD_LIBRARY_PATH for eager_code_generator

af30f545

Y

remove nightly ut from parallel ut list (#38163) · 3d7de712
由 YUNSHEN XIE 提交于 12月 16, 2021

3d7de712

Add arc hyperbolic function op (#37076) · 36b7368d

由 xiaoting 提交于 12月 16, 2021

* add activation

* update activation_op

* add unitest for activation

* fix acosh for init, test=develop

36b7368d

Conv transpose eltwiseadd bn fuse pass (#37800) · e64f0997

由 feng_shuai 提交于 12月 16, 2021

* conv_transpose_eltwiseadd_bn_fuse_pass

* change timeout

* add TIMEOUT

* add random num for group and dilation

* change PassCompat

e64f0997

王

Revert "modify the fix_seed attribute in dropout op is a def... · 464f2af8

由王明冬提交于 12月 16, 2021

Revert "modify the fix_seed attribute in dropout op is a def attribute.test=develop (#38100)" (#38127)

This reverts commit f44add7b.

464f2af8

Add tests for PaddleInference Pass (#37676) · 96597a85

由 yeliang2258 提交于 12月 16, 2021

* add test for conv_elementwise_add2_act_fuse_pass and conv_elementwise_add_act_fuse_pass

* Add conv_eltwiseadd_bn_fuse_pass test and fix test_conv_elementwise_addX_act_fuse_pass

* add tests for conv_act_mkldnn_fuse_pass

* add test for conv_bias_mkldnn_fuse_pass

* update code

* add conv_act_mkldnn_fuse_pass for relu, relu6, swish, leaky_relu

* update test

* update

* update bug

* update

* update pattern_detector

* fix test_conv_eltwiseadd_bn_fuse_pass

* add diff display notest;test=windows_ci_inference

* fix

* remove test_conv_act_mkldnn_fuse_pass.py

* ifix

96597a85

[PTen] Unify device context entrance in pten part 2 (#38182) · e02537f9

由 Chen Weihang 提交于 12月 16, 2021

* unify device context entrance

* move all_context include to header

* polish cmake relay for device_context

* fix npu compile failed

* fix npu compile failed

e02537f9

Y

add defaults value for disable_ut (#38110) · 55509ae7
由 YUNSHEN XIE 提交于 12月 16, 2021

55509ae7
Z

change tar mode (#38119) · 10ebe231
由 zhangchunle 提交于 12月 16, 2021

10ebe231

[PTen] Add register_ctx_kernel marco and move scale kernel (#38121) · af498677

由 Chen Weihang 提交于 12月 16, 2021

* add register_ctx_kernel and move scale kernel

* polish details by reviewer comment

* fix xpu compile failed

* fix cmake error

af498677

W

Arg weight of lerp support float in static mode, test=develop (#38080) · 58b4bc72
由 wuhuanzhou 提交于 12月 16, 2021

58b4bc72
J
support eager switch system (#38170) · 8305c2be
由 Jiabin Yang 提交于 12月 16, 2021
```
* support eager switch system

* polish code
```
8305c2be
D
[psgpu]add checknan print and fix trainer device (#38131) · 092839d6
由 danleifeng 提交于 12月 16, 2021
```
* trainer_device fix and checknan tool for psgpu;test=develop

* disable show_one_table;test=develop
```
092839d6
T

check build trt8 docker images (#37920) · 25c1b623
由 tianshuo78520a 提交于 12月 16, 2021

25c1b623

Adapt host event recorder to profiler (#37766) · 5b6be4d7

由 liutiexing 提交于 12月 16, 2021

* add align for WorkQueue

* add spinlock

* merge develop

* merge

* Add EventsWaiter

* Revert "Add EventsWaiter"

This reverts commit e206173aa9be7401b83a53581627bfaf557c8fb2.

* add os_info

* update

* update

* update

* update

* update

* update for bugfix

* update

* update

* update
Co-authored-by: Nliutiexing <liutiexing@google.com>

5b6be4d7

L
Add fmax and fmin operators (#37826) · dd3afc9d
由 LJQ❤️ 提交于 12月 16, 2021
```
Add elementwise_fmax and elementwise_fmin operators
```
dd3afc9d

Add sparse_attention mask ,test=develop (#37973) · fa463b90

由 Liu-xiandong 提交于 12月 16, 2021

Add key_padding_mask and attn_mask in sparse_attention Api

1.Key padding mask is a tensor with dimensions [batch_size, seq_len], and attention mask is a tensor with dimensions [seq_len, seq_len]. The data types of the two masks are consistent with Q, K, and V, which are float32 or float64. If the value in Mask is 0, it means that the position needs to be masked.

2.The changed files are mainly paddle/fluid/operators/sparse_attention_op.cu and python/paddle/fluid/tests/unittests/test_sparse_attention_op.py. sparse_attention has three parts: sddmm, softmax, and dsd. Adding the mask operation only needs to modify the softmax. It has no effect on the other two parts. In addition, in order to test the mask function, related tests has been added.

fa463b90

N
Add the transformop parameter in TensorReduceFunctorImpl (#38135) · 524389ee
由 niuliling123 提交于 12月 16, 2021
```
* Add the transformop parameter in TensorReduceFunctorImpl
```
524389ee

[Pten]Modify registered kernel name (#38109) · be874c08

由 YuanRisheng 提交于 12月 16, 2021

* Reduce reshape kernel functions in pten

* delete notes

* fix bugs when compile

* modify register name

* fix compile bugs

be874c08

[PTen] Unify device context entrance in pten part 1 (#38172) · 047ee26c

由 Chen Weihang 提交于 12月 15, 2021

* unify device context entrance

* move all_context include to header

* polish cmake relay for device_context

* fix npu compile failed

* fix npu compile failed

* revert part of change

047ee26c

C

fix header match error (#38175) · 4ef59f08
由 Chen Weihang 提交于 12月 15, 2021

4ef59f08

pylayer support tuple/list type args and fix check args bug (#38146) · 861053eb

由 chentianyu03 提交于 12月 16, 2021

* Revert "Revert "pylayer support tuple/list type args (#37727)" (#37956)"

This reverts commit d848ff04.

* move check args,kwargs before forward execute

861053eb

Add float16 type for scatter op. (#38136) · 9bac4a76

由 Li Min 提交于 12月 16, 2021

* Add float16 type for scatter op.

* Add fp16 test for scatter op.

* Add int and int64 support for scatter_grad on gpu.

* Add int and int64 for check_variable_and_dtype routine.

* Minors.

* Code format.

9bac4a76

Enabled Eager AutoCodeGen for All Existing Operators & Possible Future Operators (#37969) · 08482a86

由 Zhanlue Yang 提交于 12月 16, 2021

* Rearranged Eager AutoCodeGen directory structure

* Removed USE_OP in Eager AutoCodeGen

* Enabled generation for Operators without Grad/Inputs/Outputs

* Resolved operators without input

* Fixed merge conflicts

* Enabled Eager AutoCodeGen for 10+ more operators

* Refactored Eager AutoCodeGen with more organized helper objects

* Enabled Eager AutoCodeGen for operators with multiple OpBases

* Adjusted Eager AutoCodeGen to Enable Passing Output Tensor as Input Argument

* Handled Dispensable Inputs/Outputs in Eager AutoCodeGen

* Enabled Eager AutoCodeGen for All Existing Operators & Possible Future Operators

* Fixed CI issues

08482a86

15 12月, 2021 8 次提交

B
update mkldnn scale_matmul fuse pass ut (#37210) · bce1e572
由 baoachun 提交于 12月 15, 2021
```
* update mkldnn scale_matmul fuse pass ut

* update mkldnn scale_matmul_fuse_pass ut
```
bce1e572

add mkldnn conv3d_bias_mkldnn_fuse_pass ut (#37700) · 0456e003

由 baoachun 提交于 12月 15, 2021

* add mkldnn conv3d_bias_mkldnn_fuse_pass ut

* update conv3d_bias_mkldnn_fuse_pass ut

* disable conv3d_bias_mkldnn_fuse_pass

0456e003

Y
Change a comment in pten header to avoid the disturb to op benchmark ci. (#38165) · cecea8e6
由 Yiqun Liu 提交于 12月 15, 2021
```
test=document_fix
```
cecea8e6
Y
Fix bugs in Translated Layer when change mode from train/eval to eval/train (#38141) · f616d9a7
由 YuanRisheng 提交于 12月 15, 2021
```
* fix bugs in Translated layer when change train/eval

* fix python converage
```
f616d9a7

[Dy2stat]Fix error in tensor_shape_transformer. (#37999) · 50822491

由 0x45f 提交于 12月 15, 2021

* fix error when tensor_shape_transformer. Before in stmt like `if len(paddle.shape(x)[0]) > 0`, `paddle` will be used as a variable

* handle other call like `fluid.layers.mean` and `fluid.layers.shape`

* add unit test

50822491

ipu_inference (#37102) · 141b2854

由 jianghaicheng 提交于 12月 15, 2021

* add ipu_inference

* resovle commments

* resolve comments

* add EnableIpu introduction

* rm line

* restore npu update

* add ernie and resnet50 test

* fix copyright time
Co-authored-by: Nyaozhixin <522190855@qq.com>

141b2854

[new-exec] add standalone executor test (#38101) · ab6daf84

由 Leo Chen 提交于 12月 15, 2021

* refine test

* add download_program target

* update ut code

* refine code

* disable profiler

* add comments

* refine cmake

* skip coverage ci

ab6daf84

Z

add use_warning of amp (#38086) · 3a2093a5
由 zhangbo9674 提交于 12月 15, 2021

3a2093a5

机器未来 / Paddle 与 Fork 源项目一致

机器未来 / Paddle
与 Fork 源项目一致