提交 · 5c85f1a7d621f6b588a87abec38649b478975187 · BaiXuePrincess / Paddle

24 10月, 2022 1 次提交

Support BF16 training for sharding (#46846) (#47246) · 5c85f1a7

由 Ghost Screaming 提交于 10月 24, 2022

* Fix bug of reduce_sum op. When input.numel() > INT32_MAX, its result
is wrong.

* support pure bfloat16

* support bf16 linear

* update PR to pass CI

* tiny fix where_grad_kernel.cu

* Support bfloat16 type for reducer and sharding.

* Fix some bug.

* Polish code.

* Polise code.

* Add bfloat16 datatype in fill_grad kernels.
Co-authored-by: Nsneaxiy <sneaxiy@126.com>
Co-authored-by: Nsneaxiy <sneaxiy@126.com>

5c85f1a7

20 10月, 2022 1 次提交
- L
  Add value check & error message for gather_tree (#47051) (#47221) · 6712e262
  由 liu zhengxi 提交于 10月 20, 2022
```
Add value check & error message for gather_tree
cherry-pick #47051
```
  6712e262
11 10月, 2022 1 次提交
- F
  
  set_value_op: add support for complex types (#46885) · b051455f
  由 Feiyu Chan 提交于 10月 11, 2022
  
  b051455f
20 9月, 2022 2 次提交

H
[PolishComments] Polish some code comments (#46032) (#46261) · 42e56f65
由 HongyuJia 提交于 9月 20, 2022
```
* polish code comments

* polish data_device_transform.cc
```
42e56f65

[Cherry-pick] Fix amp error cp (#46272) · da173c40

由 Jiabin Yang 提交于 9月 20, 2022

* [Eager] Fix ocr (#46124)

* fix linspace error in amp

* fix log

* fix amp error

* fix ocr error which caused by amp

* add more check

* rename dtype ns

* [Eager Bug fix]Fix Detection (#46147)

* fix linspace error in amp

* fix log

* fix amp error

* Revert "Simplify size op impl (#45808)"

This reverts commit c252b1de.

* fix_seg

* fix detection
Co-authored-by: NChen Weihang <sunny_cwh@163.com>
Co-authored-by: NChen Weihang <sunny_cwh@163.com>

da173c40

19 9月, 2022 2 次提交
- R
  [vision.ops.nms] Fix return order error and duplicate results with specific... · be84cac7
  由 RichardWooSJTU 提交于 9月 19, 2022
```
[vision.ops.nms] Fix return order error and duplicate results with specific inputs (#46148) (#46193)

* fix return order error and duplicate results with specific inputs
```
  be84cac7
- C
  Revert "Simplify size op impl (#45808)" (#46168) · dabb8f23
  由 Chen Weihang 提交于 9月 19, 2022
```
This reverts commit c252b1de.
```
  dabb8f23
13 9月, 2022 1 次提交
- J
  
  cherry pick softmax infer kernel (#45957) · 0903020d
  由 JingZhuangzhuang 提交于 9月 13, 2022
  
  0903020d
09 9月, 2022 1 次提交
- C
  Simplify size op impl (#45808) · c252b1de
  由 Chen Weihang 提交于 9月 09, 2022
```
* simplify size op

* trans to cuda manuly

* fix copy error
```
  c252b1de
07 9月, 2022 1 次提交
- W
  [OpAttr]Adapt tensor output_size for conv2d_transpose and depthwise_conv2d_transpose (#45620) · fe169bf1
  由 WangZhen 提交于 9月 07, 2022
```
Adapt tensor output_size for conv2d_transpose and depthwise_conv2d_transpose
```
  fe169bf1
06 9月, 2022 2 次提交
- Y
  
  migrate unsqueeze kernels to phi, test=kunlun (#45673) · 4acf1ef7
  由 ykkk2333 提交于 9月 06, 2022
  
  4acf1ef7
- W
  
  Completes basic dtypes for collective api in eager mode (#45574) · 7a92e74b
  由 Wen Sun 提交于 9月 06, 2022
  
  7a92e74b
01 9月, 2022 1 次提交

[phi] Migrate uniform_random XPU kernel to PHI (#45583) · ded33b58

由 HongyuJia 提交于 9月 01, 2022

* copy kernel file to phi

* delete some code

* migrate uniform_random, test=kunlun

* fix input error, test=kunlun

* fix gpu register error, test=kunlun

* add include file, test=kunlun

* try fix error from CI, test=kunlun

* polish other PR

* fix CI-coverage error, test=kunlun

ded33b58

31 8月, 2022 6 次提交

D
enhance grid_sampler cpu kernel to 5D input (#45578) · 663ebd5f
由 duanyanhui 提交于 8月 31, 2022
```
* enhance grid_sampler cpu kernel to 5D input

* fix bug when 5D input tensor running on the cudnn kernel
```
663ebd5f

[PHI]Move elementwise div/mul of XPU kernel to PHI (#45581) · f41b8566

由 YuanRisheng 提交于 8月 31, 2022

* move elementwise test=kunlun

* move add/sub/mul/div kernel to elementwise_kernel, test=kunlun

* fix ci bugs,test=kunlun

* fix ci bugs

* test=kunlun

f41b8566

[phi] Migrate truncated_gaussian_random XPU kernel to PHI (#45529) · c2942144

由 HongyuJia 提交于 8月 31, 2022

* migrate truncated_gaussian_random kernel to phi, test=kunlun

* reuse CPU kernel, test=kunlun

* debug kernel, test=kunlun

* migrate truncated_gaussian_random kernel to phi, test=kunlun

* split truncated_normal, test=kunlun

* try fix error from CI, test=kunlun

c2942144

A
[OpAttr]output_size of unpool support Tensor type (#45543) · 236ac0d0
由 Aurelius84 提交于 8月 31, 2022
```
* [OpAttr]output_size of unpool support Tensor type

* fix coverage

* fix contain_var

* fix coverage
```
236ac0d0

Fix split api bug (#45396) · 4a25b60d

由 Charles-hit 提交于 8月 31, 2022

* fix split bug

* solve function redefine

* fix fluid.layers.split and add unit test

* delete splitInferMeta register in unary.cc

* modify test_split_op GPU unit test

* modify test_split_op GPU unit test place param

* refactor split op and fix infershape bugs

* add () in && and ||

* fix split C++ unit test

* fix split infershape

4a25b60d

L

Add index add API (#45176) · 45171911
由 Li Min 提交于 8月 31, 2022

45171911

30 8月, 2022 4 次提交
- W
  [OpAttr]Adapt tensor axis for argmin/max (#45453) · 6fc15986
  由 WangZhen 提交于 8月 30, 2022
```
* Adapt tensor axis for argmin/max

* Add UT

* Polish UT
```
  6fc15986
- W
  [OpAttr]Adapt tensor axis for reduce_min/max/mean/sum/prod (#45078) · 32f42e94
  由 WangZhen 提交于 8月 30, 2022
```
* [OpAttr]Adapt tensor axis for reduce_min/max/mean/sum/prod
```
  32f42e94
- W
  
  Adapt tensor num_samples for multinomial (#45522) · c857841e
  由 WangZhen 提交于 8月 30, 2022
  
  c857841e
- C
  
  rename mod c api name (#45476) · ad96fe2c
  由 Chen Weihang 提交于 8月 29, 2022
  
  ad96fe2c
29 8月, 2022 1 次提交

[geometric]Move graph-related incubate api to geometric (#44970) · 8f657f74

由 Siming Dai 提交于 8月 29, 2022

* move incubate to geometric

* add paddle.geometric

* fix unittest bug

* add float16 support for segment op

* change reindex and sample neighbors flag name

* add heter graph reindex

* move sample_neighbors.py to neighbors.py

* delete khop_sampler in geometric

* delete unused code

* change sample_neighbors api input order

* fix en doc

* fix unittest

* fix unittest

* change reindex

* fix division by 0

* delete unnecessary input argument

* delete final_state

8f657f74

25 8月, 2022 4 次提交

Enable OMP multithreading in lookup_table_v2 (#45249) · 0c363de8

由 piotrekobi 提交于 8月 25, 2022

* Add omp parallel for directives

* Revert "Add omp parallel for directives"

This reverts commit f4e4f8ddb12454018d9c1e49c074af2543659de6.

* Add #pragma omp parallel for to correct file

* Add check for _OPENMP definition

* Disable omp on gpu

* Trigger CI

* Readd check for _OPENMP definition

* Change macro disabling changes on GPU

* Improve macro readability

0c363de8

A
[OpAttr]min/max of uniform_random support Tensor type (#45417) · c8955d0d
由 Aurelius84 提交于 8月 25, 2022
```
* [OpAttr]min/max of Uniform_rand support Tensor type

* fix typo
```
c8955d0d
S
make full_like support double_max in dygraph (#45385) · edd66f2e
由 Sing_chan 提交于 8月 25, 2022
```
* make full_like support double_max in dygraph

* fix bug
```
edd66f2e
R

[triu_indices] add triu_indices_op (#45168) · a410c397
由 Rayman 提交于 8月 25, 2022

a410c397

24 8月, 2022 3 次提交

make tensor_util contains no cuda code (#45256) · 78916a7a

由 Leo Chen 提交于 8月 24, 2022

* make tensor_util contains no cuda code

* refine isfinite

* revert ut

* move isfinite function to its op

* fix test

* fix compile

* std::isnan is not defined for int type on windows

* fix windows compile

* fix fp16

* fix rocm compile

* revert gradient node

78916a7a

W

Adapt tensor axis for cumsum (#45372) · 7f49b9ba
由 WangZhen 提交于 8月 24, 2022

7f49b9ba
W
[OpAttr]Adapt tensor minlength for bincount (#45342) · 12917c8c
由 WangZhen 提交于 8月 24, 2022
```
* Adapt minlength attr for bincount
```
12917c8c

23 8月, 2022 1 次提交

[Phi]Move distribute_fpn_proposals to PHI (#45212) · 8f8ed7de

由 YuanRisheng 提交于 8月 23, 2022

* move distribute_fpn_proposals

* fix some code

* fix yaml bugs

* add set dtype

* move proposal_impl to funcs

* fix compile bugs

8f8ed7de

22 8月, 2022 1 次提交
- W
  [Eager] some python c api use final state (#45221) · d2ef888b
  由 wanghuancoder 提交于 8月 22, 2022
```
some python c api use final state
```
  d2ef888b
18 8月, 2022 2 次提交

[phi] Transfer fluid trilinear_interp_v2 to phi trilinear_interp (add yaml) (#45145) · 6150fade

由 HongyuJia 提交于 8月 18, 2022

* transfer trilinear op to phi, change name from trilinear_interp_v2 to trilinear_interp

* reserve linear_interp param

* change testcase scale if-branch

* testcase test_imperative_case

* fix trilinear testcase

* import paddle in test_trilinear_interp_v2

6150fade

[phi] Transfer fluid bilinear_interp_v2 to phi bilinear_interp (add yaml) (#45140) · 2c2137bb

由 HongyuJia 提交于 8月 18, 2022

* transfer bilinear op to phi, change bname from bilinear_interp_v2 to bilinear_interp

* reserve linear_interp param

* fix cross device import

2c2137bb

17 8月, 2022 1 次提交

[phi] Transfer fluid bicubic_interp_v2 to phi bicubic_interp (add yaml) (#45151) · f4da2d4d

由 HongyuJia 提交于 8月 17, 2022

* transfer bicubic_interp op to phi, change name from bicubic_interp_v2 to bicubic_interp

* test final_state_bicubic_interp api

* testcase match imperative case

f4da2d4d

16 8月, 2022 3 次提交

[Phi] Move amp ops into phi (#45079) · b4f67757

由 Chen Weihang 提交于 8月 16, 2022

* move check finite and unscale kernel into phi

* move infershape into phi

* move update_loss_scaling kernel into phi

* remove original kernels

* move update loss scaling infershape into phi

* add header for xpu and npu

* solve coverage failed

* fix npu test failed

* remove mutable data in cu file

* fix new executor failed

* add valid check for meta tensor output

b4f67757

[geometric]Add paddle.geometric.send_uv API (#44848) · 88724a53

由 Siming Dai 提交于 8月 16, 2022

* initial commit

* fix op maker bug

* fix mul grad bug

* add unittest

* fix add grad bug, add cpu kernel

* add paddle.geometric.message_passing

* add paddle.geometric.send_uv api, add unittest

* add fp16 judgement

* fix file typo, move compute_type to message_op

* add impl file

* fix unittest timeout time

* add review revise

88724a53

H

transfer nearest_interp op to phi, change name from nearest_interp_v2 to nearest_interp (#45148) · 6452ab3b
由 HongyuJia 提交于 8月 16, 2022

6452ab3b

15 8月, 2022 1 次提交

[phi] change op name linear_interp_v2 to linear_interp (#45128) · 6de3bdb3

由 HongyuJia 提交于 8月 15, 2022

* change name linear_interp_v2 to linear_interp

* fix deprecated_op_names

* deprecated_op_names add linear_interp_grad

6de3bdb3

BaiXuePrincess / Paddle 与 Fork 源项目一致

BaiXuePrincess / Paddle
与 Fork 源项目一致