提交 · 787273eda531ae0cd651532eb39850a3614bfd67 · 机器未来 / Paddle

24 9月, 2021 1 次提交

Added elementwise_sub_mkldnn operator (#35662) · 787273ed

由 piotrekobiIntel 提交于 9月 24, 2021

* Add elementwise_sub_mkldnn_op without grad

* Add test to static_mode_white_list

* Refactor code, change license years

* Remove invalid grad implementation

* Fix element_wise_sub_op test

* Fix CI Approval error

* Remove unnecessary EltwiseSubMKLDNNGradKernel class

* Fix CI Approval 2

* Fix CI Approval 3

* Fix CI Approval Attempt #4

* Fix CI Approve Attempt #5

* Fix CI Approval Attempt #6

* Fix CI Approval Attemt #7

* Change test names containing add to sub

* Fix old tests testing add instead of sub

* Copy grad implementation from elementwise_add_mkldnn

* CI test fix attempt

* Revert "CI test fix attempt"

This reverts commit c647cacf41e6a87c715385a185de5cbf65fc8900.

* Fix CI attempt 2

* Fix elementwise_sub tests, temporary mkldnn broadcast test disable

* Add working implementation of elementwise_sub grad

* Fix build errors caused by pull

* Fix format error

* Fix format error 2

* Disable elementwise_sub_mkldnn test on GPU

* Apply fix for paddle.fluid import

* Revert changes of test_elementwise_sub and Fix mkldnn test

* Revert "Apply fix for paddle.fluid import"

This reverts commit fc3b122fec8e12f2bcb32928a2685ba4d20fd742.

* fix bug of module 'paddle' has no attribute 'fluid' for python3.6 (#35862)

* Add changes suggested by reviewers

* Change @unittest.skipIf... to @OpTestTool.skip_if_not_cpu_bf16() to satisfy Approval CI

* Remove check_dygraph=False to satisify CI Approval
Co-authored-by: Nzhangbo9674 <82555433+zhangbo9674@users.noreply.github.com>

787273ed

17 9月, 2021 1 次提交

Support EMA in Paddle2.x and Fleet (#35673) · fb4d5689

由 Haohongxiang 提交于 9月 17, 2021

* Support EMA in Paddle2.x and Fleet

* update

* update

* update

* modify ut of ema

* modify docs

* modify bugs

* update

* update

* update

* modify ut

fb4d5689

14 9月, 2021 1 次提交
- Z
  add paddle.Tensor api fill_(inplace), zero_(inplace) (#33829) · efeec79b
  由 zhiboniu 提交于 9月 14, 2021
```
add fill_ backward
```
  efeec79b
10 9月, 2021 1 次提交
- Z
  
  add api_op fill_diagonal_tensor (#34515) · 98d047d7
  由 zhiboniu 提交于 9月 10, 2021
  
  98d047d7
27 8月, 2021 1 次提交
- G
  sparse_momentum_op is used to save w@GRAD memory for gather_op (#34942) · 234ce932
  由 Guoxia Wang 提交于 8月 27, 2021
```
* sparse_momentum_op is used to save w@GRAD memory for gather_op when gather from a large parameter
```
  234ce932
18 8月, 2021 1 次提交
- G
  support class center sample of PartialFC (#34106) · 100db44f
  由 Guoxia Wang 提交于 8月 18, 2021
```
* support class center sample of PartialFC
```
  100db44f
16 8月, 2021 1 次提交
- G
  support margin loss (arcface, cosface, sphereface) for single GPU and cross GPUs (#34247) · b0cb4148
  由 Guoxia Wang 提交于 8月 16, 2021
```
* support margin loss (arcface, cosface, sphereface)
```
  b0cb4148
07 7月, 2021 1 次提交
- J
  Added PRelu BF16/FP32 FWD/BWD kernels (#33878) · 375e5618
  由 jakpiase 提交于 7月 07, 2021
```
* added prelu bf16/fp32 fwd/bwd kernel
```
  375e5618
30 6月, 2021 1 次提交

Added matmul_v2 BF16/FP32 FWD kernel (#33750) · 24783c84

由 jakpiase 提交于 6月 30, 2021

* added matmul_v2 bf16/fp32 FWD kernel

added matmul_v2 bf16/fp32 FWD kernel

* added formatting

* removed some tests due to timeout in CI

* refactored tests

* merged tests classes into one file

* minor change

* removed test guard for CUDA

* remove skipIf

* changes after review

* formated one file

* minor change

* added skipping UT in CUDA place

24783c84

24 6月, 2021 1 次提交
- C
  supplet several interface of static Variable to consistent with dygraph Tensor (#33330) · af9dcb2d
  由 CtfGo 提交于 6月 24, 2021
```
As the title
```
  af9dcb2d
23 6月, 2021 1 次提交

Added split op bf16/fp32 oneDNN kernel (#33584) · 68106509

由 jakpiase 提交于 6月 23, 2021

* base changes for split op

* 90% of split functionality added

* full fp32 functionality

* added bf16 test

* added submemory caching

* added bf test to static mode whitelist

* minor change

* enabled split op for inference

* minor fix

* minor fix

68106509

17 6月, 2021 2 次提交

Add bf16 support for save and load ops (#33173) · 832a014c

由 joanna.wozna.intel 提交于 6月 17, 2021

* Add bf16 support for save and load ops

* Add bf16 test condition

* Add matmul and chagne fluid.io to paddle.static

* Reduce the test duration

832a014c

Add lookup_table_v2 BF16 op (#33172) · 9d6c8bdf

由 joanna.wozna.intel 提交于 6月 16, 2021

* Add lookup_table_v2 BF16

* Reuse lookup table UT

* Change op_type to op_version

* Remove check_dygraph

* Remove skip_check_grad_ci

9d6c8bdf

07 6月, 2021 1 次提交
- S
  [HybridParallel]Fix c_split op for TensorParallel (#33207) · 902c6f98
  由 ShenLiang 提交于 6月 07, 2021
```
* fix c_split bug

* fix utest

* add c_embedding for tensorparallel
```
  902c6f98
26 5月, 2021 2 次提交

Y

Marker op for profiling (#33034) · 5c79dbb2
由 Yuang Liu 提交于 5月 26, 2021

5c79dbb2

Added cast op oneDNN kernel for bf16/fp32 datatypes casting(FWD/BWD) (#33056) · a2a45d8d

由 jakpiase 提交于 5月 26, 2021

* added op cast functionality for fp32/bf16

* added newline

* added entries in static mode white list and unity build

* fixed failing tests

* changes after review

* added formatting

* upgraded tests file as reviewer suggested

* changes after review

* minor change

a2a45d8d

25 5月, 2021 1 次提交
- J
  
  Added scale op FP32/BF16 FWD/BWD kernels (#32975) · 86ea8dce
  由 jakpiase 提交于 5月 25, 2021
  
  86ea8dce
29 4月, 2021 1 次提交

Add BF16 uniform random initializer (#32468) · f46f15a0

由 joanna.wozna.intel 提交于 4月 29, 2021

* Add bf16 uniform random initializer

* Remove duplicated section

* Change UT to CPU place only

* Put detail functions into anonymous namespace

f46f15a0

21 4月, 2021 1 次提交
- J
  
  Added bilinear and nearest interp v2 oneDNN FP32 kernels (#32312) · 5d19f8d8
  由 jakpiase 提交于 4月 21, 2021
  
  5d19f8d8
14 4月, 2021 2 次提交

J

Added oneDNN reduce_op FWD kernel (#31816) · 3a804a0e
由 jakpiase 提交于 4月 14, 2021

3a804a0e

adds new CPU kernel for SGD op supporting BF16 data type (#32162) · 3ac6c189

由 Adam Osewski 提交于 4月 14, 2021

* Initial draft for SGD BG16 kernel.

* Unit tests for SGD with BF16 data type.

* Add VLOG message to SGD BF16 op CPU kernel.

* Enhance error messages and error types.

* Refactor SGD op kernels to leverage some common code.

* Make easier to add new kerne invoke code.

* Fix SGD op kernel for sparse grad.

* Unify quotes style.

* Fix error for ROCM compilation.

* Use specialized PADDLE_ENFORCE_xx functions.

3ac6c189

30 3月, 2021 1 次提交
- J
  
  Added int8 kernel for oneDNN LSTM op (#31894) · 6dca7a1d
  由 jakpiase 提交于 3月 30, 2021
  
  6dca7a1d
22 3月, 2021 1 次提交
- A
  
  [oneDNN] Initial bf16 amp integration (#31093) · 7ccf6b60
  由 arlesniak 提交于 3月 22, 2021
  
  7ccf6b60
19 3月, 2021 1 次提交
- A
  
  [oneDNN] lookup_table op with support for BF16 data type. (#31558) · a4a2b77d
  由 Adam Osewski 提交于 3月 19, 2021
  
  a4a2b77d
04 3月, 2021 1 次提交
- J
  
  Added LSTM BF16 and fixed GRU BF16 (#31234) · 5b4f8aac
  由 jakpiase 提交于 3月 04, 2021
  
  5b4f8aac
02 3月, 2021 1 次提交

lamb_op_xpu;test=kunlun (#31012) · d79fdc3d

由 Gradie 提交于 3月 02, 2021

* lamb_op_xpu;test=kunlun

* modify lamb_op_xpu.cc;test=kunlun

* delete atol lamb_op_xpu; test=kunlun

* update xpu.cmake;test=kunlun

* test_error 1e-5,lamb_op_xpu;test=kunlun

* error1e-5,lamb_op_xpu,test=kunlun

* delete atol lamb_xpu;test=kunlun

* modify atol,lamb_op_xpy;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu, XPUOptest;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu,modify xpu_cmake; test=kunlun

* lamb_op_xpu;test=kunlun

* lamb_op_xpu,modify xpucmake;test=kunlun

d79fdc3d

18 2月, 2021 1 次提交

Add Conv Transpose BF16 (#30877) · caf9d398

由 joanna.wozna.intel 提交于 2月 18, 2021

* Add conv transpose BF16

* Share function GetWeightsTz

* Adjust to review and fix op compatibility

* Add bias to unique handler name

* Remove errors related to paddle enforce

* Add conv2d_transpose to bf16 list and kernel refator

caf9d398

03 2月, 2021 1 次提交
- A
  
  Layer normalization fuse pass. (#30721) · 4f066e31
  由 Adam Osewski 提交于 2月 03, 2021
  
  4f066e31
27 1月, 2021 1 次提交

REUPLOAD Added vanilla LSTM and LSTM with peepholes oneDNN fp32 kernel (#30719) · f8da5536

由 jakpiase 提交于 1月 27, 2021

* added external reorder to profiler

* resolved conflict

* added enable_static

* initial version of lstm, not working yet

* added lstm to operators.cmake

* added vanilla lstm mkldnn op

* added peephole weights integration

* minor changes

* added formatting

* added fusion_lstm_mkldnn to static_whitelist

* added formatting

* removed comment

* moved use_peepholes attribute inside is_cached block

* reverted wrong changes

* minor formatting change

* minor changes

* changed stream handling

* minor change

* added datatype to GetExpectedKernelType()

* added reading stream from TLS

f8da5536

26 1月, 2021 2 次提交

T
Revert "Added vanilla LSTM and LSTM with peepholes oneDNN fp32 kernel (#30661)" (#30708) · 824a79d3
由 Tao Luo 提交于 1月 26, 2021
```
This reverts commit d834f4e6.
```
824a79d3

Added vanilla LSTM and LSTM with peepholes oneDNN fp32 kernel (#30661) · d834f4e6

由 jakpiase 提交于 1月 26, 2021

* added external reorder to profiler

* resolved conflict

* added enable_static

* initial version of lstm, not working yet

* added lstm to operators.cmake

* added vanilla lstm mkldnn op

* added peephole weights integration

* minor changes

* added formatting

* added fusion_lstm_mkldnn to static_whitelist

* added formatting

* removed comment

* moved use_peepholes attribute inside is_cached block

* reverted wrong changes

* minor formatting change

* minor changes

d834f4e6

31 12月, 2020 1 次提交

Add mkldnn nearest_interp and bilinear_interp op (#30016) · c3c064a8

由 cc 提交于 12月 31, 2020

* Add mkldnn nearest_interp and bilinear_interp op
* don't run mkldnn interpolate in default
* add interpolate_mkldnn_pass

c3c064a8

21 12月, 2020 1 次提交
- L
  Add double grad for conv_transpose (#29706) · e5af650b
  由 LielinJiang 提交于 12月 21, 2020
```
* add double grad for conv_transpose
```
  e5af650b
09 12月, 2020 1 次提交

remove addcmul (#28937) · dc8bb76c

由 Wei Shengyu 提交于 12月 09, 2020

* remove addcmul

* remove unittest and other related code of addcmul

* fix bug

* fix merge conflict

dc8bb76c

26 11月, 2020 1 次提交
- J
  Add bf16 pool2d and unify bf16 unit tests (#29039) · b0d1ac16
  由 joanna.wozna.intel 提交于 11月 26, 2020
```
* Add bf16 pool2d and unify bf16 unit tests

* Add change default ops test
```
  b0d1ac16
25 11月, 2020 1 次提交
- W
  Add multi_gru_fuse_pass and tests (#28601) · 7b5a8e46
  由 Wojciech Uss 提交于 11月 25, 2020
```
* Add multi_gru_fuse_pass and tests

* fix date

* cleaned up headers
```
  7b5a8e46
24 11月, 2020 1 次提交
- W
  Add multi_gru_seq_fuse_pass and tests (#28604) · 991345b3
  由 Wojciech Uss 提交于 11月 24, 2020
```
* Add multi_gru_seq_fuse_pass and tests

* fix date

* removed unused functions
```
  991345b3
20 11月, 2020 1 次提交
- J
  Add bf16 matmul, fc, elementwise add and mul (#28729) · 8c0ea4bf
  由 joanna.wozna.intel 提交于 11月 20, 2020
```
* Add bf16 matmul, fc, elementwise add and mul

* Correct unit test
```
  8c0ea4bf
19 11月, 2020 1 次提交
- W
  Add multi_gru op and tests (#28591) · 04bcc13f
  由 Wojciech Uss 提交于 11月 19, 2020
```
* Add multi_gru op and tests

* removed redundant disable_dygraph()
```
  04bcc13f
17 11月, 2020 1 次提交
- J
  
  [oneDNN] Layer norm bf16 kernel (#28619) · 6d8d3d4c
  由 Jacek Czaja 提交于 11月 17, 2020
  
  6d8d3d4c

机器未来 / Paddle 与 Fork 源项目一致

机器未来 / Paddle
与 Fork 源项目一致