提交 · 96d530c1286936c78b5ad6869d926e159b4563b5 · PaddlePaddle / Paddle

23 2月, 2022 1 次提交
- C
  
  move array_ref_test and small_vector_test into paddle/utils and format header macro define (#39831) · 96d530c1
  由 chentianyu03 提交于 2月 23, 2022
  
  96d530c1
22 2月, 2022 2 次提交

C

import llvm::ArrayRef and add test (#39802) · a167a143
由 chentianyu03 提交于 2月 22, 2022

a167a143

change Vector to std::vector and provide MixVector class as a helper … (#39559) · 728c0624

由 xiongkun 提交于 2月 22, 2022

* change Vector to std::vector and provide MixVector class as a helper wrapper class

* solve the multi-gpu hang problem

* remove the duplicate template instantialize

* Copy vector to cpu

* add CopyToCPU

* xxx

* final version: fix the problem of all reduce

* remove mixvector dependence

* fix

* merge

* fix code

* fix by CI

728c0624

25 1月, 2022 1 次提交

[PTen] Migrate string tinyformat errors and part of enforce into pten (#39051) · 6ca49164

由 xiongkun 提交于 1月 25, 2022

* transfer: string tinyformat errors and part of enforce into pten

* remove comment

* fix by code review

* assert is not compile in -DNDEBUG

* add string as dependences of paddle_inference

6ca49164

11 1月, 2022 1 次提交
- S
  support vs2019 compilation in windows (#38719) · 0ad363b1
  由 Sing_chan 提交于 1月 11, 2022
```
* support vs2019 compilation in windows

* not modify pow_op's original compute logic
```
  0ad363b1
01 11月, 2021 1 次提交

Paddle Tensor Operation Library initial implementation (#34425) · b9fdd3bc

由 Chen Weihang 提交于 11月 01, 2021

* initial tensor design & sign kernel demo

* add move constructor for meta & add lodtensor

* add dirs & sign xpu kernel

* add mean cpu&cuda kernel impl

* move sign & mean xpu & npu kernel

* add selected_rows basic impl

* refactor design, BaseTensor to DenseTensor, etc.

* add scale mkldnn kernel

* polish xpu & npu impl details

* fix mkldnn reuse compile failed

* change tensor operation lib name

* rename util filename

* add more comments

* change TensorImplInterface to TensorInterface

* add kernel key and factory

* remove MKLDNNTensorMeta, add MKLDNNDenseTensor

* change XXDeviceContext to XXContext

* add base kernel registrar utils & test on sign

* replace boost::any by paddle::any

* fix several ci failed

* fix npu compile error

* add ordered map util

* fix multiple ordered_map compile errors

* move dev into include dir

* support sign op in static op run

* fix static op run error

* fix new executor compile failed

* add dygraph branch & remove sign_op.h

* fix test_infer_no_need_buffer_slots

* fix rocm compile link error

* fix unitybuild error & clear glog

* fix npu compile failed

* skip quant trans test

* fix part windows compile problem

* fix xpu enforce error

* fix inference test failed

* remove ordered_map to solve quant failed

* fix part of rcom compile faild

* add more register kernels

* revert scale kernel temporarily

* fix code format error

* add new kernel registrar marco

* rename top to tcmpt

* revert xpu, npu, mkldnn impl & remove op def

* add kernel args parse functor to auto parse args

* revert some change & add scale kernels

* add op proto in dygraph kernelcontext building

* polish kernel dispatch logic & nameing rule

* fix scale kernel match error

* fix scale test failed

* add mean API and unittest

* test mean api success

* add branch to solve compiled error

* skip clang format error

* add mean skip rule in op_library

* add dot kernel, api and unittest (#6)

* remove old kernel and add symbol link

* fix dot compiled failed

* add merco for module declare

* fix npu and xpu compile error

* revert sign, mean, scale, dot kernel removing

* add comment for keeping old kernel impl

* fix mutable_data error

* fix bfloat16 conflit

* fix inference undef error

* adapt to msvc compile rules

* polish comment for template inst

* add cmake template instantiation for win

* fix backend to place device id bug

* fix ifdef error

* Op2functor (#7)

* add kernel args maker class

* make args maker non-const

* remove debug log

* modify codes by review options

* split constructPrKernelContext function

* fix output name bug

* fix test_mean_op test_sign_op failed

* fill_any_like kernel refactor (#10)

* fill_any_like kernel refactor

* remove useless code of full_like c++ api

* skip dtype for fill_any_like

* add attrs for kernel key constrcut

* add use_pt_kernel Flags to control whether to use pt kernel (#13)

* add use_pt_kernel Flags to control whether to use pt kernel

* change the default value to true for cheking pt kernels

* fix mutable_data cuda place error

* move high level apis into hapi

* remove selectedrows adapting temporarily

* Support Scalar in Tensor Compute Library (#14)

* fill_any_like kernel refactor

* remove useless code of full_like c++ api

* Support Scalar in Tensor Compute Library

* add scalar in dygraph and static graph mode

* keep the basic type for attr, instead of using scalar for all

* merge the code

* remove mkldnn tensor & polish details

* use flat_hash_map and small_vector in kernel factory

* Refactor flatten kernel (#12)

* refactor flatten kernel

* update infershape function

* fix compile bugs

* fix bugs when merge

* fix compiler bugs

* fix bugs when run test_flatten_api

* fix bugs when run test

* Revert "use flat_hash_map and small_vector in kernel factory"

This reverts commit 23091495cfdd3df8cc1be592d30f09ea66a7c72b.

* Move cpu, cuda and other device code into kernels (#15)

* fill_any_like kernel refactor

* remove useless code of full_like c++ api

* Support Scalar in Tensor Compute Library

* add scalar in dygraph and static graph mode

* keep the basic type for attr, instead of using scalar for all

* merge the code

* start refactor matmul

* move cpu, cuda and other device modules into kernels

* merge code

* polish code in operator.cc

* Perfect unitests (#16)

* perfect unittest

* update license

* replace with flat_hash_map, small_vector (#19)

* fix small_vector build error on windows platform

* replace with flat_hash_map, small_vector

* remove todo

* Perfect unitests (#20)

* perfect unittest

* update license

* fix bug when run tcmpt_utils_test

* refactor execution adapting impl

* fix insert conflit

* Fix CI bug of test_yolov3 (#21)

* fill_any_like kernel refactor

* remove useless code of full_like c++ api

* Support Scalar in Tensor Compute Library

* add scalar in dygraph and static graph mode

* keep the basic type for attr, instead of using scalar for all

* merge the code

* start refactor matmul

* move cpu, cuda and other device modules into kernels

* merge code

* polish code in operator.cc

* Fix CI bug of test_yolov3

* add the tensor base class, test=develop (#17)

* update the tensor base class, test=develop

* remove two funcs, test=develop

* update the error msg, test=develop
Co-authored-by: NChen Weihang <chenweihang@baidu.com>

* [no-verify] commit backend and tensor signature changes

* Rename tcmpt to pten (#23)

* rename tcmpt to pten

* update omitted files for rename to pten

* update omitted file for rename to pten

* remove k of all enum var

* remove kernel_instantiate (#26)

* remove symbols and spatial_tensor

* change common to functions

* readd share tensor impl methods

* add a candidate dense tensor class, test=develop (#28)

* change all Pt to Pten

* resolve conflit with xiaowei

* Op2functor opt1 (#27)

* replace to small vector and change to const &

* add std::move
Co-authored-by: NChen Weihang <chenweihang@baidu.com>

* polish kernel factory and kernel registry

* fix operator test error msg mismatch

* remove tensor signature and backend set member

* move scalar and polish enforce

* revert dtype layout change to fix error

* fix enum operator override error

* add several base unittests

* add pten utils tests

* polish some details

* Dev/op2func refactor 3 (#30)

* add a candidate dense tensor class, test=develop

* remove TensorBase::backend(), test=develop

* remove some ops, test=develop

* cherry-pick the pr of tensor meta, test=develop

* moves the dense tensor and some ops, test=develop

* update the linalg operator, test=develop

* update other operators, test=develop

* fix errors, test=develop

* fix bugs, test=develop

* try to resolve the problem of windows ci, test=develop

* updates codes, test=develop

* fix the tensor_utils.cc, test=develop

* modify the dense tensor, test=develop

* fix the data type, test=develop
Co-authored-by: Nshixiaowei02 <39303645+Shixiaowei02@users.noreply.github.com>

* polish some details

* polish kernel signature details

* fix a bug about offsets of the tensor, test=develop (#31)
Co-authored-by: Nshixiaowei02 <39303645+Shixiaowei02@users.noreply.github.com>

* polish some details
Co-authored-by: Nchentianyu03 <ctychentianyu@gmail.com>
Co-authored-by: Nzyfncg <1370305206@qq.com>
Co-authored-by: NYuanRisheng <yuanrisheng@baidu.com>
Co-authored-by: N石晓伟 <39303645+Shixiaowei02@users.noreply.github.com>

b9fdd3bc

10 9月, 2021 2 次提交

add llvm::SmallVector to paddle (#34832) · 11965bca

由 chentianyu03 提交于 9月 10, 2021

* add llvm::SmallVector to paddle

* rename small vector file

* merge paddle small vector to one file

* add small_vector_test

* modify smallvector test argument type

* add string header

11965bca

import ska flat_hash_map (#34464) · 3d9603dc

由 chentianyu03 提交于 9月 10, 2021

* import ska flat_hash_map

* add define NOMINMAX macro to fix windows build failed bug

* add brackets to std::max in flat_hash_map

* move flat_hash_map directions

* modify namespace to paddle

* modify namespace to paddle

* modify namespace to paddle

* modify namespace to paddle

* rm not used map.h and replace with op_info

3d9603dc

17 8月, 2021 1 次提交

Copy boost optional to Paddle (#34780) · 9be41447

由 chentianyu03 提交于 8月 17, 2021

* copy boost optional.hpp to paddle

* copy boost optional.hpp to paddle

* move directions

* del fluid/utils

* modify .hpp to .h

* move directions

* modify to paddle::optional

* add modification description

* format code stype for the files in paddle/utils

* format code stype

9be41447

10 8月, 2021 1 次提交

copy boost/any.hpp to utils and replace boost::any with self defined any (#34613) · 12892929

由 chentianyu03 提交于 8月 10, 2021

* add any.hpp to utils and replace boost::any with self defined paddle::any

* add copy any.hpp to custom op depends

* modify any.hpp include path

* remove boost from setup.py.in

* add copy any.hpp to custom op depends

* move any.hpp to paddle/utils/ dirs

* move any.h to extension/include direction

* copy utils to right directions

12892929

03 7月, 2018 1 次提交
- X
  
  hide utils to legacy · 94cb59ad
  由 Xin Pan 提交于 7月 03, 2018
  
  94cb59ad
29 5月, 2018 1 次提交
- W
  fix develop build issue (#10978) · 8f7b020b
  由 Wu Yi 提交于 5月 29, 2018
```
* fix develop build issue

* fix google style

* cpplint check only fluid
```
  8f7b020b
24 5月, 2018 1 次提交
- Y
  
  Remove cpplint in cmake · a229734c
  由 yuyang18 提交于 5月 24, 2018
  
  a229734c
27 4月, 2018 1 次提交
- F
  
  fix mac compile errors · 31373370
  由 fengjiayi 提交于 4月 27, 2018
  
  31373370
16 4月, 2018 1 次提交
- Y
  
  add tensorrt build support(#9891) · 18665979
  由 Yan Chunwei 提交于 4月 16, 2018
  
  18665979
06 4月, 2018 1 次提交
- L
  
  Build: generate all the build related files into one directory. (#9512) · 09b4a1a3
  由 Lei Wang 提交于 4月 05, 2018
  
  09b4a1a3
21 1月, 2018 1 次提交

"fix decode bug" (#7711) · e983cc90

由 dzhwinter 提交于 1月 21, 2018

* "fix decode bug"

* "follow commnet"

* "fix error"

* "fix hook bug"

* fix based comment

* fix copyright

* fix based on comment

e983cc90

15 1月, 2018 1 次提交

Feature/hooks (#7513) · b9b75377

由 dzhwinter 提交于 1月 15, 2018

* add copyright hook

* add copyright hook

* refine copyright hook

* "test copyright hook"

* fix check style

* fix ci

b9b75377

04 1月, 2018 2 次提交
- T
  
  default disable use_mkl_packed · 3b5e4e0a
  由 tensor-tang 提交于 1月 04, 2018
  
  3b5e4e0a
- T
  
  add flag use_mkl_packed · 042f3524
  由 tensor-tang 提交于 1月 04, 2018
  
  042f3524
12 12月, 2017 1 次提交
- T
  
  unify MKL macro definition · 69b44f2f
  由 tensor-tang 提交于 12月 12, 2017
  
  69b44f2f
06 11月, 2017 1 次提交
- Y
  
  Enable the build for iOS simulator. (#5211) · bba62235
  由 Yiqun Liu 提交于 11月 06, 2017
  
  bba62235
05 10月, 2017 2 次提交

Y

Use PADDLE_WITH_CUDA instead of PADDLE_WITH_GPU · 4558807c
由 Yi Wang 提交于 10月 04, 2017

4558807c

Change `PADDLE_ONLY_CPU` to `PADDLE_WITH_GPU` · 84500f94

由 Yu Yang 提交于 10月 04, 2017

By shell command

```bash
sed -i 's#ifdef PADDLE_ONLY_CPU#ifndef PADDLE_WITH_GPU#g' `find ./paddle/ -name '*.h' -o -name '*.cc' -o -name '*.cpp' -o -name '*.c' -o -name '*.cu'`
sed -i 's#ifndef PADDLE_ONLY_CPU#ifdef PADDLE_WITH_GPU#g' `find ./paddle/ -name '*.h' -o -name '*.cc' -o -name '*.cpp' -o -name '*.c' -o -name '*.cu'`
```

84500f94

07 9月, 2017 1 次提交
- H
  
  Fix android-16 compile. · 49661746
  由 hedaoyuan 提交于 9月 07, 2017
  
  49661746
29 8月, 2017 1 次提交
- L
  
  Seperate the codes that cannot and don't need to build for iOS devices. · fb38e662
  由 Liu Yiqun 提交于 8月 29, 2017
  
  fb38e662
17 8月, 2017 1 次提交
- T
  
  remove flag use_mkldnn_wgt · e08651f9
  由 tensor-tang 提交于 8月 17, 2017
  
  e08651f9
10 8月, 2017 1 次提交
- L
  
  Rename PROJ_ROOT to PADDLE_SOURCE_DIR and PROJ_BINARY_ROOT to PADDLE_BINARY_DIR · 7a56d46a
  由 liaogang 提交于 8月 10, 2017
  
  7a56d46a
08 8月, 2017 1 次提交
- T
  
  add test case use_mkldnn_wgt · 6373291c
  由 tensor-tang 提交于 8月 08, 2017
  
  6373291c
04 8月, 2017 1 次提交
- T
  
  add use_mkldnn flag · 3c3a11a0
  由 tensor-tang 提交于 8月 04, 2017
  
  3c3a11a0
27 7月, 2017 1 次提交

Fix bug in SequenceSoftmax · f4e57b4b

由 Yu Yang 提交于 7月 27, 2017

Also remove operator bool in Error. The Error should be removed
later because it is not necessary for Paddle. We are now using Enforce
to handle error.

f4e57b4b

15 7月, 2017 1 次提交
- L
  
  FIX: cppint code style · 569f7e83
  由 liaogang 提交于 7月 15, 2017
  
  569f7e83
05 7月, 2017 1 次提交
- Y
  Correct GLOG CHECK in Paddle · 5eb8bf03
  由 Yu Yang 提交于 7月 05, 2017
```
Use CHECK instead of PCHECK, because PCHECK is used for errno.
```
  5eb8bf03
04 7月, 2017 1 次提交

Remove buggy BarrierStat · 1ecddd81

由 Yu Yang 提交于 7月 04, 2017

The implementation of BarrierStat is buggy, and it is not necessary
for Paddle to diagnose which node in cluster is slow.

1ecddd81

28 6月, 2017 3 次提交
- Y
  
  Fix TravisCI · 64b78b16
  由 Yu Yang 提交于 6月 28, 2017
  
  64b78b16
- Y
  Add pb_cc_library in generic.cmake · b1a311c4
  由 Yu Yang 提交于 6月 28, 2017
```
Fix #2567
```
  b1a311c4
- Y
  
  Remove must_check in paddle::platform · 9ad846ec
  由 Yu Yang 提交于 6月 28, 2017
  
  9ad846ec
26 6月, 2017 1 次提交

Adding platform/must_check.h · d76d2feb

由 Yu Yang 提交于 6月 26, 2017

__must_check is a macro mark of function return value. It let developer
must check the return value is legal or not.

d76d2feb

20 6月, 2017 1 次提交

Fix bugs for rnn generation · 3438d650

由 xuwei06 提交于 6月 19, 2017

1. v2.layer.parse_network does not correctly handle the generation output.
2. GatherAgentLayer does not correctly handle generation output when batch_size > 1
3. Fix CustomStackTrace for rnn group

3438d650

27 5月, 2017 1 次提交
- L
  
  Support native build on NVIDIA DRIVE PX2 (arm64 + GPU). · 07ac67ec
  由 Liu Yiqun 提交于 5月 26, 2017
  
  07ac67ec

PaddlePaddle / Paddle 1 年多 前同步成功

PaddlePaddle / Paddle
1 年多前同步成功