提交 · 220eef602e0e0af0d5566231397766cff5d080db · Crayon鑫 / Paddle

24 7月, 2019 1 次提交

Extend Matmul to support matrix multiplication with multiple heads (#18570) · 220eef60

由 Bob Zhu 提交于 7月 24, 2019

* extend matmul op to support multiple head multiplication

With the support of multiple head, the multiplication of two big matrixes is
split into multiplication of several (head_number) small matrixes. e.g. if
Mat A is [3, 24] and Mat B is [24, 4], when multiple A and B with head_number
as 4, Mat A will be split as 4 matrix of [3, 6] and Mat B will be 4 matrix of
[6, 4]. The result of final matrix will be 4 matrix of [3, 4], i.e. [3, 16].

220eef60

21 3月, 2019 1 次提交
- P
  
  fix matmul shape check; test=develop · 0e402989
  由 phlrain 提交于 3月 21, 2019
  
  0e402989
18 9月, 2018 1 次提交
- S
  
  modification · 0718113a
  由 sneaxiy 提交于 9月 18, 2018
  
  0718113a
17 9月, 2018 1 次提交
- S
  
  tiny change to save memory · abf9832c
  由 sneaxiy 提交于 9月 17, 2018
  
  abf9832c
10 5月, 2018 1 次提交
- Y
  
  matmul support float16/double · 27197290
  由 yuyang18 提交于 5月 10, 2018
  
  27197290
08 5月, 2018 2 次提交

Clean OpProtoAndCheckerMaker · 0e78cb69

由 Yu Yang 提交于 5月 08, 2018

Do not use ctor

* Reduce line of codes.
* We can use virtual function for Maker now.
* The implementation does not care what maker holds, it is easier to
refactor later.

0e78cb69

Y

Follow comments and polish code names · fcd31d61
由 Yu Yang 提交于 5月 08, 2018

fcd31d61

07 5月, 2018 1 次提交
- Y
  
  Rewrite Matmul, make code cleaner · c6a6d87f
  由 Yu Yang 提交于 5月 07, 2018
  
  c6a6d87f
19 4月, 2018 1 次提交
- Y
  add semicolon to op registry (#10034) · e04c43d5
  由 Yang Yang(Tony) 提交于 4月 18, 2018
```
* script to add semicolon

* fix typo
```
  e04c43d5
17 4月, 2018 1 次提交
- Y
  
  script to fix all · ce7c2e86
  由 Yang Yang 提交于 4月 16, 2018
  
  ce7c2e86
12 4月, 2018 1 次提交
- S
  Fix cpplint errors for a set of operators (#9837) · 8d3ce01f
  由 Siddharth Goyal 提交于 4月 11, 2018
```
* Fix cpplint errors, round2

* Fix pointer issue
```
  8d3ce01f
12 2月, 2018 1 次提交
- Q
  
  Fix the grammar in copyright. (#8403) · 24509f4a
  由 qingqing01 提交于 2月 12, 2018
  
  24509f4a
10 2月, 2018 2 次提交
- Y
  
  Correct #include path · fc374821
  由 Yi Wang 提交于 2月 09, 2018
  
  fc374821
- Y
  
  Move file to fluid/; Edit CMakeLists.txt · 90648f33
  由 Yi Wang 提交于 2月 09, 2018
  
  90648f33
21 1月, 2018 1 次提交
- C
  
  follow comments · 782ddc5f
  由 chengduoZH 提交于 1月 21, 2018
  
  782ddc5f
19 1月, 2018 1 次提交
- C
  
  follow comments · 0468422d
  由 chengduoZH 提交于 1月 19, 2018
  
  0468422d
18 1月, 2018 3 次提交
- C
  
  modify doc · 259858b4
  由 chengduoZH 提交于 1月 18, 2018
  
  259858b4
- C
  
  code refine · 578d60bf
  由 chengduoZH 提交于 1月 18, 2018
  
  578d60bf
- C
  
  add 4-d for matmul_op · 2edc136c
  由 chengduoZH 提交于 1月 18, 2018
  
  2edc136c
20 12月, 2017 1 次提交
- Y
  Move framework.proto to proto namespace (#6718) · e445b3ff
  由 Yu Yang 提交于 12月 20, 2017
```
* Move framework.proto to proto namespace

* Fix compile

* Fix compile

* Fix Compile
```
  e445b3ff
12 12月, 2017 1 次提交

Refine device context (#6433) · 61ec0b95

由 QI JUN 提交于 12月 12, 2017

There are mainly following fixes:

- take `DeviceContext` as the template parameter of math functors and OpKernel instead of `Place`
- remove `eigen_device` interface in base class  `DeviceContext`
- remove `GetEigenDevice` interface in `ExecutionContext` and base class `DeviceContext`
- remove unused `platform::EigenDeviceConverter`
- rename `REGISTER_OP_GPU_KERNEL` to `REGISTER_OP_CUDA_KERNEL`
- rename `USE_GPU_ONLY_OP` to `USE_CUDA_ONLY_OP`

61ec0b95

05 11月, 2017 1 次提交
- K
  Polish Operator Doc (m) (#5375) · cb0118f3
  由 kexinzhao 提交于 11月 04, 2017
```
* fix m_ops

* fix activation op
```
  cb0118f3
18 10月, 2017 1 次提交

MatMul operator (#4856) · 16489827

由 Markus Kliegl 提交于 10月 17, 2017

* initial matmul operator

Similar to np.matmul, but also has transpose_X and transpose_Y flags,
and only supports tensors from rank 1 to 3 inclusive.

For GPU, uses cublas?gemmStridedBatched. For CPU, uses
cblas_?gemm_batch if available via MKL; otherwise a simple serial
implementation that loops over the batch dimension is employed for now.

16489827

Crayon鑫 / Paddle 与 Fork 源项目一致

Crayon鑫 / Paddle
与 Fork 源项目一致