提交 · d3f627d22ea5c89e4adb382aee3cae5e71363d3a · PaddlePaddle / Paddle-Lite

21 2月, 2020 1 次提交
- X
  fix: move quant op to basic (#2951) · c4111f5a
  由 xiaogang 提交于 2月 21, 2020
```
fix: move quant op to basic
```
  c4111f5a
14 2月, 2020 1 次提交
- X
  
  fix: move some necessary op from extra to basic, (#2872) · 9ae56306
  由 xiaogang 提交于 2月 14, 2020
  
  9ae56306
27 12月, 2019 1 次提交
- Y
  
  move flatten op from extra to basic, test=develop (#2659) · d47f309a
  由 yiicy 提交于 12月 27, 2019
  
  d47f309a
23 12月, 2019 2 次提交
- W
  add sequence_pool_concat fuse and kernel test=develop (#2645) · 04ab34b6
  由 Wilber 提交于 12月 23, 2019
```
add sequence_pool_concat fuse pass

add fuse kernel
```
  04ab34b6
- Y
  
  [ARM] add grid_sampler op and ut, test=develop (#2598) · 9875843e
  由 yiicy 提交于 12月 23, 2019
  
  9875843e
18 12月, 2019 1 次提交
- J
  Support Mask RCNN2 (#2588) · a39f92ea
  由 juncaipeng 提交于 12月 18, 2019
```
* Support Mask RCNN2 (#2588)
```
  a39f92ea
13 12月, 2019 1 次提交
- H
  [LITE][NPU][XPU] Refine subgraph pass, and support NPU/XPU model generation at... · 1dbcd51d
  由 hong19860320 提交于 12月 13, 2019
```
[LITE][NPU][XPU] Refine subgraph pass, and support NPU/XPU model generation at execution time (#2576)
```
  1dbcd51d
10 12月, 2019 1 次提交
- Y
  
  [ARM] add instance norm op and ut, test=develop (#2578) · 933a3724
  由 yiicy 提交于 12月 10, 2019
  
  933a3724
07 12月, 2019 1 次提交

由 juncaipeng 提交于 12月 07, 2019

* add arm split lod tensor, test=develop

* add arm merge lod tensor, test=develop

* update split merge lod tensor, test=develop

* add reduce_prob op, test=develop

* support mask_rcnn succeed, test=develop

cf31b835

26 11月, 2019 1 次提交
- H
  
  add graph_op into basic type test=develop (#2504) · 69ebd957
  由 huzhiqiang 提交于 11月 26, 2019
  
  69ebd957
25 11月, 2019 1 次提交

arrange ops&kernels to support basic models (#2482) · eff81539

由 huzhiqiang 提交于 11月 25, 2019

* arrange ops&kernels to support basic models
* add relu and iocopy into basic
* add fusion_elementwise_activation_ops into basic
* add dropout,layout,io_copyonce kernel&op into basic

eff81539

19 11月, 2019 3 次提交
- Z
  add search_seq_softmax op; regist search_seq_softmax x86 kernel and cuda kernel (#2445) · d7aeb376
  由 zhupengyang 提交于 11月 19, 2019
```
test=develop
```
  d7aeb376
- H
  add x86 kernels: search_fc and sequence_topk_ave_pooling (#2443) · 7c6a9495
  由 huzhiqiang 提交于 11月 18, 2019
```
* add x86 op and kernel : search_fc and sequence_topk_avg_pooling   for content-dnn model test=develop
```
  7c6a9495
- Z
  [X86][CUDA] add attention_padding_mask op, x86 kernel, cuda kernel and unit tests (#2437) · 4048c261
  由 zhupengyang 提交于 11月 19, 2019
```
* [X86] add attention_padding_mask op, x86 kernel and unit test

test=develop

* [CUDA] add attention_padding_mask cuda kernel and unit test

test=develop
```
  4048c261
18 11月, 2019 2 次提交
- P
  add search_group_padding op and x86 kernel, test=develop (#2440) · 33d0cbcd
  由 Pei Yang 提交于 11月 18, 2019
```
add search_group_padding op and x86 kernel
```
  33d0cbcd
- Z
  [X86][CUDA] add sequence_arithmetic op , x86 kernel, cuda kernel and unit test (#2436) · 73296cb6
  由 zhupengyang 提交于 11月 18, 2019
```
* [X86][CUDA] add sequence_arithmetic op , x86 kernel, cuda kernel and unit test

test=develop

* add sequence_arithmetic cuda kernel unit test

test=develop
```
  73296cb6
16 11月, 2019 1 次提交
- H
  
  [LITE][X86] Add search_aligned_mat_mul and search_seq_fc op for X86 (#2428) · 2148bf49
  由 hong19860320 提交于 11月 16, 2019
  
  2148bf49
15 11月, 2019 1 次提交

Add content-dnn ops (#2429) · 7f408ee8

由 juncaipeng 提交于 11月 15, 2019

* add search_seq_depadding x86 and cuda
* add match_matrix_tensor x86
* add search_grnn x86, no test

7f408ee8

14 11月, 2019 2 次提交
- W
  add var_conv_2d op, x86 kernel and unit test test=develop (#2422) · 74f0c8cc
  由 Wilber 提交于 11月 14, 2019
```
- add var_conv_2d op

- add var_conv_2d x86 kernel

- add var_conv_2d x86 test
```
  74f0c8cc
- P
  Update lookup_table op on arm x86，add lookup_table_v2_op (#2405) · 24029728
  由 Pei Yang 提交于 11月 14, 2019
```
* update lookup_table arm x86, test=develop

* add lookup_table_v2_op for compatibility, test=develop
```
  24029728
13 11月, 2019 2 次提交

W
add sequence_reverse op and kerenl for arm and cuda test=develop (#2397) · 3132ad03
由 Wilber 提交于 11月 13, 2019
```
- add sequence_reverse op

- add sequence_reverse kernel for x86 and cuda

- add sequence_reverse_test for x86 and cuda
```
3132ad03

add sequence_concat op kernel and test test=develop (#2414) · 46bb9703

由 Wilber 提交于 11月 13, 2019

- add sequence_concat op

- add sequence_concat kernel for x86 and cuda

- add sequence_concat_test for x86 and cuda

46bb9703

08 11月, 2019 1 次提交

Move muliti class kernel back to basic (#2396) · 7ea34b1b

由 huzhiqiang 提交于 11月 07, 2019

* move multiclass_nms kernel back to host test=develop

* move layer_norm OP and arm_kernel into extra type since it's added after release/v2.0-beta1 and not related with CV test=develop

* fix code_style test=develop

7ea34b1b

07 11月, 2019 1 次提交

check arm kernels type to make sure all_library_links work normally (#2386) · 6b38eab8

由 huzhiqiang 提交于 11月 06, 2019

We have changed 11 arm_kernels into extra type in #2347 , which has caused test_compiling failure. In this PR , we move their 11 related arm_kernel_test into build_extra=ON

6b38eab8

06 11月, 2019 1 次提交
- J
  add channel_wise_dequantized_max_abs op and ChannelWiseDequantOpFuser (#2368) · 41559865
  由 juncaipeng 提交于 11月 06, 2019
```
* add channel_wise_dequantized_max_abs op and ChannelWiseDequantOpFuser, test=develop
```
  41559865
05 11月, 2019 1 次提交
- L
  fix StepRNN model run related bugs (#2300) · 99d4f70e
  由 lijianshe02 提交于 11月 05, 2019
```
* fix step rnn model run bugs test=develop
```
  99d4f70e
04 11月, 2019 1 次提交
- H
  Move new op kernel into extra (#2348) · a438d0dc
  由 huzhiqiang 提交于 11月 04, 2019
```
* move some basic ops into extra type to reduce library size test=develop (#2347)
```
  a438d0dc
22 10月, 2019 2 次提交

Optimize quant_dequant (#2215) · aefb4ea3

由 juncaipeng 提交于 10月 22, 2019

* Add DeleteQuantOpFuser
* Add fake_quantize_dequantize_moving_avg_abs_max_op
* Add DeleteQuantDequantOpFuser

aefb4ea3

Transformer pr (#2214) · 330644b0

由 TianXiaogang 提交于 10月 22, 2019

* feat: add beam_search_special function for support nlp model

* fix: add beam_search_compute kernel input and output

* feat: add assign op & copy_compute kernel

* feat: add fill_const_batch_size_like op & kernel

* feat: add layer_norm op and kernel and ut

* fix: fix some bugs
    fix mul_op infer_shape bug when x_dim_idx = 2, x_dims.size()=3 & y_dim_idx = 1, y_dims.size()=2
    fix elementwise_compute bug when y axis is all 1
    fix beam_search choose math_func wrong bug
    fix layer_norm get attr bug
    fix fill_constant_batch_size_like shape_set bug

* feat: add gather op and kernel & and transform ut

* feats: add ops and fix bugs to support transformer op
       fix type_cast passes to skip `while`
       fix elementwise infer_shape bug when x.dims=3 and y.dims={1} & axis=0
       fix lookup_table compute bug
       fix read_from_array/beam_search/increment/compate/gather ops data_type problems

* fix:
    transfomer ut add word read inferface
    fix copy/gather/norm/layer_norm include path problem

* fix:debug info

* fix: fix input reshape bug

* fix: fix norm bug

* style: style fix & test=develop

* style: fix operators cmakelist

* style: fix operators cmakelist; test=develop

* fix and test=develop

* fix and test=develop

* style: style fix; test=develop

330644b0

11 10月, 2019 1 次提交

CUDA: can run yolov3 int8 (#2172) · 29f448c6

由 Zhaolong Xing 提交于 10月 11, 2019

* add conv int8 support(in condition which the input or output channel not be the times of 4)
add add_kernel for cuda.

* can run yolov3 fp32
test=develop

* 1. fix bug with yolov3 run
test=develop

* can run yolov3 int8 test=develop

29f448c6

17 9月, 2019 1 次提交
- L
  add fill_constant_batch_size_like op and add its unittest (#2044) · b91850b9
  由 liu zhengxi 提交于 9月 17, 2019
```
* add fill_constant_batch_size_like op and add its unittest
```
  b91850b9
16 9月, 2019 1 次提交
- L
  Gru op (#2002) · 1cb36af6
  由 lhl960107 提交于 9月 16, 2019
```
* add x86 gru&&relu&&sequence_expand_as op test=develop
```
  1cb36af6
12 9月, 2019 1 次提交

add unsqueeze and range op (x2paddle) (#1988) · cca0aec6

由 Wilber 提交于 9月 12, 2019

* add unsqueeze and range op. modify concat op test=develop

* modify exception in range_test_x86

cca0aec6

09 9月, 2019 1 次提交
- J
  add assign_value and hard_sigmoid, add fluid_type (#1983) · 9796c57d
  由 juncaipeng 提交于 9月 09, 2019
```
* add assign_value op, arm kernel and test, add fluid_type, test=develop

* add hard_sigmoid, test=develop
```
  9796c57d
07 9月, 2019 1 次提交

add lite x86 ops for ASR test=develop (#1981) · 25b775d6

由 lijianshe02 提交于 9月 07, 2019

* add lite x86 ops for ASR test=develop

* add lite x86 ops for ASR test=develop

* fix x86 ci run test problems test=develop

* fix mkl path for CI test=develop

25b775d6

02 9月, 2019 1 次提交

Add ops and fix bugs for Faster RCNN (#1942) · cfd5abe5

由 juncaipeng 提交于 9月 02, 2019

* add ops for faster rcnn

* disable test for generate_proposals and roi_align, test=develop

* remove .swp file

* remove log in tensor slice

* finish the unit test for roi_align, test=develop

* add box_clip op and fix tensor slice bug

* remove add four op twice

* rewrite the implement for box_coder and sequence_expand, add faster_rcnn_test, test=develop

* fix test bug of box_clip in x86 server, test=develop

cfd5abe5

01 9月, 2019 1 次提交

[ARM][CPU] Fix time counter of arm cpu profiler (#1925) · c096de0e

由 Yuan Shuai 提交于 9月 01, 2019

* Fix timer of arm cpu profiler. test=develop

* Fix un-added op in cmake.test=develop

* fix cmake error

* fix cmake error, test=develop

* Fix pass sequence. test=develop

* replace option with lite_option. test=develop

* disable profile mode by default. test=develop

* Fix error option name. test=develop

c096de0e

29 8月, 2019 2 次提交

L

add stack op and add reduce_mean op and their unit tests (#1888) · 8ccd01a6
由 liu zhengxi 提交于 8月 29, 2019

8ccd01a6

ad ops for faster rcnn, including affine_channel, anchor_generator,... · f3035827

由 juncaipeng 提交于 8月 29, 2019

ad ops for faster rcnn, including affine_channel, anchor_generator, generate_proposals and roi_align (#1895)

* add ops for faster rcnn

* disable test for generate_proposals and roi_align, test=develop

* remove .swp file

* remove log in tensor slice

* finish the unit test for roi_align, test=develop

f3035827

28 8月, 2019 1 次提交
- J
  Modify cast op and remove warning in argmax_test (#1894) · 5fe41d5c
  由 juncaipeng 提交于 8月 28, 2019
```
* modify cast op, test=develop

* modify cast op and remove warning in argmax_test, test=develop
```
  5fe41d5c