- 10 5月, 2019 1 次提交
- 
- 
由 zhaoyuchen2018 提交于refine code fuse cublas calling and kernels into one cuda kernel. test=develop Signed-off-by: Nzhaoyuchen <zhaoyuchen01@baidu.com>
 
- 
- 07 5月, 2019 1 次提交
- 
- 
由 Kaipeng Deng 提交于* add attr axis infershape. test=develop * add CUDA kernel. test=develop * fix unittest. test=develop * fix unittest for soft_label. test=develop * fix fp16 unittest. test=develop * remove comment code. test=develop * refine test for axis. test=develop * add python api. test=develop * fix doc. test=develop * fix fp16 unittest. test=develop * fix ngraph test. test=develop * fix ENFORCE for test_imperative_transformer. test=develop * fit for ngraph test. test=develop * fix after rebase develop. test=develop * fix doc. test=develop * fix API.spec. test=develop * fix test_layers. test=develop * fix format. test=develop 
 
- 
- 20 4月, 2019 1 次提交
- 
- 
由 Yibing Liu 提交于* Support seq len equal to 0 in sequence ops test=develop * Add more test cases * Fix some comments test=develop * Fix py3 error test=develop 
 
- 
- 17 4月, 2019 1 次提交
- 
- 
由 Kevin 提交于* fix overflow by int32 mul test=develop * fix reference nullptr * fix codestyle test=develop * modify to point in ContextProjectFunctor test=develop * modify to point in ContextProjectFunctor test=develop * modify . to -> test=develop 
 
- 
- 12 4月, 2019 3 次提交
- 
- 
由 Qiao Longfei 提交于
- 
由 Qiao Longfei 提交于
- 
由 Qiao Longfei 提交于
 
- 
- 25 3月, 2019 1 次提交
- 
- 
由 dengkaipeng 提交于
 
- 
- 20 3月, 2019 2 次提交
- 
- 
由 phlrain 提交于
- 
由 dengkaipeng 提交于
 
- 
- 18 3月, 2019 2 次提交
- 
- 
由 dengkaipeng 提交于
- 
由 phlrain 提交于
 
- 
- 14 3月, 2019 2 次提交
- 
- 
由 sneaxiy 提交于test=develop 
- 
由 Zeng Jinle 提交于test=develop 
 
- 
- 12 3月, 2019 1 次提交
- 
- 
由 sneaxiy 提交于test=develop 
 
- 
- 08 3月, 2019 3 次提交
- 
- 
由 tensor-tang 提交于test=develop 
- 
由 Yiqun Liu 提交于Make parent_idx a dispensable output for beam_search op to support models saved by older paddle version. (#16106) test=develop 
- 
由 Yiqun Liu 提交于Make parent_idx a dispensable output for beam_search op to support models saved by older paddle version. (#16106) test=develop 
 
- 
- 07 3月, 2019 1 次提交
- 
- 
由 tensor-tang 提交于test=develop 
 
- 
- 04 3月, 2019 3 次提交
- 
- 
由 Yiqun Liu 提交于test=develop 
- 
由 Yihua Xu 提交于test=develop 
- 
由 Qiao Longfei 提交于
 
- 
- 28 2月, 2019 1 次提交
- 
- 
由 Yiqun Liu 提交于test=develop 
 
- 
- 26 2月, 2019 1 次提交
- 
- 
由 Yihua Xu 提交于test=develop 
 
- 
- 22 2月, 2019 2 次提交
- 
- 
由 tensor-tang 提交于* Revert "Optimze Gelu with MKL Erf function (#15770)" This reverts commit 676995c8. * test=develop 
- 
由 Yihua Xu 提交于* Optimize for gelu operator * Set up the low accuracy mode of MKL ERF function. test=develop * Only enable MKLML ERF when OS is linux * Use the speical mklml version included vmsErf function to verify gelu mkl kernel. test=develop * Add the CUDA macro to avoid NVCC's compile issue. test=develop * Add the TODO comments for mklml library modification. test=develop * Clean Code test=develop * Add the comment of marco for NVCC compiler. test=develop 
 
- 
- 19 2月, 2019 1 次提交
- 
- 
由 xuezhong 提交于test=develop 
 
- 
- 11 2月, 2019 1 次提交
- 
- 
由 xuezhong 提交于test=develop 
 
- 
- 02 2月, 2019 1 次提交
- 
- 
由 peizhilin 提交于test=develop 
 
- 
- 30 1月, 2019 5 次提交
- 29 1月, 2019 3 次提交
- 
- 
由 tensor-tang 提交于test=develop 
- 
由 tensor-tang 提交于test=develop 
- 
由 tensor-tang 提交于test=develop 
 
- 
- 24 1月, 2019 2 次提交
- 
- 
由 Yiqun Liu 提交于* Refine the beam_search op and test. * A basic CUDA implementation of beam_search for small batch_size. * Implement CUDA kernel for beam_search_op. * Use multiple CUDA threads in the same block to select the top beam. * Update the python api of beam_search op. * Enable extend function in CPU kernel of beam_search op. * Unify the CUDA codes. test=develop * Unify the CPU kernel of beam_search op. * Ensure the seletced items of beam_search_op's CPU kernel sorted by scores. * Update the description of beam_search in API.spec. * Enable the use of CUDA kernel in beam_search op. * Exclude the beam_search's CUDA unittest when there is no CUDA gpu, and delete some debuging statements. test=develop * Follow comments. test=develop * Call the CPU kernel for beam_search op when batch_size > 4. test=develop * Remove the except of is_empty op in PrepareData. test=develop 
- 
由 tangwei12 提交于* nce add check sample lables, test=develop 
 
- 
- 21 1月, 2019 1 次提交
- 
- 
由 Dun 提交于* mem opt * test=develop * test=develop * test=develop * test=develop * test=develop * test=develop * test=develop * refine code test=develop * refine code test=develop * refine code test=develop * refine code test=develop * refine with cub test=develop * fix mkldnn test && remove comments && test=develop * polish code && test=develop * add only_forward test && test=develop 
 
- 
