1. 12 2月, 2019 1 次提交
  2. 24 1月, 2019 2 次提交
    • Y
      Add the CUDA kernel for beam_search op (#15020) · 3008fa12
      Yiqun Liu 提交于
      * Refine the beam_search op and test.
      
      * A basic CUDA implementation of beam_search for small batch_size.
      
      * Implement CUDA kernel for beam_search_op.
      
      * Use multiple CUDA threads in the same block to select the top beam.
      
      * Update the python api of beam_search op.
      
      * Enable extend function in CPU kernel of beam_search op.
      
      * Unify the CUDA codes.
      test=develop
      
      * Unify the CPU kernel of beam_search op.
      
      * Ensure the seletced items of beam_search_op's CPU kernel sorted by scores.
      
      * Update the description of beam_search in API.spec.
      
      * Enable the use of CUDA kernel in beam_search op.
      
      * Exclude the beam_search's CUDA unittest when there is no CUDA gpu, and delete some debuging statements.
      test=develop
      
      * Follow comments.
      test=develop
      
      * Call the CPU kernel for beam_search op when batch_size > 4.
      test=develop
      
      * Remove the except of is_empty op in PrepareData.
      test=develop
      3008fa12
    • T
      nce add check sample lables, test=develop (#15463) · 5cfc40de
      tangwei12 提交于
      * nce add check sample lables, test=develop
      5cfc40de
  3. 23 1月, 2019 1 次提交
  4. 22 1月, 2019 1 次提交
  5. 21 1月, 2019 1 次提交
    • D
      Memory optimization of depthwise conv op and group norm op (#15313) · 9f8f0fc2
      Dun 提交于
      * mem opt
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * refine code  test=develop
      
      * refine code  test=develop
      
      * refine code  test=develop
      
      * refine code  test=develop
      
      * refine with cub test=develop
      
      * fix mkldnn test && remove comments && test=develop
      
      * polish code && test=develop
      
      * add only_forward test && test=develop
      9f8f0fc2
  6. 18 1月, 2019 1 次提交
    • Z
      Tree conv op (#15217) · e2ba9668
      zhaozhehao 提交于
      * refactor tree2col operator with new memory mechanism test=develop
      
      * test=develop
      
      * test=develop
      
      * Modified API according to panyx0718 test=develop
      
      * fix API change according to heavengate test=develop
      
      * Modify API comment test=develop
      e2ba9668
  7. 16 1月, 2019 1 次提交
  8. 15 1月, 2019 1 次提交
  9. 14 1月, 2019 1 次提交
  10. 13 1月, 2019 2 次提交
  11. 09 1月, 2019 1 次提交
  12. 04 1月, 2019 1 次提交
  13. 29 12月, 2018 1 次提交
  14. 28 12月, 2018 1 次提交
  15. 27 12月, 2018 6 次提交
  16. 26 12月, 2018 6 次提交
    • Q
      enable unit test for test_nce · 0384f330
      Qiao Longfei 提交于
      test=develop
      0384f330
    • Q
      fix · 031995cf
      Qiao Longfei 提交于
      031995cf
    • Q
      add init once for assign layer · b53eb7dc
      Qiao Longfei 提交于
      b53eb7dc
    • W
      Make topk op support variable k. (#15044) · 2314f2eb
      whs 提交于
      * Make topk op support variable k.
      test=develop
      
      * Fix tensor type.
      test=develop
      2314f2eb
    • W
      Fp16 training (#14992) · 856f0da0
      Wu Yi 提交于
      * wip
      
      * wip
      
      * wip
      
      * wip for test
      
      * add fp16 tests test=develop
      
      * fix cpu build test=develop
      
      * fix test=develop
      
      * fix py3 tests test=develop
      
      * fix lr_scheduler dtype test=develop
      
      * fix test=dvelop
      
      * test fix ci compile test=develop
      
      * fix build and merge test=develop
      
      * fallback momentumop change to general test=develop
      
      * make fp16 lr schedule simple test=develop
      
      * fix ut test=develop
      
      * fix tests test=develop
      
      * remove fp16 learning rate cast test=develop
      856f0da0
    • Y
      Fix the unstack layer (#15047) · a28df3eb
      Yibing Liu 提交于
      test=develop
      a28df3eb
  17. 25 12月, 2018 1 次提交
  18. 24 12月, 2018 2 次提交
  19. 20 12月, 2018 3 次提交
  20. 19 12月, 2018 4 次提交
  21. 18 12月, 2018 2 次提交