1. 24 1月, 2019 1 次提交
    • Y
      Add the CUDA kernel for beam_search op (#15020) · 3008fa12
      Yiqun Liu 提交于
      * Refine the beam_search op and test.
      
      * A basic CUDA implementation of beam_search for small batch_size.
      
      * Implement CUDA kernel for beam_search_op.
      
      * Use multiple CUDA threads in the same block to select the top beam.
      
      * Update the python api of beam_search op.
      
      * Enable extend function in CPU kernel of beam_search op.
      
      * Unify the CUDA codes.
      test=develop
      
      * Unify the CPU kernel of beam_search op.
      
      * Ensure the seletced items of beam_search_op's CPU kernel sorted by scores.
      
      * Update the description of beam_search in API.spec.
      
      * Enable the use of CUDA kernel in beam_search op.
      
      * Exclude the beam_search's CUDA unittest when there is no CUDA gpu, and delete some debuging statements.
      test=develop
      
      * Follow comments.
      test=develop
      
      * Call the CPU kernel for beam_search op when batch_size > 4.
      test=develop
      
      * Remove the except of is_empty op in PrepareData.
      test=develop
      3008fa12
  2. 22 1月, 2019 1 次提交
  3. 21 1月, 2019 2 次提交
    • Y
      fea/infer memory optim2 (#14953) · 885c4e57
      Yan Chunwei 提交于
      885c4e57
    • D
      Memory optimization of depthwise conv op and group norm op (#15313) · 9f8f0fc2
      Dun 提交于
      * mem opt
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * test=develop
      
      * refine code  test=develop
      
      * refine code  test=develop
      
      * refine code  test=develop
      
      * refine code  test=develop
      
      * refine with cub test=develop
      
      * fix mkldnn test && remove comments && test=develop
      
      * polish code && test=develop
      
      * add only_forward test && test=develop
      9f8f0fc2
  4. 20 1月, 2019 1 次提交
  5. 19 1月, 2019 1 次提交
  6. 17 1月, 2019 1 次提交
  7. 15 1月, 2019 1 次提交
  8. 14 1月, 2019 5 次提交
  9. 13 1月, 2019 2 次提交
  10. 12 1月, 2019 2 次提交
    • X
      fix · 50b4ac08
      Xin Pan 提交于
      test=develop
      50b4ac08
    • X
      try fix py2 · a1bfb35d
      Xin Pan 提交于
      test=develop
      a1bfb35d
  11. 11 1月, 2019 3 次提交
  12. 10 1月, 2019 5 次提交
  13. 09 1月, 2019 1 次提交
  14. 08 1月, 2019 5 次提交
  15. 07 1月, 2019 9 次提交