- 14 2月, 2019 1 次提交
-
-
由 Yancey1989 提交于
-
- 31 1月, 2019 2 次提交
-
-
由 liuwei1031 提交于
* expose peak gpu memory API to python test=develop * add unittest for peak gpu memory monitoring test=develop * add pybind change test=develop * add mutex to gpu mem usage monitor test=develop * update benchmark flag definition file test=develop * tweak unittest for memory monitoring test=develop
-
由 Yan Chunwei 提交于
-
- 29 1月, 2019 1 次提交
-
-
由 Krzysztof Binias 提交于
test=develop
-
- 28 1月, 2019 1 次提交
-
-
由 Yan Chunwei 提交于
-
- 26 1月, 2019 3 次提交
-
-
由 gongweibao 提交于
-
由 baojun-nervana 提交于
-
由 baojun-nervana 提交于
-
- 25 1月, 2019 6 次提交
-
-
由 JiabinYang 提交于
-
由 JiabinYang 提交于
-
由 gongweibao 提交于
-
由 JiabinYang 提交于
-
由 JiabinYang 提交于
-
由 baojun 提交于
* enable ngraph_engine_op test=develop * merge develop test=develop * avoid const_cast test=develop * rm ngraph_operator test=develop * Added TODO to move EnableNgraph test=develop * Add TODO to remove const_cast test=develop
-
- 24 1月, 2019 2 次提交
-
-
由 Yiqun Liu 提交于
* Refine the beam_search op and test. * A basic CUDA implementation of beam_search for small batch_size. * Implement CUDA kernel for beam_search_op. * Use multiple CUDA threads in the same block to select the top beam. * Update the python api of beam_search op. * Enable extend function in CPU kernel of beam_search op. * Unify the CUDA codes. test=develop * Unify the CPU kernel of beam_search op. * Ensure the seletced items of beam_search_op's CPU kernel sorted by scores. * Update the description of beam_search in API.spec. * Enable the use of CUDA kernel in beam_search op. * Exclude the beam_search's CUDA unittest when there is no CUDA gpu, and delete some debuging statements. test=develop * Follow comments. test=develop * Call the CPU kernel for beam_search op when batch_size > 4. test=develop * Remove the except of is_empty op in PrepareData. test=develop
-
由 nhzlx 提交于
1. graph and program_desc alignment 2. trt stream test=develop
-
- 22 1月, 2019 1 次提交
-
-
由 sneaxiy 提交于
test=develop
-
- 21 1月, 2019 2 次提交
-
-
由 Yan Chunwei 提交于
-
由 Dun 提交于
* mem opt * test=develop * test=develop * test=develop * test=develop * test=develop * test=develop * test=develop * refine code test=develop * refine code test=develop * refine code test=develop * refine code test=develop * refine with cub test=develop * fix mkldnn test && remove comments && test=develop * polish code && test=develop * add only_forward test && test=develop
-
- 20 1月, 2019 1 次提交
-
-
由 WangZhen 提交于
-
- 19 1月, 2019 1 次提交
-
-
由 WangZhen 提交于
-
- 17 1月, 2019 1 次提交
-
-
由 gongweibao 提交于
-
- 15 1月, 2019 1 次提交
-
-
由 mozga-intel 提交于
test=develop
-
- 14 1月, 2019 5 次提交
-
-
由 peizhilin 提交于
-
由 peizhilin 提交于
test=develop
-
由 tensor-tang 提交于
test=develop
-
由 tensor-tang 提交于
-
- 13 1月, 2019 2 次提交
-
-
由 tensor-tang 提交于
test=develop
-
由 tensor-tang 提交于
-
- 12 1月, 2019 2 次提交
- 11 1月, 2019 3 次提交
-
-
由 Zhaolong Xing 提交于
-
由 chengduozh 提交于
test=develop This reverts commit 064512aa.
-
由 chengduo 提交于
* remove workspace_handle in conv2d_cudnn test=develop * remove workspace_handle test=develop * fix bug test=develop * make test_conv2d_op SERIAL test=develop * save memory in conv_cudnn test=develop * enhance thread safety test=develop * enhance temporary allocator test=develop * Add excess fraction test=develop * follow comments test=develop * fix bug and code refine test=develop * fix memory size check test=develop * rename reuse_tmp_allocation_excess_fraction test=develop
-
- 10 1月, 2019 5 次提交
-
-
由 tensor-tang 提交于
test=develop
-
由 tensor-tang 提交于
test=develop
-
由 flame 提交于
-
由 tensor-tang 提交于
test=develop
-
由 tensor-tang 提交于
test=develop
-