- 28 1月, 2019 6 次提交
-
-
由 Xin Pan 提交于
imperative supports multi grad ops
-
由 Tao Luo 提交于
Performance and functional fixes to LRN
-
由 Haihao Shen 提交于
* Enable mobilenet UT in separate test class; use download cache by paddle download utility and cache unzip; and fix typo; test=develop * Extract cache_unzipping function for reuse; format code style; test=develop * Simplify the test code by define a combined function for both downloading and unzipping; test=develop
-
由 Zhaolong Xing 提交于
add trt int8 support
-
由 Jiabin Yang 提交于
test=develop, refine_error_message for data type
-
由 Kaipeng Deng 提交于
-
- 27 1月, 2019 4 次提交
-
-
由 乔龙飞 Qiao Longfei 提交于
Fix prefetch one parameter
-
由 Jacek Czaja 提交于
test=develop
-
由 Jacek Czaja 提交于
test=develop
-
由 Qiao Longfei 提交于
-
- 26 1月, 2019 6 次提交
-
-
由 gongweibao 提交于
-
由 Tao Luo 提交于
Refine INT8 calibration API
-
由 Tao Luo 提交于
move ngraph_bridge to ngraph directory
-
由 nhzlx 提交于
test=develop
-
由 Yan Chunwei 提交于
-
由 baojun-nervana 提交于
-
- 25 1月, 2019 22 次提交
-
-
由 Qiao Longfei 提交于
-
由 Qiao Longfei 提交于
-
由 ruri 提交于
Add Shuffle Channel Operator
-
由 Jacek Czaja 提交于
Added reading dst mem pd from lrn pd coding style fixes test=develop
-
由 Haihao Shen 提交于
-
由 nhzlx 提交于
test=develop
-
由 tensor-tang 提交于
jit benchmark use tensor with alignment
-
由 Qiao Longfei 提交于
-
由 乔龙飞 Qiao Longfei 提交于
Optimize cpp reader
-
由 gongweibao 提交于
-
由 Qiao Longfei 提交于
-
由 nhzlx 提交于
test=develop
-
由 JiabinYang 提交于
-
由 Qiao Longfei 提交于
-
由 gongweibao 提交于
-
由 tangwei12 提交于
* fix mistakes in merge_ids, test=develop
-
由 Zhaolong Xing 提交于
Add check: conv_fusion op runs with cudnn version > 7100 .
-
由 Xin Pan 提交于
test=develop
-
由 baojun 提交于
* enable ngraph_engine_op test=develop * merge develop test=develop * avoid const_cast test=develop * rm ngraph_operator test=develop * Added TODO to move EnableNgraph test=develop * Add TODO to remove const_cast test=develop
-
由 chengduo 提交于
test=develop
-
由 chengduo 提交于
* Revert "set constant for loss" This reverts commit 167933f678ccbb3563e949710279efe004a27731. * Revert "remove workspace_handle" test=develop This reverts commit b4aca8ede9e685bce1dfb1c59e63919f33432572.
-
由 tensor-tang 提交于
test=develop
-
- 24 1月, 2019 2 次提交
-
-
由 Xin Pan 提交于
test=develop
-
由 Yiqun Liu 提交于
* Refine the beam_search op and test. * A basic CUDA implementation of beam_search for small batch_size. * Implement CUDA kernel for beam_search_op. * Use multiple CUDA threads in the same block to select the top beam. * Update the python api of beam_search op. * Enable extend function in CPU kernel of beam_search op. * Unify the CUDA codes. test=develop * Unify the CPU kernel of beam_search op. * Ensure the seletced items of beam_search_op's CPU kernel sorted by scores. * Update the description of beam_search in API.spec. * Enable the use of CUDA kernel in beam_search op. * Exclude the beam_search's CUDA unittest when there is no CUDA gpu, and delete some debuging statements. test=develop * Follow comments. test=develop * Call the CPU kernel for beam_search op when batch_size > 4. test=develop * Remove the except of is_empty op in PrepareData. test=develop
-