1. 28 1月, 2022 1 次提交
    • F
      [PSLIB] Add Metrics Module, Support User-defined Add Metric (#38230) · 7460a891
      Fan Zhang 提交于
      * 12.3 first add metrics module
      
      * add Mask/MultiTask
      
      * add WuAUC
      
      * [PSLIB] Update WuAUC Compute
      
      * [PSLIB] Change WuAUC Compute Mehod
      
      * [PSLIB] Clean WuAUC Compute
      
      * [PSLIB] Clean Metric Module Unused Code
      
      * mv metric instance
      
      * [PSLIB] Add Metrics Module, Support User-defined Add Metric (#38789)
      
      * [PSLIB] Add Metrics Module, Support User-defined Add Metric
      
      * [PSLIB] Modify According to CI
      
      * [PSLIB] Modify According to CI
      
      * [PSLIB] Modify According to CI
      
      * [PSLIB] Modify According to CI Coverage
      
      * [PSLIB] Modify According to CI
      
      * [PSLIB] Modify According to CI
      
      * [PSLIB] Modify According to CI
      
      * [PSLIB] Modify According to CI
      
      * [PSLIB] Modify According to CI
      
      * [PSLIB] Modify According to CI Coverage
      
      * [PSLIB] Modify According to CI Coverage
      
      * [PSLIB] Modify According to CI Coverage
      
      * modify role_maker
      
      * update CMakeLists.txt
      7460a891
  2. 26 10月, 2021 1 次提交
    • F
      [CPU-PSLIB] Fix bug for consistency insepection of op's embedding name and... · 15cb05c8
      Fan Zhang 提交于
      [CPU-PSLIB] Fix bug for consistency insepection of op's embedding name and sparse table name in config_fleet.py (#36215)
      
      * [CPU-PSLIB] Add consistency insepection of use_var_list and data_generator data
      
      * [CPU-PSLIB] Fix bug for consistency insepection of op's embedding name and sparse table name in config_fleet.py
      15cb05c8
  3. 07 9月, 2021 2 次提交
  4. 17 8月, 2021 1 次提交
  5. 12 8月, 2021 1 次提交
  6. 26 7月, 2021 1 次提交
  7. 18 5月, 2021 2 次提交
  8. 30 3月, 2021 1 次提交
    • S
      Scale 1.8 (#31940) · c36c22fe
      Shang Zhizhou 提交于
      * add n-d input support for trt scale converter (#31316)
      
      * add n-d input support for trt scale converter
      
      * add flatten for ut
      
      * fix dims
      
      * fix batchnorm when inpu dims < 3 (#31933)
      
      * fix batchnorm when inpu dims < 3
      
      * add unittest for batchnorm dims = 2
      
      * fix unittest
      Co-authored-by: NPei Yang <peiyang@baidu.com>
      c36c22fe
  9. 24 3月, 2021 1 次提交
  10. 07 12月, 2020 1 次提交
    • S
      cherry-pick PR #27933 (#29377) · 9a6ecb03
      Shang Zhizhou 提交于
      * cherry-pick PR #27933
      
      * fix: cuda version is in varibale CUDA_VERSION in 1.8 cuda.cmake
      
      * close unittest failed temporarily
      
      * cherry-pick PR #27544, fix layer_norm and softmax bug in tensorRT
      9a6ecb03
  11. 01 12月, 2020 1 次提交
  12. 13 11月, 2020 1 次提交
    • S
      Skip layernorm to 1.8 (#28583) · ec672e88
      Shang Zhizhou 提交于
      * 裁剪transformer模型trt支持;修复tensorRT不支持DeletePass的bug (#28517)
      
      * skip_layernorm_op done
      
      * add unittest
      
      * slice op convertor support trt < 6
      
      * skip_layernorm only work in ernie
      
      * fix unittest
      
      * fix unittest
      ec672e88
  13. 09 11月, 2020 1 次提交
  14. 05 11月, 2020 1 次提交
    • S
      Ernie varlen to 1.8 (#28400) · 78d68d59
      Shang Zhizhou 提交于
      * Fix TRT plugin registry without TRT lib (#25982)
      
      * fix trt plugin registry without trt lib
      
      * support trt4
      
      * refine code style
      
      * pick ea851796 from develop
      
      * cherry-pick develop PR  #26273 && #27796
      
      * fix unittest error
      
      * fix unittest error
      
      * remove const_cast
      Co-authored-by: NPei Yang <peiyang@baidu.com>
      78d68d59
  15. 13 10月, 2020 1 次提交
  16. 12 10月, 2020 1 次提交
  17. 10 10月, 2020 2 次提交
  18. 28 9月, 2020 1 次提交
  19. 27 9月, 2020 2 次提交
  20. 23 9月, 2020 1 次提交
    • P
      Optimize slice trt plugin (#26970) (#27456) · 8e1712a7
      Pei Yang 提交于
      * optimize slice TRT plugin
      
      This patch removes unnecessary barrier for data transfer of needed offset,
      so data transfer can be overlap with GPU kernel execution.
      
      This patch also fixes incorrect name of slice plugin. That is, replaces
      "layernorm" with "slice"
      
      test=develop
      
      * add serialize/deserialize to slice plugin
      
      * add static shape slice trt plugin
      
      * fix slice trt op convertor dynamic shape bug
      
      * fix format by clang-format
      
      * fix pylint format error
      
      * fix problems commented by peiyang
      Co-authored-by: NRyan Jeng <rjeng@nvidia.com>
      Co-authored-by: NShang Zhizhou <shangzhizhou@baidu.com>
      Co-authored-by: NRyan Jeng <rjeng@nvidia.com>
      8e1712a7
  21. 22 9月, 2020 2 次提交
  22. 17 9月, 2020 1 次提交
  23. 15 9月, 2020 1 次提交
  24. 11 9月, 2020 1 次提交
  25. 10 9月, 2020 1 次提交
  26. 08 9月, 2020 1 次提交
  27. 03 9月, 2020 1 次提交
  28. 27 8月, 2020 1 次提交
  29. 21 8月, 2020 1 次提交
  30. 20 8月, 2020 1 次提交
  31. 19 8月, 2020 1 次提交
  32. 18 8月, 2020 1 次提交
  33. 17 8月, 2020 3 次提交