1. 30 11月, 2021 1 次提交
  2. 26 5月, 2021 1 次提交
    • T
      ut fix (#33102) · e05a7a49
      tangwei12 提交于
      
      Change-Id: I2e82dfcee6a1d0512b94cebc32281123fa5bf597
      
      * pretty print for datafeed error
      
      Change-Id: I056a8b6f03608e96679a83846c97aed289cef7e6
      
      * fix fleet dist infer ut
      e05a7a49
  3. 30 12月, 2020 1 次提交
    • T
      fix ut (#29989) · ed856d25
      tangwei12 提交于
      * fix ut
      
      Change-Id: I151e152919a1863db07792bffb42d0ca68995756
      ed856d25
  4. 24 12月, 2020 1 次提交
  5. 08 9月, 2020 1 次提交
  6. 02 9月, 2020 1 次提交
  7. 20 8月, 2020 1 次提交
  8. 19 8月, 2020 1 次提交
  9. 07 8月, 2020 1 次提交
  10. 30 7月, 2020 1 次提交
  11. 17 2月, 2020 1 次提交
  12. 12 2月, 2020 1 次提交
  13. 17 1月, 2020 1 次提交
  14. 06 1月, 2020 1 次提交
  15. 19 12月, 2019 1 次提交
  16. 15 10月, 2019 1 次提交
    • C
      Fix communicator slow bug & fix communicator stop bug (#20366) · 940c6ff1
      Chengmo 提交于
      * test=develop,Fix communicator slow bug
      
      * test=develop, delete if() in stop_worker()
      
      * test=develop
      
      * fix UT, test=develop
      
      * fix bug in fetch handler, test=develop
      
      * fix bug in fetch handler, test=develop
      
      * test=develop, fix fetch barrier bug
      
      * test=develop, bug fix
      
      * test=develop, bug fix
      
      * test=develop, fix bug
      940c6ff1
  17. 07 10月, 2019 1 次提交
  18. 27 9月, 2019 1 次提交
  19. 28 8月, 2019 1 次提交
    • T
      Fix the correctness of async mode at distributed training (#18863) · 65c73684
      tangwei12 提交于
      * fix correctness of the communicator
      
      * fix a bug in send thread when sending var context is empty, test=develop
      
      * add lookup_table_prefetch_op and prefetch optimize, test=develop
      
      * remove remote prefetch GPU supported
      
      * word2vec force with CPU, test=develop
      
      * test dist remote lookup table force with CPU, test=develop
      65c73684
  20. 22 7月, 2019 1 次提交
  21. 12 6月, 2019 1 次提交
  22. 29 10月, 2018 1 次提交
    • W
      [1.1] [project] train imagenet using large batch size (#13766) · 26200f2e
      Wu Yi 提交于
      * fix nccl2 lars dist support
      
      * put lars in momentum op
      
      * add tests lars
      
      * fix ci
      
      * fix cpu kernel
      
      * soft warning
      
      * remove lars in test_recognize_digits.py
      
      * move to another op
      
      * add file
      
      * update api.spec test=develop
      
      * update test=develop
      
      * fix api.spec test=develop
      
      * wip
      
      * wip, finish grad merge ops
      
      * wip, finish graph build
      
      * wip test running
      
      * work on 1 gpu
      
      * workable version
      
      * update
      
      * fix tests
      
      * fuse broadcast op
      
      * fix compile failed
      
      * refine
      
      * add batch merge test mnist
      
      * fix CI test=develop
      
      * fix build
      
      * use independent bn params for batch merge test=develop
      
      * update api.spec
      
      * follow comments and for test
      
      * wip
      
      * refine tests test=develop
      
      * follow comments test=develop
      
      * remove startup bn modify test=develop
      
      * follow comments test=develop
      
      * fix merge test=develop
      26200f2e