1. 11 12月, 2020 1 次提交
    • Z
      improve dropout (#29465) · 6702040e
      Zhang Ting 提交于
      * improve drop out
      
      * add VectorizedRandomGeneratorWithGenerator
      
      * fix bug
      
      * modify according to comments
      6702040e
  2. 04 9月, 2020 1 次提交
  3. 13 4月, 2020 1 次提交
  4. 10 12月, 2019 1 次提交
  5. 03 9月, 2019 1 次提交
  6. 20 8月, 2019 1 次提交
    • W
      optimize the realization of cuda dropout (#19136) · 6e326ca2
      wangchaochaohu 提交于
      * cuda optimie for dropout
      
      * remove tmp swp file
      
      * fix compile error test=develop
      
      * test=develop optimize the cuda realization of dropout op
      
      * remove unsed code test=develop
      
      * remove tmp file test=develop
      6e326ca2
  7. 28 4月, 2019 1 次提交
    • Z
      Refine dropout gpu memory (#17095) · 28d69d71
      Zeng Jinle 提交于
      * refine_dropout_mem,test=develop
      
      * # This is a combination of 14 commits.
      # The first commit's message is:
      remove ut test_dist_word2vec in mac ci, will fix it in private, test=develop (#17066)
      
      # This is the 2nd commit message:
      
      Fleet unify distributed training (#16791)
      
      * implement distributed transpiler with fleet
      # This is the 3rd commit message:
      
      ParallelDyGraph with GPU collective mode (#16827)
      
      implement dygraph.parallel.DataParallel to hook reduce op.
      
      # This is the 4th commit message:
      
      Init mixed precision training interface (#16856)
      
      * Init mixed precision training interface
      
      * Add fp16 test script
      
      test=develop
      
      * All initializers support float16
      
      test=develop
      
      * Code cleanup & add more code annotations
      
      test=develop
      
      * Update API spec
      
      test=develop
      
      * Add usage example in doc
      
      test=develop
      
      # This is the 5th commit message:
      
      fix reference_count_pass,test=develop (#17060)
      
      test=develop
      # This is the 6th commit message:
      
      Speedup roi_perspective_transform op by caching the information of linear interpolation in forward (#17090)
      
      * Cache the information of linear interpolation in forward and use it in backward.
      test=develop
      
      * Fix cuda kernel.
      test=develop
      
      # This is the 7th commit message:
      
      remove unnecessary prepare_data (#17080)
      
      test=develop
      # This is the 8th commit message:
      
      fix interpolate cu. test=develop (#17101)
      
      # This is the 9th commit message:
      
      test=develop, double backward leaky_relu (#17067)
      
      backward of backward: leaky_relu
      # This is the 10th commit message:
      
      fix fuse optimizer ops (#17102)
      
      test=develop
      # This is the 11th commit message:
      
      truncated_gaussian_random supported in distributed training, test=develop (#17091)
      
      # This is the 12th commit message:
      
       Detailed coordinate description for yolov3 loss (#17007)
      
      * Detailed coordinate description for yolov3 loss
      
      test=develop
      
      * modified api.spec
      
      test=develop
      
      * modified loss name
      
      * fix api.spec
      
      test=develop
      
      * polish description
      
      test=develop
      
      * modified api.spec
      
      test=develop
      
      # This is the 13th commit message:
      
      fix test_weight_decay (#17109)
      
      test=develop
      # This is the 14th commit message:
      
      Path flag (#17105)
      
      * fix python/paddle/fluid/__init__.py detecting problems
      28d69d71
  8. 30 1月, 2019 1 次提交
  9. 11 12月, 2018 1 次提交
  10. 24 10月, 2018 1 次提交
  11. 23 10月, 2018 1 次提交
  12. 20 4月, 2018 1 次提交
  13. 19 4月, 2018 1 次提交
    • D
      accelerate dropout (#9902) · 2e331c65
      dzhwinter 提交于
      * accelerate dropout
      
      * accelerate dropout
      
      * "fix the dropout test"
      
      * "rerun ci"
      
      * "fix ci"
      
      * "rerun ci"
      
      * "fix ci"
      
      * "fix"
      
      * "stage"
      
      * disable
      2e331c65
  14. 27 3月, 2018 1 次提交
  15. 22 3月, 2018 1 次提交
  16. 20 3月, 2018 2 次提交
  17. 12 2月, 2018 2 次提交
  18. 10 2月, 2018 2 次提交
  19. 30 1月, 2018 1 次提交
  20. 26 12月, 2017 2 次提交
  21. 21 12月, 2017 1 次提交
  22. 12 12月, 2017 1 次提交
    • Q
      Refine device context (#6433) · 61ec0b95
      QI JUN 提交于
      There are mainly following fixes:
      
      - take `DeviceContext` as the template parameter of math functors and OpKernel instead of `Place`
      - remove `eigen_device` interface in base class  `DeviceContext`
      - remove `GetEigenDevice` interface in `ExecutionContext` and base class `DeviceContext`
      - remove unused `platform::EigenDeviceConverter`
      - rename `REGISTER_OP_GPU_KERNEL` to `REGISTER_OP_CUDA_KERNEL`
      - rename `USE_GPU_ONLY_OP` to `USE_CUDA_ONLY_OP`
      61ec0b95
  23. 24 11月, 2017 1 次提交
  24. 28 9月, 2017 1 次提交
  25. 20 9月, 2017 1 次提交
  26. 19 9月, 2017 2 次提交
  27. 16 9月, 2017 1 次提交
  28. 03 9月, 2017 1 次提交
  29. 02 9月, 2017 1 次提交
  30. 08 8月, 2017 1 次提交
  31. 07 8月, 2017 1 次提交
  32. 04 8月, 2017 1 次提交
  33. 02 8月, 2017 1 次提交
  34. 31 7月, 2017 1 次提交
  35. 25 7月, 2017 1 次提交