1. 14 4月, 2023 1 次提交
    • F
      1. modify set_value op, use Scalars to represent attr `values`, instead of a... · dd2a749a
      Feiyu Chan 提交于
      1. modify set_value op, use Scalars to represent attr `values`, instead of a bunch of attributs of various types; (#52408)
      2. add program converter and set_value op as an example, which provides the functionality to convert `paddle::framework::ProgramDesc` between old and new formats(the differences are mainly some operators with incompatible updates in the definition);
      3. program version and operator version map now are always saved when serializing `paddle::framework::ProgramDesc` to identify the version;
      3. provide an option `legacy_format=false` in  serialization of `paddle::framework::ProgramDesc`, it decided whether to convert ProgramDesc back to a legacy format, which is compatible for paddle 2.4.2 or earlier versions to load and execute;
      4. deserialization of `paddle::framework::ProgramDesc` is now automatically detecting whether the bytes it receives is in legacy format(contains any of the operators that has been incompatibly updated and have any attribute of type `Scalar`) and convert it to new format. But if you want a faithful deserialization without the automatic conversion, you can use protobuf's deserialization instead. Though it is not recommended, it can be used for the purpose of testing.
  2. 25 3月, 2023 1 次提交
  3. 02 3月, 2023 1 次提交
  4. 22 2月, 2023 1 次提交
  5. 12 1月, 2023 1 次提交
  6. 09 11月, 2022 1 次提交
  7. 08 11月, 2022 1 次提交
  8. 23 10月, 2022 1 次提交
  9. 27 9月, 2022 1 次提交
  10. 16 8月, 2021 1 次提交
  11. 22 6月, 2021 1 次提交
  12. 19 10月, 2020 1 次提交
  13. 30 8月, 2020 1 次提交
  14. 13 8月, 2020 1 次提交
  15. 07 8月, 2020 1 次提交
  16. 06 8月, 2020 1 次提交
    • T
      add heter ps mode (#25682) · 0cb60c70
      Thunderbrook 提交于
      * add heter ps mode
      * code style
      * add with_pslib
      * unitest
      * code style
      * code style
      * code style
      * code style
      * code style
      * code style
      * code style
      * code style
      * test monitor
      * prepare trainer
      * code style
  17. 30 7月, 2020 1 次提交
  18. 14 4月, 2020 1 次提交
  19. 18 3月, 2020 1 次提交
  20. 17 3月, 2020 1 次提交
  21. 02 2月, 2020 1 次提交
  22. 26 11月, 2019 1 次提交
  23. 10 9月, 2019 1 次提交
  24. 06 9月, 2019 1 次提交
  25. 16 8月, 2019 1 次提交
  26. 12 8月, 2019 1 次提交
  27. 02 8月, 2019 1 次提交
    • J
      support filelist size < trainer num && fix pull dense (#18956) · 02c370c3
      jiaqi 提交于
      * support filelist size < trainer num
      * pull dense when stop, to make sure local dense params are same as pserver, so save paddle model will save dense model same as pserver
      *  enable QueueDataset train same filelist for serveral times
  28. 22 7月, 2019 1 次提交
  29. 27 6月, 2019 1 次提交
    • H
      supports collective communicated training (#18175) · b7128bac
      HaoRen 提交于
      * fix prepare context redundant code problem, optimize executor by caching create_varaiables
      * supports collective training in executor
      * make fetch_list runable with variables, add more unittest for use_program_cache
      * fix comment
      * use unique name for nccl_id
      * supports output to stream in program_to_code
      * insert sync_comm_stream before regularization; add skip_op_callstack capability in program_to_code
      * set op role in collective training
      * add collective op role
      * remove orig file
      * add build optimizer by strategy
      * add collective strategy
      * refine collective strategy
      * add multi-process role maker
      * refine strategy building factory so that we can easily plugin more strategy
      * scale loss grad in collective sgd transpiler
      * add support for distributed fc
      * code format
      * revert some features for dist fc
      * add support for distributed fc training
      * fix prepare context redundant code problem, optimize executor by caching create_varaiables
      * supports collective training in executor
      * make fetch_list runable with variables, add more unittest for use_program_cache
      * use unique name for nccl_id
      * supports output to stream in program_to_code
      * insert sync_comm_stream before regularization; add skip_op_callstack capability in program_to_code
      * set op role in collective training
      * add collective op role
      * fix comment
      * remove orig file
      * add build optimizer by strategy
      * add collective strategy
      * refine collective strategy
      * add multi-process role maker
      * refine strategy building factory so that we can easily plugin more strategy
      * scale loss grad in collective sgd transpiler
      * add support for distributed fc
      * code format
      * revert some features for dist fc
      * add support for distributed fc training
      * test=develop
      add collective op unittest standard
      * test=develop
      remove the test_collective directory
      * test=develop
      remove the test_collective directory
      * remove slicegather test
      * code format for reducescatter
      * update attr of shard_index_op
      * Modify macro nccl_helper
      * remove test without distribute
      * macro collective_helper
      * marcro update
      * test=develop
      update support python3.5
      * test=develop change gpu memory use to 0.1 when test
      * test=develop
      update ut equal func
      * test=develop
      set flags to 1.5
      * test=develop fix pickle dumple  py35
      * test=develop
      fix divide in slice and add sync_comm_stream
      update atol and rtol to 1e-05
      rm shard_index op and test
      modify read input from file to read from memory
      remove origin_program in framework and add i/o in c_sync_calc_stream
      * test=develop update unittest sync operator I/O
  30. 17 6月, 2019 2 次提交
  31. 12 6月, 2019 1 次提交
  32. 23 5月, 2019 1 次提交
  33. 09 5月, 2019 1 次提交
  34. 25 4月, 2019 1 次提交