- 14 1月, 2018 1 次提交
- 
- 
由 dzhwinter 提交于* "unified operators" * "add CUDNN register" * "add use cudnn attribute" * "add attribute" * "test conv tranpose op" * "remove duplicated attr" * "fix op test" * "add attribute to set cudnn" * "add more log" * "need layout op register support" * "add more log" * "change GetExpectedKernelType " * "fix Get attr in conv_op" * "fix CI" * "fix tests" * "removed kernel priority fallback" * "fix CI" * "fix stack pointer bug" * "refine buggy interface" * "add const cast to save life" * "fix get_output_with_grad" * "fix op test with dataformat" * ""fix pooling * "fix pooling test" * "fix CI" * "fix with_gpu error" * "add transform needed functional check" * "fix unpack list error" * "comment out parallel.do temporary" * "fix CI" * "fix compile doc error" * "make threshold larger" 
 
- 
- 05 1月, 2018 1 次提交
- 
- 
由 dzhwinter 提交于* "add c++ side kernel selection" * "add multiple kernel op test" * "kernel selection only support cudnn" * "better formatter" * "small fix with UseCPU" * "depends on change interface Get(Place, Library)" * "fix CI" * "fix python cudnn test" * "leave the register cudnn op to another PR" * "fix CI" * "use all kernel by default" * "fix CI" 
 
- 
- 03 1月, 2018 1 次提交
- 
- 
由 Yang Yu 提交于
 
- 
- 25 12月, 2017 2 次提交
- 24 12月, 2017 1 次提交
- 
- 
由 dzhwinter 提交于* "change operator interface" * "move devicepool to device_context" * "fix operator test" * "fix op_registry Run interface" * "net op passed. Need to fix nccl multi-Context" * "add nccl group function" * "add nccl group function" * "fix gpu count exceed 32 error" * "fix recurrent op, nccl op" * "change the other operators interface with Place" * "fix typo" * "fix pybind" * "fix device in python side" * "fix pybind failed" * "add init for test" * "fix CI" 
 
- 
- 21 12月, 2017 2 次提交
- 20 12月, 2017 2 次提交
- 19 12月, 2017 2 次提交
- 
- 
由 qiaolongfei 提交于
- 
由 fengjiayi 提交于
 
- 
- 18 12月, 2017 2 次提交
- 
- 
由 dzhwinter 提交于* "add DeviceContextPool" * "add devicecontextpool in pybind" * "add comments in python side " * "fix static link error" * "fix CI error" * "add executor.py" * "fix CI error" * "add with gpu macro" * "remove comment out codes" * "add TODO items" * "update init devices" 
- 
由 fengjiayi 提交于
 
- 
- 14 12月, 2017 2 次提交
- 27 11月, 2017 1 次提交
- 
- 
由 dangqingqing 提交于
 
- 
- 26 11月, 2017 1 次提交
- 
- 
由 dzhwinter 提交于* "make global tensor function independently" * "replace functor" * "fix inline template error" * "fix tensor array with CopyFrom" * "fix other case use CopyFrom" * "move the op interface hardly" * "fix operators" * "fix typo" * "delete dynamic recurrent rnn and fix gru_unit in debugmode" * "fix unique_ptr copy" * "fix cuda copy" * "fix namespace error" * "removed nccl python test" * "fix include error" * "fix typo" * fix copy util test 
 
- 
- 24 11月, 2017 1 次提交
- 
- 
由 QI JUN 提交于* is_training to is_test in dropout op * handle dropout and batch_norm operator when prune pdesc in testing mode * handle dropout and batch_norm operator when prune pdesc in testing mode * add get_inference_program method * fix dropout op * fix ci * test data after each batch training * refine code * refine test_book3 * fix ci * follow comments 
 
- 
- 14 11月, 2017 1 次提交
- 
- 
由 Qiao Longfei 提交于
 
- 
- 08 11月, 2017 1 次提交
- 
- 
由 Yu Yang 提交于* Compare Operator * Follow comments 
 
- 
- 06 11月, 2017 1 次提交
- 
- 
由 Yu Yang 提交于* Use stable_sort in lod_rank_table It is easy to debug and test when use `stable_sort`and the time complexity is not changed. * Add LoDTensorArray 
 
- 
- 04 11月, 2017 1 次提交
- 
- 
由 Yu Yang 提交于* Add LoDRankTable LoD Rank Table stores the `level` of `lod` which is ordered by sequence length in descending order. It is useful when implement dynamic RNN and is shared by dynamic RNN memory, dynamic RNN slice input and dynamic RNN slice output operators. * Add InferVarType 
 
- 
- 02 11月, 2017 1 次提交
- 
- 
由 Yu Yang 提交于* Init commit * Make executor use ProgramDescBind * Change Attribute from BlockDesc to BlockDescBind * Since we will get the program desc in RNN, just BlockDesc is not enough. * Add DeviceContext to Executor API * Rewrite RNN * Pass Python * AddBiasOp does not care num_flatten_dims * Stash * Fix MacOS Compile * Pass RNN forward * add python test * refactor test * Make compile pass * add gradopmaker * First draft done * Polish code * add grad op maker and grad infershape * Polish code * Fix backward.cc bug * Fix infershape * Rename function * add backward test * simplify recurrent test * Update * Pass unittest * Add comments & refine test * Add comments * refactor test * Complete Unittest * fix StepScopes enforce * Remove unused unittest * no type error * Update * Make RNN Pass unittest 
 
- 
- 01 11月, 2017 1 次提交
- 
- 
由 Yu Yang 提交于* Init commit * Make executor use ProgramDescBind * Change Attribute from BlockDesc to BlockDescBind * Since we will get the program desc in RNN, just BlockDesc is not enough. 
 
- 
- 31 10月, 2017 2 次提交
- 
- 
由 Qiao Longfei 提交于* improve unique_name, uniq id is related to prefix * fix join 
- 
由 QI JUN 提交于* add init_gflags interface * refine code * follow comments 
 
- 
- 28 10月, 2017 1 次提交
- 
- 
由 fengjiayi 提交于* Add `dump_to_file()` for ProgrameDescBind in pybind * Update * Add utility.py * typo * Fix bugs * Move add_feed/fetch_components to untility.py * Compelete dump * Follow comments * Change output of Prune() from inference to pointer * Expose Prune() to Python * Compelete save/load API of inference model * Fix errors * Debuging * Compelete unit tests * follow comments 
 
- 
- 27 10月, 2017 3 次提交
- 
- 
由 Dong Zhihong 提交于
- 
由 Dong Zhihong 提交于
- 
由 Dong Zhihong 提交于
 
- 
- 25 10月, 2017 1 次提交
- 
- 
由 Dong Zhihong 提交于
 
- 
- 24 10月, 2017 1 次提交
- 
- 
由 Yi Wang 提交于* Add print_operators_doc.cc * Update Escape * Correct a bug * Remove OpInfoMap::Iterate * Update the print_operators_doc.cc * Escape tab * Use auto& * Use auto& * Remove trailing , * clang-format C++ 
 
- 
- 21 10月, 2017 2 次提交
- 
- 
由 Yu Yang 提交于
- 
由 Yan Chunwei 提交于
 
- 
- 20 10月, 2017 4 次提交
- 
- 
由 Yu Yang 提交于* Unify `set_feed_variable` to one method * Move global scope to python, not in C++ 
- 
由 Yu Yang 提交于
- 
由 Yu Yang 提交于* Remove template parameter for Tensor methods * Also check the type is correct when data() * Simplize holder_ * Fix accuracy_op * Register Code 
- 
由 Yu Yang 提交于* Implement FC layer with helper * Update LayerHelper * Add debug string for Python ProtoBuf and Rename `Sync` to `Flush` * Add check of ProtoBuf initialization * Layer wrapper for FC * Fix unittest * Fix CI * Add code generator * AttributeChecker Better error log and speicalize bool Since lots of types can be cast to bool * Complete mlp, fit_a_line * Expose get global scope * Make global scope not thread-safe 1. It is no need to make global scope thread-safe, since it will be invoked in Python main thread. 2. Do not free the global scope when C++ exit. Let the OS free memories, otherwise, we need to handle the destroy dependencies. See https://google.github.io/styleguide/cppguide.html#Static_and_Global_Variables * Fix * Implementation of simple conv_2d layer * Stash * Remove private data members in OpRegister * Fix bugs * Stash * Expose FeedFetchList as VarType * Change ProgramDesc not a global variable * Polish code style * Stash * Correct implement BlockDesc destructor * Correct implement BlockDesc destructor * Unify program as parameter name * Fix bugs * Add unittest * Fix unit test error * Remove unused functions * Add clone for Python Program * Working on executor * Stash * Add glog as dependencies of ops * Use VLOG to logging some information is helpful when we debug Paddle * Expose VarDesc::persistable to Python * Test executor * Complete unittest * Polish code * Fix merge error * Follow comment * Polish Python Code 
 
- 
- 19 10月, 2017 1 次提交
- 
- 
由 Yu Yang 提交于* Change ProgramDesc not a global variable * Polish code style * Correct implement BlockDesc destructor * Unify program as parameter name 
 
- 
