merg to local by junjun315 · Pull Request #24 · junjun315/Paddle

junjun315 · 2019-05-27T06:23:00Z

No description provided.

* fix sqrt_grad_grad unittest. test=develop * disable sqrt_grad_grad unittest. test=develop

* add set_not_owned function for graph * add scope set. test=develop * add scope_ptr enforce not null before setting.test=develop

* init auto loss scaling test=develop * change API.spec * change ifelse to switch and use reduce_sum to optimize checking isfinite test=develop * Remove redundant code test=develop

Refine the Executor when the num_thread=1

This reverts commit aca60e9.

add inductive shape index

…ory (#17374) * improve the API Sample of DataFeeder, memory_optimize and release_memory, test=develop * update API.spec, test=develop, test=document_preview * tweak the code format of feed API, test=develop * update API.spec, test=develop * improve doc for DataFeeder and default_main_program, test=develop

* improve gru unit performance. refine code test=develop Signed-off-by: zhaoyuchen <[email protected]> * Add conditional compile for gru opt Not enable gru opt if compute ability < 700 test=develop Signed-off-by: zhaoyuchen <[email protected]> * refine code. test=develop Signed-off-by: zhaoyuchen <[email protected]>

* add cache_update_mutex_ for operator

* test=develop, add gradient sort backward strategy * test=develop, fix test by add FLAGS_cudnn_deterministic on new tests * test=develop, fix memory leak in dygraph mode * test=develop, fix memory leak in dygraph mode * test=develop, polish code * test=develop, polish code * test=develop, polish code

* add var grad hook test=develop

test=develop

* add record_event test=develop * remove csp test=develop

* examples use code-block in dataset.py test=develop test=document_preview * add QueueDataset example test=develop test=document_preview

* fix data_feed_desc.py example run error test=develop test=test=document_preview * fix data_feed_desc.py example display error test=develop test=document_preview * update API.spec for DataFeedDesc test=develop test=document_preview

add elementwise_sub_grad_grad op for backward of backward calculation

test=develop

* double backward, elementwise_div * fix dx empty. test=develop * bug fix (#17392) fix secure bug * Eanble stack operator for a Ngraph, test=develop (#17406) * fix sqrt_grad_grad unittest. test=develop (#17410) * fix sqrt_grad_grad unittest. test=develop * disable sqrt_grad_grad unittest. test=develop * test=develop, fix unittest * test=develop, fix unittest * test=develop, fix unittest * test=develop, fix bug * fix unittest. test=develop * fix unittest dx. test=develop * tmp fix! for test... test=develop * reduce tmp, test=develop * test=develop, reduce tmp * fix broadcast unittest. test=develop * fix format. test=develop * refine code. test=develop * refine code. test=develop * refine GetDoubleGradSafeTensor. test=develop * fix format. test=develop

* fix the random compilation failure on windows

test=develop

* improve the doc of paddle.fluid.memory_optimize, test=develop * fix typo, test=develop

…ebug, test=develop (#17491)

* optimize communicator flag * change flags in init py test=develop

* fix quantize_squash_pass segfault when there is no tensor linked do Bias input test=develop * add googlenet test test=develop * fix concat CreateKey not using input format test=develop

* add conv_concat_relu fuse test=develop * add test code test=develop * added missing include with unordered_map test=develop * review fixes for wojtuss test=develop * remove 'should (not) be fused' comment statements one of them was invalid anyway test=develop

Python examples of fluid.layers.io.double_buffer and some BuildStrategy's methods.

* This PR adds broadcast for multi-process. And it could be used in dynamic graph to broadcast parameters.

* Fix the example code in some Python API. test=develop * Fix the example code in some Python API by adding import. test=develop

test=develop

add Run Prepared Ctx, fix pybind problem

* fix DecayedAdagrad example; test=develop test=document_preview * add space; test=develop

…expand_as ] (#17210) * fix example; test=develop * fix api spec; test=develop * fix api spec; test=develop * add doc check test=develop test=document_preview * test=develop,test=document_preview add blank line to fix format, add one more "import" * fix bug; test=develop * fix bug; test=develop

test=develop

* fix unique_name growth bug in dygraph mode,test=develop * change generate_tmp to generate_with_ignorable_key,test=develop

#17588) * add __str__ method for tensor and lodtensor to support print test=develop

#17530) * BugFix: fix api comment of paddle.fluid.clip.GradientClipByValue * test=develop, test=document_preview

* fuse mul and elementwise add to fc * Reimplement the FC forward operator * Fix FC MKLDNN integration by transposing weights * Add FC MKLDNN Pass test=develop * FC MKLDNN Pass: change memcpy to std::copy * Fix MKLDNN FC handling of mismatch input and weights dims * Lower tolerance for MKL-DNN in resnet50 test test=develop * Adjust FC to support MKLDNN Op placement test=develop * Adjust Placement Op to set use_mkldnn attribute for graph test=develop * MKLDNN FC: fix weights format so that gemm version is called test=develop * FC MKLDNN: Remove tolerance decrease from tester_helper * FC MKL-DNN: Refactor the code, change input reorder to weight reorder * MKL-DNN FC: Introduce operator caching test=develop * FC MKL-DNN: Fix the tensor type in ExpectedKernelType test=develop * FC MKL-DNN: fix style changes test=develop * FC MKL-DNN: fallback to native on non-supported dim sizes test=develop * FC MKLDNN: fix CMake paths test=develop * FC MKLDNN: Refine placement pass graph mkldnn attribute test=develop * Fix Transpiler error for fuse_conv_eltwise test=develop * Fix missing STL includes in files test=develop * FC MKL-DNN: Enable new output size computation Also, refine pass to comply with newest interface. test=develop * FC MKL-DNN: enable only when fc_mkldnn_pass is enabled * FC MKL-DNN: Allow Weights to use oi or io format * FC MKL-DNN: Adjust UT to work with correct dims test=develop * Enable MKL DEBUG for resnet50 analyzer test=develop * FC MKL-DNN: Improve Hashing function test=develop * FC MKL-DNN: Fix shape for fc weights in transpiler * FC MKL-DNN: Update input pointer in re-used fc primitive * Add log for not handling fc fuse for unsupported dims test=develop * FC MKL-DNN: Move transpose from pass to Op Kernel test=develop * FC MKL-DNN: Disable transpose in unit test test=develop * FC MKL-DNN: Remove fc_mkldnn_pass from default list * Correct Flag for fake data analyzer tests test=develop * FC MKL-DNN: Add comment about fc mkldnn pass disablement test=develop * FC MKL-DNN: Disable fc in int8 tests test=develop

* fluid int8 train and trt int8 predict align. trt int8 predict init op converter * 2. align fluid int8 train and trt int8 inference. enhance quant dequant fuse pass enhance op converter, trt engine, trt engine op, trt subgraph pass. * 3. add delete_quant_dequant_pass for trt test=develop * 4. add the missing file test=develop * 5. i modify the c++ interface, but forget to modify the pybind code fix the IS_TRT_VERSION_GE bug, and fix elementwise op converter test=develop

* Add Dockerfile for cuda9 and cuda10 Add Dockerfile for building cuda9 cuda10 images.

* gather_op support int64_t index by adding a template typename * add UT and rename typename test=develop

* add data parallel batch

* Add Dockerfile for cuda9 and cuda10

test=develop

* Revert "Revert "Fix allocator bug"" This reverts commit 174d0d0. * Revert "fix travis ci" This reverts commit 5656fa9. test=develop * add inlined_vector.h, test=develop * add inlined_vector_test,test=develop * clean code of allocator,test=develop * delete zero_size_allocator.h,test=develop * fix failed unittest,test=develop

* enhance print

… APIs. (#17639) * fix the bug that sub_scope_ may be null in AnalysisPredictor::Run. * add more directions about io APIs' docs. * update the API.spec. test=develop test=document_preview

test=develop

heavengate and others added 30 commits May 15, 2019 20:41

fix sqrt_grad_grad unittest. test=develop (#17410)

58d5c61

* fix sqrt_grad_grad unittest. test=develop * disable sqrt_grad_grad unittest. test=develop

Add setting Scope function for the graph class (#17417)

4a1b7fe

* add set_not_owned function for graph * add scope set. test=develop * add scope_ptr enforce not null before setting.test=develop

init auto loss scaling (#17194)

30e178f

* init auto loss scaling test=develop * change API.spec * change ifelse to switch and use reduce_sum to optimize checking isfinite test=develop * Remove redundant code test=develop

[Speed] Refine the Executor when the num_thread=1 (#17405)

e336dc8

Refine the Executor when the num_thread=1

Revert "remove unnecessary prepare_data (#17080)" (#17432)

5babcd0

This reverts commit aca60e9.

fix recurrent_op,test=develop (#17433)

712bfb1

add inductive shape index (#17435)

43c9561

add inductive shape index

fix assert,test=develop (#17445)

3a9ae28

test=develop, fix AdgradOptimizer example code (#17401)

15453d0

add cache_update_mutex_ for operator test=develop (#17124)

728bbaa

* add cache_update_mutex_ for operator

polish parallel dygraph code (#17164)

0217555

* add var grad hook test=develop

support sparse table get shard_num from TableParameter (#17443)

05df39a

test=develop

Add record event And remove CSP (#17447)

5a6ab38

* add record_event test=develop * remove csp test=develop

examples use code-block in dataset.py (#17451)

e32f4c4

* examples use code-block in dataset.py test=develop test=document_preview * add QueueDataset example test=develop test=document_preview

fix data_feed_desc.py example run error (#17452)

75cda4d

* fix data_feed_desc.py example run error test=develop test=test=document_preview * fix data_feed_desc.py example display error test=develop test=document_preview * update API.spec for DataFeedDesc test=develop test=document_preview

support elementwise_sub double backward (#17476)

977e9fc

add elementwise_sub_grad_grad op for backward of backward calculation

fix sqrt unittest. test=develop (#17440)

14f2236

fix recurrent fwd bug when no backward and scope clear (#17460)

3d4e826

Fix compiling error with cuDNN 5.1 (#17458)

97f0ec2

test=develop

fix the random compilation failure on windows test=develop (#17475)

ca3ba37

* fix the random compilation failure on windows

add clear ops in dygraph optimizers,test=develop (#17484)

65dd7ec

remove unused expected_kernel_cache_pass (#17486)

32da5e9

test=develop

improve the doc of paddle.fluid.memory_optimize, test=develop (#17473)

f82e4d7

* improve the doc of paddle.fluid.memory_optimize, test=develop * fix typo, test=develop

remove two useless flags: enable_subgraph_optimize, memory_optimize_d…

c3949f5

…ebug, test=develop (#17491)

fix uniform_random op,test=develop (#17492)

9eb19df

Optimize communicator flags (#17494)

287de41

* optimize communicator flag * change flags in init py test=develop

heavengate and others added 29 commits May 24, 2019 11:31

refine shape and split test. test=develop (#17545)

3db9c8c

fix quantize_squash_pass segfault when no tensor linked to Bias (#17292)

bccb0ba

* fix quantize_squash_pass segfault when there is no tensor linked do Bias input test=develop * add googlenet test test=develop * fix concat CreateKey not using input format test=develop

BuildStrategy api comment (#17348)

2280f18

Python examples of fluid.layers.io.double_buffer and some BuildStrategy's methods.

Add broadcast operators (#17503)

b5f4d5e

* This PR adds broadcast for multi-process. And it could be used in dynamic graph to broadcast parameters.

Fix the example code in some Python API. (#17343)

2a7b321

* Fix the example code in some Python API. test=develop * Fix the example code in some Python API by adding import. test=develop

Fix trust ratio in lamb (#17614)

e8990e6

test=develop

add Run Prepared Ctx (#17616)

326bf82

add Run Prepared Ctx, fix pybind problem

Fix decayed adagrad example (#17390)

e53119f

* fix DecayedAdagrad example; test=develop test=document_preview * add space; test=develop

Enable logical operators for the nGraph Bridge. (#17543)

e9216d0

test=develop

Fix dygraph unique name bug (#17592)

887a39f

* fix unique_name growth bug in dygraph mode,test=develop * change generate_tmp to generate_with_ignorable_key,test=develop

add __str__ method for tensor and lodtensor to support print test=dev… (

6724a65

#17588) * add __str__ method for tensor and lodtensor to support print test=develop

[DOC][PYTHON] Fix api comment of paddle.fluid.clip.GradientClipByValue (

21138eb

#17530) * BugFix: fix api comment of paddle.fluid.clip.GradientClipByValue * test=develop, test=document_preview

Enable elementwise pow operator for ngraph (#17526)

2b83d75

Add Dockerfile for cuda9 and cuda10 (#17600)

febc07f

* Add Dockerfile for cuda9 and cuda10 Add Dockerfile for building cuda9 cuda10 images.

Gather Op Index Support int64_t datatype (#17610)

1670db5

* gather_op support int64_t index by adding a template typename * add UT and rename typename test=develop

Add data distributed_sampler (#17573)

9322216

* add data parallel batch

test=develop (#17643)

9f85afb

fix conflicts,test=develop (#17186)

bbd6e43

Remove Docker build for CI tasks (#17650)

afc3d85

* Add Dockerfile for cuda9 and cuda10

Fix the usage of out_grad lod in sequence_slice_op. (#17625)

430e256

test=develop

Polish Print Op (#17651)

3430173

* enhance print

Fix the bug in the AnalysisPredictor and add more directions about io…

8bd651b

… APIs. (#17639) * fix the bug that sub_scope_ may be null in AnalysisPredictor::Run. * add more directions about io APIs' docs. * update the API.spec. test=develop test=document_preview

[NGraph] Enable gelu operator for the nGraph Bridge. (#17547)

b1bd483

test=develop

Add multi-ncclcomm and 2D ncclallreduce support. (#17263)

65bbf95

junjun315 merged commit 426f940 into junjun315:develop May 27, 2019

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

merg to local#24

merg to local#24
junjun315 merged 103 commits intojunjun315:developfrom
PaddlePaddle:develop

junjun315 commented May 27, 2019

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

20 participants

Conversation

junjun315 commented May 27, 2019

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

20 participants