Add paddle::variant and replace paddle::any #42139

chenwhql · 2022-04-22T13:42:21Z

PR types

Performance optimization

PR changes

Others

Describe

Add paddle::variant and replace paddle::any in kernel context to improve performance

Why not use boost::variant?
- Because phi cannot depend on boost in principle, nor can it depend on the entire boost lib just because of one variant

C++ std::variant vs std::any

原先fluid和phi兼容执行用的Context内用any存储attribute，分析发现，attr的构造析构引入了多次堆内存分配，极端情况下，attr的构建比kernel执行还要耗时，因此这里改为使用variant，避免不必要的堆内存分配

以paddle.slice（attr较复杂较多）为例：

替换前

替换后

可以看到明显去掉了一段any引入的new操作

TODO: replace paddle::any in InferMetaContext

paddle-bot-old · 2022-04-22T13:43:38Z

你的PR提交成功，感谢你对开源项目的贡献!
请关注后续CI自动化测试结果，详情请参考Paddle-CI手册。
Your PR has been submitted. Thanks for your contribution!
Please wait for the result of CI firstly. See Paddle CI Manual for details.

XiaoguangHu01

LGTM

* add variant and replace any * split attribute

* Add paddle::variant and replace paddle::any (#42139) * add variant and replace any * split attribute * Optimize dygraph GetExpectedKernelType perf (#42154) * opt dygraph scheduling * revert part impl * fix variant compile error (#42203) * replace any by variant in infermeta (#42181)

* bind elementwise_mod_op_xpu *test=kunlun * add more supported dtypes and UTs *test=kunlun * fix datatype error * add op to in xpu1_op_list * Update Mac cmake version >=3.15 (#41456) * Update Mac cmake version >=3.15 * notest;read test1 notest;read test2 notest;read test3 * fix inference link error * fix inference link error * fix windows link error * fix cmake_policy * fix build big size * Add paddle::variant and replace paddle::any (#42139) * add variant and replace any * split attribute * disable unittest failed in eager CI in temporary (#42101) * test=py3-eager * test=py3-eager * test=py3-eager * combine graph_table and feature_table in graph_engine (#42134) * extract sub-graph * graph-engine merging * fix * fix * fix heter-ps config * test performance * test performance * test performance * test * test * update bfs * change cmake * test * test gpu speed * gpu_graph_engine optimization * add dsm sample method * add graph_neighbor_sample_v2 * Add graph_neighbor_sample_v2 * fix for loop * add cpu sample interface * fix kernel judgement * add ssd layer to graph_engine * fix allocation * fix syntax error * fix syntax error * fix pscore class * fix * change index settings * recover test * recover test * fix spelling * recover * fix * move cudamemcpy after cuda stream sync * fix linking problem * remove comment * add cpu test * test * add cpu test * change comment * combine feature table and graph table * test * test * pybind * test * test * test * test * pybind * pybind * fix cmake * pybind * fix * fix * add pybind * add pybind Co-authored-by: DesmonDay <908660116@qq.com> * [CustomDevice] add eager mode support (#42034) * fix FlattenContiguousRangeOpConverter out dim error (#42087) * fix FlattenContiguousRangeOpConverter out dim error * update code * fix python3.10 compile bug on windows (#42140) * Optimize dygraph GetExpectedKernelType perf (#42154) * opt dygraph scheduling * revert part impl * fix incorrect usages of std::move and other compile errors (#41045) * fix bug of std::move and others * fix an compile error in debug mode * fix wrong copy assignment operator Signed-off-by: tiancaishaonvjituizi <452565578@qq.com> * reformat Signed-off-by: tiancaishaonvjituizi <452565578@qq.com> * reformat Signed-off-by: tiancaishaonvjituizi <452565578@qq.com> * fix ArrayRef constructor following llvm * fix format * fix conflict with master * fix variant compile error (#42203) * [Eager] Support numpy.ndarry in CastNumpy2Scalar (#42136) * [Eager] Remove redundancy code, fix fp16 case (#42169) * [Eager] Support div(scalar) in eager mode (#42148) * [Eager] Support div scalar in eager mode * Updated and remove debug logs * Remove list, use 'or' directly * Remove useless statement * fix recompute (#42128) * fix recompute * modify return * add LICENSE in wheel dist-info package (#42187) * replace any by variant in infermeta (#42181) * 【PaddlePaddle Hackathon 2】24、为 Paddle 新增 nn.ChannelShuffle 组网 API (#40743) * Add infermeta for ChannelShuffle * Create channel_shuffle_grad_kernel.h * Create channel_shuffle_kernel.h * Create channel_shuffle_sig.cc * Create channel_shuffle_op.cc ChannelShuffle算子的描述 * Create channel_shuffle_kernel_impl.h ChannelShuffle核函数的实现 * Create channel_shuffle_grad_kernel_impl.h ChannelShuffle反向核函数的实现 * Add kernel register of channel shuffle and grad 注册ChannelShuffle及其反向的核函数 * add nn.functional.channel_shuffle * add nn.ChannelShuffle * Create test_channel_shuffle.py * Update example of ChannelShuffle in vision.py * Update test_channel_shuffle.py * 修改channel_shuffle核函数的实现位置 * 修正代码格式 * 删除多余空格 * 完善channel_shuffle的错误检查 * Update unary.cc * Update channel_shuffle_op.cc * Update test_channel_shuffle.py * Update unary.cc * add channel_shuffle * Update test_channel_shuffle.py * Update vision.py * 调整代码格式 * Update channel_shuffle_sig.cc * 更新ChannelShuffle的文档 * 更新channel_shuffle的文档 * remove ChannelShuffleOpArgumentMapping * add ChannelShuffleGradInferMeta * Update channel_shuffle_op.cc * 调整channel_shuffle及其梯度的核函数的位置 * Do not reset default stream for StreamSafeCUDAAllocator (#42149) * remove redundant computation in Categorical.probs (#42114) * Downloading data for test_analyzer_vit_ocr (#42041) * Change server URL * update config * add test to parallel UT rule * add checksum to ensure files are downloaded * change downloading target * reuse existing variable * change target directory * fix en docs of some Apis (gradients, scope_guard, cuda_places, name_scope, device_guard, load_program_state, scale, ParamAttr and WeightNormParamAttr) (#41604) * Update scope_guard; test=document_fix * gradients; test=document_fix * gradients; test=document_fix * name_scope; test=document_fix * cpu_places; test=document_fix * WeightNormParamAttr; test=document_fix * cuda_places; test=document_fix * load_program_state; test=document_fix * device_guard; test=document_fix * device_guard; test=document_fix * ParamAttr; test=document_fix * scale; test=document_fix * scale; test=document_fix * update code example；test=document_fix Co-authored-by: Chen Long <1300851984@qq.com> * fix datatype error add op to in xpu1_op_list *test=kunlun * fix elementwise_mod op path error *test=kunlun * fix elementwise_mod UT error *test=kunlun * fix datatype error add op to in xpu1_op_list *test=kunlun add op to in xpu1_op_list fix elementwise_mod op path error *test=kunlun fix elementwise_mod UT error *test=kunlun Co-authored-by: tianshuo78520a <707759223@qq.com> Co-authored-by: Chen Weihang <chenweihang@baidu.com> Co-authored-by: pangyoki <pangyoki@126.com> Co-authored-by: seemingwang <seemingwang@users.noreply.github.com> Co-authored-by: DesmonDay <908660116@qq.com> Co-authored-by: ronnywang <524019753@qq.com> Co-authored-by: baoachun <962571062@qq.com> Co-authored-by: Zhou Wei <1183042833@qq.com> Co-authored-by: tiancaishaonvjituizi <452565578@qq.com> Co-authored-by: Weilong Wu <veyron_wu@163.com> Co-authored-by: Roc <30228238+sljlp@users.noreply.github.com> Co-authored-by: BrilliantYuKaimin <91609464+BrilliantYuKaimin@users.noreply.github.com> Co-authored-by: Ruibiao Chen <chenruibiao@baidu.com> Co-authored-by: Feiyu Chan <chenfeiyu@baidu.com> Co-authored-by: Sławomir Siwek <slawomir.siwek@intel.com> Co-authored-by: Yilingyelu <103369238+Yilingyelu@users.noreply.github.com> Co-authored-by: Chen Long <1300851984@qq.com>

add variant and replace any

bb8f354

split attribute

b302607

zyfncg approved these changes Apr 24, 2022

View reviewed changes

phlrain self-requested a review April 24, 2022 02:20

phlrain approved these changes Apr 24, 2022

View reviewed changes

XiaoguangHu01 approved these changes Apr 24, 2022

View reviewed changes

chenwhql merged commit 79f717d into PaddlePaddle:develop Apr 24, 2022

chenwhql mentioned this pull request Apr 24, 2022

Replace any by variant in infermeta context #42181

Merged

chenwhql added a commit to chenwhql/Paddle that referenced this pull request Apr 25, 2022

Add paddle::variant and replace paddle::any (PaddlePaddle#42139)

3dfb1cc

* add variant and replace any * split attribute

tiancaishaonvjituizi mentioned this pull request Apr 26, 2022

fix false positive warning of gcc>=9 #42265

Merged

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Add paddle::variant and replace paddle::any #42139

Add paddle::variant and replace paddle::any #42139

chenwhql commented Apr 22, 2022 •

edited

Loading

paddle-bot-old bot commented Apr 22, 2022

XiaoguangHu01 left a comment

Add paddle::variant and replace paddle::any #42139

Add paddle::variant and replace paddle::any #42139

Conversation

chenwhql commented Apr 22, 2022 • edited Loading

PR types

PR changes

Describe

paddle-bot-old bot commented Apr 22, 2022

XiaoguangHu01 left a comment

Choose a reason for hiding this comment

chenwhql commented Apr 22, 2022 •

edited

Loading