PoseFusion: SelectLSTM 实现强大的物体在手姿态估计 (PoseFusion: Robust Object-in-Hand Pose Estimation with SelectLSTM) - 专知论文

会员服务 ·

0

遮挡 · 数据集 · 融合 · 姿态估计 · 触觉传感器 ·

2023 年 4 月 10 日

PoseFusion: Robust Object-in-Hand Pose Estimation with SelectLSTM

翻译：PoseFusion: SelectLSTM 实现强大的物体在手姿态估计

Yuyang Tu,Junnan Jiang,Shuang Li,Norman Hendrich,Miao Li,Jianwei Zhang

Accurate estimation of the relative pose between an object and a robot hand is critical for many manipulation tasks. However, most of the existing object-in-hand pose datasets use two-finger grippers and also assume that the object remains fixed in the hand without any relative movements, which is not representative of real-world scenarios. To address this issue, a 6D object-in-hand pose dataset is proposed using a teleoperation method with an anthropomorphic Shadow Dexterous hand. Our dataset comprises RGB-D images, proprioception and tactile data, covering diverse grasping poses, finger contact states, and object occlusions. To overcome the significant hand occlusion and limited tactile sensor contact in real-world scenarios, we propose PoseFusion, a hybrid multi-modal fusion approach that integrates the information from visual and tactile perception channels. PoseFusion generates three candidate object poses from three estimators (tactile only, visual only, and visuo-tactile fusion), which are then filtered by a SelectLSTM network to select the optimal pose, avoiding inferior fusion poses resulting from modality collapse. Extensive experiments demonstrate the robustness and advantages of our framework. All data and codes are available on the project website: https://elevenjiang1.github.io/ObjectInHand-Dataset/

翻译：准确估计物体与机器人手之间的相对姿态对于许多操作任务至关重要。然而，大多数现有的物体在手姿态数据集使用两指夹持器，并且还假设物体保持在手中而没有任何相对运动，这并不代表真实世界的情况。为解决这个问题，我们提出了一种基于模拟遥控的人形影子手的六维物体在手姿态数据集。我们的数据集包括 RGB-D 图像、本体感知和触觉数据，涵盖了不同的抓握姿势、手指接触状态和物体遮挡。为了克服真实场景中显著的手部遮挡和有限的触觉传感器接触，我们提出了 PoseFusion，这是一种混合多模态融合方法，它将视觉和触觉感知通道的信息集成在一起。PoseFusion 从三个估计器（仅触觉、仅视觉和视觉触觉融合）生成三个候选物体姿态，然后通过 SelectLSTM 网络过滤选择最佳姿态，避免由于模态崩溃而导致的次优融合姿态。广泛的实验证明了我们框架的鲁棒性和优点。所有数据和代码都可在项目网站上获得：https://elevenjiang1.github.io/ObjectInHand-Dataset/

0

相关内容

【CVPR 2022】基于实例深度估计的统一深度感知全景分割 PanopticDepth: Per-Instance Depth Estimation for Unified Depth-Aware Panoptic Segmentation

【CVPR 2022】基于实例深度估计的统一深度感知全景分割 PanopticDepth: Per-Instance Depth Estimation for Unified Depth-Aware Panoptic Segmentation

专知会员服务

18+阅读 · 2022年3月19日

【CVPR 2022】基于时空解耦与重耦的RGB-D动作识别 Decoupling and Recoupling Spatiotemporal Representation for RGB-D-based Motion Recognition

【CVPR 2022】基于时空解耦与重耦的RGB-D动作识别 Decoupling and Recoupling Spatiotemporal Representation for RGB-D-based Motion Recognition

专知会员服务

14+阅读 · 2022年3月19日

【三维物体和手部姿态估计】综述论文最新进展，Recent Advances in 3D Object and Hand Pose Estimation

【三维物体和手部姿态估计】综述论文最新进展，Recent Advances in 3D Object and Hand Pose Estimation

专知会员服务

21+阅读 · 2020年6月13日

【CVPR2020-Facebook】从检测到3D目标，FroDO: From Detections to 3D Objects

【CVPR2020-Facebook】从检测到3D目标，FroDO: From Detections to 3D Objects

专知会员服务

33+阅读 · 2020年5月12日

【CVPR2020-台大】透视眼：学会透过障碍物看东西，Learning to See Through Obstructions

【CVPR2020-台大】透视眼：学会透过障碍物看东西，Learning to See Through Obstructions

专知会员服务

27+阅读 · 2020年4月3日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

运动物体检测与运动相机:一个全面的综述：Moving Objects Detection with a Moving Camera: A Comprehensive Review

运动物体检测与运动相机:一个全面的综述：Moving Objects Detection with a Moving Camera: A Comprehensive Review

专知会员服务

27+阅读 · 2020年1月17日

【论文推荐】小样本视频合成，Few-shot Video-to-Video Synthesis

【论文推荐】小样本视频合成，Few-shot Video-to-Video Synthesis

专知会员服务

24+阅读 · 2019年12月15日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

CVPR 2019 论文大盘点—目标检测篇

CVPR 2019 论文大盘点—目标检测篇

极市平台

33+阅读 · 2019年7月1日

CVPR 2019 | 34篇 CVPR 2019 论文实现代码

CVPR 2019 | 34篇 CVPR 2019 论文实现代码

AI科技评论

21+阅读 · 2019年6月23日

CVPR 2019 | 重磅！34篇 CVPR2019 论文实现代码

CVPR 2019 | 重磅！34篇 CVPR2019 论文实现代码

AI研习社

11+阅读 · 2019年6月21日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

CVPR2019| 05-14更新8篇论文及代码合集（含手势姿态估计/人脸/数据集等）

CVPR2019| 05-14更新8篇论文及代码合集（含手势姿态估计/人脸/数据集等）

极市平台

12+阅读 · 2019年5月14日

CVPR2019| 05-09更新10篇论文及代码合集（含图像恢复/图神经网络/人体形状重建等）

CVPR2019| 05-09更新10篇论文及代码合集（含图像恢复/图神经网络/人体形状重建等）

极市平台

17+阅读 · 2019年5月9日

CVPR2019 | 15篇论文速递（涵盖目标检测、语义分割和姿态估计等方向）

CVPR2019 | 15篇论文速递（涵盖目标检测、语义分割和姿态估计等方向）

AI研习社

15+阅读 · 2019年5月8日

【论文推荐】最新五篇度量学习相关论文—无标签、三维姿态估计、主动度量学习、深度度量学习、层次度量学习与匹配

【论文推荐】最新五篇度量学习相关论文—无标签、三维姿态估计、主动度量学习、深度度量学习、层次度量学习与匹配

专知

20+阅读 · 2018年4月5日

【推荐】YOLO实时目标检测(6fps)

【推荐】YOLO实时目标检测(6fps)

机器学习研究会

20+阅读 · 2017年11月5日

2017-最全手势识别/跟踪相关资源大列表分享（论文、数据集、比赛等）

2017-最全手势识别/跟踪相关资源大列表分享（论文、数据集、比赛等）

深度学习与NLP

64+阅读 · 2017年10月29日

基于深层特征学习的RGB-D人体行为识别方法

国家自然科学基金

4+阅读 · 2015年12月31日

5G极化码译码算法理论与实现关键技术研究

国家自然科学基金

0+阅读 · 2015年12月31日

CSP I-plus 修饰的内皮抑制素靶向抑制肝细胞癌转移的研究

国家自然科学基金

0+阅读 · 2015年12月31日

旋转飞行物体的状态估计与轨迹预测

国家自然科学基金

0+阅读 · 2014年12月31日

汽车行驶状态与道路参数非线性一体化估计方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

Affordance辅助服务机器人识别形状不规则物体研究

国家自然科学基金

0+阅读 · 2013年12月31日

miR-146a靶向IRAK1与TRAF6调控非小细胞肺癌转移的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

强力学仿生细胞外基质纳米纤维支架介导DCN shRNA长效转染ASCs的肌腱缺损修复研究

国家自然科学基金

0+阅读 · 2012年12月31日

E-cadherin阳性树突状细胞在非小细胞肺癌免疫微环境中的作用及机制

国家自然科学基金

0+阅读 · 2011年12月31日

基于物体棱线线流场的三维物体运动估计与结构重建研究

国家自然科学基金

0+阅读 · 2011年12月31日

PlaNeRF: SVD Unsupervised 3D Plane Regularization for NeRF Large-Scale Scene Reconstruction

Arxiv

0+阅读 · 2023年5月26日

Uncertain Pose Estimation during Contact Tasks using Differentiable Contact Features

Arxiv

0+阅读 · 2023年5月26日

Robust Category-Level 3D Pose Estimation from Synthetic Data

Arxiv

0+阅读 · 2023年5月25日

All Points Matter: Entropy-Regularized Distribution Alignment for Weakly-supervised 3D Segmentation

Arxiv

0+阅读 · 2023年5月25日

Emergency Response Person Localization and Vital Sign Estimation Using a Semi-Autonomous Robot Mounted SFCW Radar

Arxiv

0+阅读 · 2023年5月25日

Computer Vision for Construction Progress Monitoring: A Real-Time Object Detection Approach

Arxiv

0+阅读 · 2023年5月24日

Towards Large-Scale Small Object Detection: Survey and Benchmarks

Arxiv

40+阅读 · 2022年7月28日

Recovering 3D Human Mesh from Monocular Images: A Survey

Arxiv

12+阅读 · 2022年3月8日

Deep Learning-Based Human Pose Estimation: A Survey

Arxiv

27+阅读 · 2020年12月24日

3D Backbone Network for 3D Object Detection

Arxiv

12+阅读 · 2019年1月24日

VIP会员

文章信息

相关主题

触觉传感器

相关VIP内容

【CVPR 2022】基于实例深度估计的统一深度感知全景分割 PanopticDepth: Per-Instance Depth Estimation for Unified Depth-Aware Panoptic Segmentation

【CVPR 2022】基于实例深度估计的统一深度感知全景分割 PanopticDepth: Per-Instance Depth Estimation for Unified Depth-Aware Panoptic Segmentation

专知会员服务

18+阅读 · 2022年3月19日

【CVPR 2022】基于时空解耦与重耦的RGB-D动作识别 Decoupling and Recoupling Spatiotemporal Representation for RGB-D-based Motion Recognition

【CVPR 2022】基于时空解耦与重耦的RGB-D动作识别 Decoupling and Recoupling Spatiotemporal Representation for RGB-D-based Motion Recognition

专知会员服务

14+阅读 · 2022年3月19日

【三维物体和手部姿态估计】综述论文最新进展，Recent Advances in 3D Object and Hand Pose Estimation

【三维物体和手部姿态估计】综述论文最新进展，Recent Advances in 3D Object and Hand Pose Estimation

专知会员服务

21+阅读 · 2020年6月13日

【CVPR2020-Facebook】从检测到3D目标，FroDO: From Detections to 3D Objects

【CVPR2020-Facebook】从检测到3D目标，FroDO: From Detections to 3D Objects

专知会员服务

33+阅读 · 2020年5月12日

【CVPR2020-台大】透视眼：学会透过障碍物看东西，Learning to See Through Obstructions

【CVPR2020-台大】透视眼：学会透过障碍物看东西，Learning to See Through Obstructions

专知会员服务

27+阅读 · 2020年4月3日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

运动物体检测与运动相机:一个全面的综述：Moving Objects Detection with a Moving Camera: A Comprehensive Review

运动物体检测与运动相机:一个全面的综述：Moving Objects Detection with a Moving Camera: A Comprehensive Review

专知会员服务

27+阅读 · 2020年1月17日

【论文推荐】小样本视频合成，Few-shot Video-to-Video Synthesis

【论文推荐】小样本视频合成，Few-shot Video-to-Video Synthesis

专知会员服务

24+阅读 · 2019年12月15日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

《静默安全：海军水面舰艇上的防空程序评估》

Nature综述：金融网络中的物理学

美陆军新型AI/LLM工具：提升作战效能

《数字孪生与生成式AI融合构建战术网络弹性边缘智能》

相关资讯

CVPR 2019 论文大盘点—目标检测篇

CVPR 2019 论文大盘点—目标检测篇

极市平台

33+阅读 · 2019年7月1日

CVPR 2019 | 34篇 CVPR 2019 论文实现代码

CVPR 2019 | 34篇 CVPR 2019 论文实现代码

AI科技评论

21+阅读 · 2019年6月23日

CVPR 2019 | 重磅！34篇 CVPR2019 论文实现代码

CVPR 2019 | 重磅！34篇 CVPR2019 论文实现代码

AI研习社

11+阅读 · 2019年6月21日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

CVPR2019| 05-14更新8篇论文及代码合集（含手势姿态估计/人脸/数据集等）

CVPR2019| 05-14更新8篇论文及代码合集（含手势姿态估计/人脸/数据集等）

极市平台

12+阅读 · 2019年5月14日

CVPR2019| 05-09更新10篇论文及代码合集（含图像恢复/图神经网络/人体形状重建等）

CVPR2019| 05-09更新10篇论文及代码合集（含图像恢复/图神经网络/人体形状重建等）

极市平台

17+阅读 · 2019年5月9日

CVPR2019 | 15篇论文速递（涵盖目标检测、语义分割和姿态估计等方向）

CVPR2019 | 15篇论文速递（涵盖目标检测、语义分割和姿态估计等方向）

AI研习社

15+阅读 · 2019年5月8日

【论文推荐】最新五篇度量学习相关论文—无标签、三维姿态估计、主动度量学习、深度度量学习、层次度量学习与匹配

【论文推荐】最新五篇度量学习相关论文—无标签、三维姿态估计、主动度量学习、深度度量学习、层次度量学习与匹配

专知

20+阅读 · 2018年4月5日

【推荐】YOLO实时目标检测(6fps)

【推荐】YOLO实时目标检测(6fps)

机器学习研究会

20+阅读 · 2017年11月5日

2017-最全手势识别/跟踪相关资源大列表分享（论文、数据集、比赛等）

2017-最全手势识别/跟踪相关资源大列表分享（论文、数据集、比赛等）

深度学习与NLP

64+阅读 · 2017年10月29日

相关论文

PlaNeRF: SVD Unsupervised 3D Plane Regularization for NeRF Large-Scale Scene Reconstruction

Arxiv

0+阅读 · 2023年5月26日

Uncertain Pose Estimation during Contact Tasks using Differentiable Contact Features

Arxiv

0+阅读 · 2023年5月26日

Robust Category-Level 3D Pose Estimation from Synthetic Data

Arxiv

0+阅读 · 2023年5月25日

All Points Matter: Entropy-Regularized Distribution Alignment for Weakly-supervised 3D Segmentation

Arxiv

0+阅读 · 2023年5月25日

Emergency Response Person Localization and Vital Sign Estimation Using a Semi-Autonomous Robot Mounted SFCW Radar

Arxiv

0+阅读 · 2023年5月25日

Computer Vision for Construction Progress Monitoring: A Real-Time Object Detection Approach

Arxiv

0+阅读 · 2023年5月24日

Towards Large-Scale Small Object Detection: Survey and Benchmarks

Arxiv

40+阅读 · 2022年7月28日

Recovering 3D Human Mesh from Monocular Images: A Survey

Arxiv

12+阅读 · 2022年3月8日

Deep Learning-Based Human Pose Estimation: A Survey

Arxiv

27+阅读 · 2020年12月24日

3D Backbone Network for 3D Object Detection

Arxiv

12+阅读 · 2019年1月24日

相关基金

基于深层特征学习的RGB-D人体行为识别方法

国家自然科学基金

4+阅读 · 2015年12月31日

5G极化码译码算法理论与实现关键技术研究

国家自然科学基金

0+阅读 · 2015年12月31日

CSP I-plus 修饰的内皮抑制素靶向抑制肝细胞癌转移的研究

国家自然科学基金

0+阅读 · 2015年12月31日

旋转飞行物体的状态估计与轨迹预测

国家自然科学基金

0+阅读 · 2014年12月31日

汽车行驶状态与道路参数非线性一体化估计方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

Affordance辅助服务机器人识别形状不规则物体研究

国家自然科学基金

0+阅读 · 2013年12月31日

miR-146a靶向IRAK1与TRAF6调控非小细胞肺癌转移的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

强力学仿生细胞外基质纳米纤维支架介导DCN shRNA长效转染ASCs的肌腱缺损修复研究

国家自然科学基金

0+阅读 · 2012年12月31日

E-cadherin阳性树突状细胞在非小细胞肺癌免疫微环境中的作用及机制

国家自然科学基金

0+阅读 · 2011年12月31日

基于物体棱线线流场的三维物体运动估计与结构重建研究

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员