Skip to content

Model-agnostic pose estimation for enhanced collaborative robot grasping via binocular vision

Aug 2026 · Signal, Image and Video Processing · Vol 20 · 0 citations · 43 references

TL;DR

A novel 7-DoF grasping pose generation framework that integrates sparse attention and null convolution is introduced, which enhances the model’s ability to capture fine-grained features from point clouds, significantly improving the accuracy of parallel gripping pose estimation.

View source

Similar papers

Open access Jul 2026

A Robust Visual Grasping Method for Robots in Cluttered and Stacked Scenes

An iterative closed-loop optimization framework that deeply couples SAM with FoundationPose and designs a multi-dimensional confidence assessment module that integrates both the 2D image domain and the 3D geometric domain to comprehensively evaluate the reliability of the current pose.

Zhiqiang Gao, Mengqi Li, Huihui Bai et al. · 0 citations
Conference 2026

Two-stage Monocular 6D Pose Estimation for Small Cubic Objects

This paper studies monocular 6D pose estimation of small cubic objects from a single RGB image and proposes a two-stage manipulation- oriented framework, which achieves the strongest overall balance in ADD-S, translation accuracy, rotation stability, and task-oriented usability metrics.

Xinmiao Du · 0 citations
#artificial intelligence Preprint Aug 2026

Iterative Grasp Pose Refinement: A Deep Reinforcement Learning Approach for 2D Vision

A reinforcement learning-based framework for robotic grasp refinement, integrating keypoint-based object representations with a Deep Q-Network (DQN), is proposed, offering a scalable and adaptable solution for contact-rich manipulation tasks.

Amir Arsalan Nematollahi, Shayan Ahmadi, M. T. Masouleh et al. · 0 citations
Open access 2026

Automated Model Selection for Task-Specific RGB–Tactile Fusion in In-Hand Grasp Pose Estimation

This study investigates in-hand pose estimation of a USB stick that is already held within a robotic gripper and provides a controlled task-specific analysis showing that structured model selection can improve multimodal grasp-pose regression within the evaluated setup.

A. Altenbuchner, Bsher Karbouj, Fabian Dilly et al. · 0 citations
Open access Aug 2026

A Unified Multi-Task Deep Learning Framework for Robotic Bin-Picking of Planar Objects

An innovative approach is introduced for the random bin-picking of planar objects by developing a multi-task model for instance segmentation and keypoint detection in 2D images and a grasp candidate selection strategy is proposed to enable reliable grasping in cluttered industrial environments.

Ho Chi Minh, The-Thinh Pham, Tuan-Khanh Nguyen et al. · 0 citations