Skip to content
Preprint

FoRIS: Progressive Foreground Refinement for Training-Free In-Context Segmentation

Sep 2026 · 0 citations · 56 references
Computer Science

TL;DR

FoRIS consists of three key stages: Foreground Purification, Foreground Localization, and Foreground Consolidation, which progressively suppress background distractions, localize discriminative target regions, and recover complete foreground structures through semantic aggregation.

Abstract

In-Context Segmentation (ICS) aims to precisely segment arbitrary semantic concepts, such as objects or parts, given one or a few annotated visual exemplars. In this paper, we revisit ICS from a more classical segmentation perspective, viewing it as a coarse-to-fine progressive refinement process. Rather than directly predicting the final mask through reference-query matching, we progressively refine the segmentation from coarse and ambiguous foreground responses to precise and complete foreground structures. Building upon this perspective, we propose a training-free in-context segmentation framework, termed FoRIS. Specifically, FoRIS consists of three key stages: Foreground Purification, Foreground Localization, and Foreground Consolidation, which progressively suppress background distractions, localize discriminative target regions, and recover complete foreground structures through semantic aggregation. Experimental results demonstrate that FoRIS achieves SOTA performance across semantic and part segmentation tasks, with average improvements of 4.5 and 4.8 mIoU points over existing approaches in the 1-shot and 5-shot settings, respectively. Code: https://github.com/Xi-Mu-Yu/FoRIS.

View source

Similar papers

Preprint Aug 2026

FoundYou: A Unified Model for Personalized Segmentation and Retrieval

Personalized segmentation and personalized retrieval both aim to identify the same physical object across different images. While the former localizes the object within a target image, the latter retrieves images where it appears. Despite this shared instance-level objective, the two tasks have largely evolved separate...

Gabriele Trivigno, Marcos Alfaro, C. Cuttano et al. · 0 citations
#artificial intelligence Preprint Sep 2026

ENEAS: Embedding-guided Neural Ensemble for Adaptive Segmentation

We present ENEAS, a unified, text-promptable method for instance tracking and semantic discovery. Text-promptable segmentation models, including the latest foundation models such as SAM 3, still suffer from temporal hallucinations, spatial fragmentation, and semantic misclassification: they fail to report target absenc...

J. del Pino, Salvador Rodríguez, Alejandro Garabito et al. · 0 citations
Preprint Sep 2026

SRPR-Net: Semantic and Relational Prompt Refinement for Automated SAM-based Instance Segmentation

Instance segmentation is a fundamental computer vision task with diverse real-world applications. Recently, prompt-driven foundation models have shown promising generalization. However, automated prompting remains limited by insufficient semantic guidance and inter-instance modeling. To address this challenge, we propo...

Lu Liu, Guo-Jie Li, Sun-Cheng Xiang et al. · 0 citations
Conference Open access Sep 2026

F³S: feature fused few-shot segmentation with CLIP guided semantic priors and Sinkhorn attention refinement

Few-shot segmentation (FSS) remains a significant challenge due to the scarcity of annotated data and the need for precise object localization across novel classes. Existing approaches often rely on single-backbone architectures and coarse priors, which struggle to capture detailed semantics and precise spatial alignme...

Guo-Hua Geng, Xiao-Feng Wang · 0 citations
Preprint Sep 2026

Exploiting Target Knowledge from MLLMs for Robust Few-Shot Segmentation

A novel framework that mines target knowledge using the strong reasoning capacity of Multimodal Large Language Models (MLLMs) and employs it to enhance FSS, which shows promising results and largely surpasses existing methods.

Yi-Jun Hu, Heng Fan, Li-Bo Zhang · 0 citations
Conference Open access Sep 2026

From Language to Segmentation: Collaborative Category-Guided Unsupervised Camouflaged Object Detection with SAM3

Camouflaged Object Detection (COD) aims to segment objects that are hidden within complex backgrounds. Due to the low visual contrast of camouflaged objects, annotations are costly, motivating unsupervised COD (UCOD) to eliminate labeling expenses. Most UCOD methods follow the “MLLMs + other foundation models + SAM” pa...

Hua-Feng Chen, Yueming Lyu, Caifeng Shan · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.