Aller au contenu principal
2025 conference-paper

Dual Stream Framework with Self-Correcting Memory for Semi-Supervised Video Object Segmentation

0Citations signalées, ce qui n’est pas une note de qualité
2Institutions déclarées
1Pays d’affiliation déclarés

Rattachement africain : cn. Niveau de preuve : code pays fourni par la source.

Le résumé fourni par la source

Recent memory-based approaches have shown promising performance in semi-supervised video object segmentation. However, they still face two core challenges. First, their heavy reliance on pixel-level feature matching makes them vulnerable to background distractors that closely resemble the target object, resulting in erroneous segmentations. Second, these incorrect segmentation masks are indiscriminately stored in memory, causing error accumulation and deteriorating performance over time. To address these challenges, we propose the Instance Aware Self-Correcting Model (IASC), which integrates three key components: the Dual-Stream Framework, the Online Adaptation Module (OAM), and the Self-Correcting Module (SCM). The Dual-Stream Framework leverages complementary representations by combining the Appearance Feature Stream, which recovers fine-grained pixel-level details through similarity-based matching, and the Semantic Mask Stream, which generates instance-level attention using reference masks to suppress distractors and enhance segmentation robustness. The OAM dynamically fine-tunes the key projector using the first annotated frame, improving the accuracy of pixel-level matching and the reliability of instance-level guidance. Additionally, the SCM refines unreliable predictions, effectively mitigating error propagation during memory updates. Extensive experiments on benchmark datasets demonstrate that IASC achieves state-of-the-art performance, striking an effective balance between segmentation accuracy and computational efficiency.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
Dual Stream Framework with Self-Correcting Memory for Semi-Supervised Video Object Segmentation
Date Crossref
30/06/2025
Éditeur
IEEE
Type
proceedings-article

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.

Les institutions déclarées

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Les sujets associés

Visual Attention and Saliency DetectionAdvanced Neural Network ApplicationsMultimodal Machine Learning Applications

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.