[2411.15087] Phrase-Instance Alignment for Generalized Referring

[2411.15087] Phrase-Instance Alignment for Generalized Referring Segmentation

arXiv - Machine Learning March 26, 2026 3 min read

About this article

Abstract page for arXiv paper 2411.15087: Phrase-Instance Alignment for Generalized Referring Segmentation

Computer Science > Computer Vision and Pattern Recognition arXiv:2411.15087 (cs) [Submitted on 22 Nov 2024 (v1), last revised 24 Mar 2026 (this version, v2)] Title:Phrase-Instance Alignment for Generalized Referring Segmentation Authors:E-Ro Nguyen, Hieu Le, Dimitris Samaras, Michael S. Ryoo View a PDF of the paper titled Phrase-Instance Alignment for Generalized Referring Segmentation, by E-Ro Nguyen and 3 other authors View PDF HTML (experimental) Abstract:Generalized Referring expressions can describe one object, several related objects, or none at all. Existing generalized referring segmentation (GRES) models treat all cases alike, predicting a single binary mask and ignoring how linguistic phrases correspond to distinct visual instances. To this end, we reformulate GRES as an instance-level reasoning problem, where the model first predicts multiple instance-aware object queries conditioned on the referring expression, then aligns each with its most relevant phrase. This alignment is enforced by a Phrase-Object Alignment (POA) loss that builds fine-grained correspondence between linguistic phrases and visual instances. Given these aligned object instance queries and their learned relevance scores, the final segmentation and the no-target case are both inferred through a unified relevance-weighted aggregation mechanism. This instance-aware formulation enables explicit phrase-instance grounding, interpretable reasoning, and robust handling of complex or null expressions....

Originally published on March 26, 2026. Curated by AI News.

Ai Infrastructure

UMKC Announces New Master of Science in Artificial Intelligence

UMKC announces a new Master of Science in Artificial Intelligence program aimed at addressing workforce demand for AI expertise, set to l...

AI News - General · 4 min · about 1 hour ago

Machine Learning

[D] Looking for definition of open-world ish learning problem

Hello! Recently I did a project where I initially had around 30 target classes. But at inference, the model had to be able to handle a lo...

Reddit - Machine Learning · 1 min · about 2 hours ago

Machine Learning

Mystery Shopping Meets Machine Learning: Can Algorithms Become the Ultimate Customer Experience Auditor?

Customer expectations across Africa are shifting faster than most organisations can track. A single inconsistent interaction can ignite a...

AI News - General · 8 min · about 2 hours ago

Machine Learning

GitHub to Use User Data for AI Training by Default

submitted by /u/i-drake [link] [comments]

Reddit - Artificial Intelligence · 1 min · about 3 hours ago

[2411.15087] Phrase-Instance Alignment for Generalized Referring Segmentation

About this article

Related Articles

UMKC Announces New Master of Science in Artificial Intelligence

[D] Looking for definition of open-world ish learning problem

Mystery Shopping Meets Machine Learning: Can Algorithms Become the Ultimate Customer Experience Auditor?

GitHub to Use User Data for AI Training by Default

No comments

Stay updated with AI News