AI Dynamics

Global AI News Aggregator

About

X2SAM: Image and Video Segmentation by Conversation

What if you could segment anything in both images and videos using just a conversation? Researchers from Sun Yat-Sen University, Peng Cheng Laboratory, and Meituan present X2SAM. It pairs a large language model with a Mask Memory module to generate temporally consistent masks

→ View original post on X — @jiqizhixin