A Lightweight Framework for Audio-Visual Segmentation with an Audio-Guided Space–Time Memory Network

As a multimodal fusion task, audio-visual segmentation (AVS) aims to locate sounding objects at the pixel level within a given image. This capability holds significant importance and practical value in applications such as intelligent surveillance, multimedia content analysis, and human–robot intera...

Full description

Saved in:
Bibliographic Details
Main Authors: Yunpeng Zuo, Yunwei Zhang
Format: Article
Language:English
Published: MDPI AG 2025-06-01
Series:Applied Sciences
Subjects:
Online Access:https://www.mdpi.com/2076-3417/15/12/6585
Tags: Add Tag
No Tags, Be the first to tag this record!