Національний репозитарій академічних текстів

Розширений пошук
База даних НРАТ:
Звіти у сфері наукової і науково-технічної діяльності
0
Загальна кількість
0
Повні тексти
Дисертації на здобуття наукових ступенів та автореферати
0
Загальна кількість
0
Повні тексти
Матеріали видань та локальних репозитаріїв
0
Кількість локальних репозитаріїв
0
Повні тексти

Пошук академічних текстів

Знайдено документів: 1

Інформація × Реєстраційний номер 2123U011196, Матеріали видань та локальних репозитаріїв Категорія Препринт Назва роботи Global Motion Understanding in Large-Scale Video Object Segmentation Автор Fedynyak VolodymyrFedynyak Volodymyr Дата публікації 01-01-2023 Постачальник інформації Український католицький університет Першоджерело https://hdl.handle.net/20.500.14570/4422 Видання Опис In this thesis, we show that transferring knowledge from other domains of video understanding combined with large-scale learning can improve robustness of Video Object Segmentation (VOS) under complex circumstances. Namely, we focus on integrating scene global motion knowledge to improve large-scale semi-supervised Video Object Segmentation. Prior works on VOS mostly rely on direct comparison of semantic and contextual features to perform dense matching between current and past frames, passing over actual motion structure. On the other hand, Optical Flow Estimation task aims to approximate the scene motion field, exposing global motion patterns which are typically undiscoverable during all pairs similarity search. We present WarpFormer, an architecture for semi-supervised Video Object Segmentation that exploits existing knowledge in motion understanding to conduct smoother propagation and more accurate matching. Our framework employs a generic pretrained Optical Flow Estimation network whose prediction is used to warp both past frames and instance segmentation masks to the current frame domain. Consequently, warped segmentation masks are refined and fused together aiming to inpaint occluded regions and eliminate artifacts caused by flow field imperfects. Additionally, we employ novel large-scale MOSE 2023 dataset to train model on various complex scenarios. Our method demonstrates strong performance on DAVIS 2016/2017 validation (93.0% and 85.9%), DAVIS 2017 test-dev (80.6%) and YouTube-VOS 2019 validation (83.8%) that is competitive with alternative state-of-the-art methods while using much simpler memory mechanism and instance understanding logic. Додано в НРАТ 2025-11-05 Закрити

Матеріали

Препринт

Global Motion Understanding in Large-Scale Video Object Segmentation

Fedynyak Volodymyr. Global Motion Understanding in Large-Scale Video Object Segmentation : публікація 2023-01-01; Український католицький університет, 2123U011196

Знайдено документів: 1

Оновлено: 2026-04-07

Роздрукувати цю сторінку

Звіти у сфері наукової і науково-технічної діяльності

Дисертації на здобуття наукових ступенів та автореферати

Матеріали видань та локальних репозитаріїв