通过图像合成和域对抗学习实现自监督视觉跟踪

Self-Supervised Visual Tracking via Image Synthesis and Domain Adversarial Learning.

作者信息

Geng Gu, Zhou Sida, Tang Jianing, Zhang Xinming, Liu Qiao, Yuan Di

机构信息

Guangzhou Institute of Technology, Xidian University, Guangzhou 510555, China.

School of Electrical and Information Engineering, Yunnan Minzu University, Kunming 650504, China.

出版信息

Sensors (Basel). 2025 Jul 25;25(15):4621. doi: 10.3390/s25154621.

DOI:10.3390/s25154621

PMID:40807783

原文链接:https://pmc.ncbi.nlm.nih.gov/articles/PMC12349480/

Abstract

With the widespread use of sensors in applications such as autonomous driving and intelligent security, stable and efficient target tracking from diverse sensor data has become increasingly important. Self-supervised visual tracking has attracted increasing attention due to its potential to eliminate reliance on costly manual annotations; however, existing methods often train on incomplete object representations, resulting in inaccurate localization during inference. In addition, current methods typically struggle when applied to deep networks. To address these limitations, we propose a novel self-supervised tracking framework based on image synthesis and domain adversarial learning. We first construct a large-scale database of real-world target objects, then synthesize training video pairs by randomly inserting these targets into background frames while applying geometric and appearance transformations to simulate realistic variations. To reduce domain shift introduced by synthetic content, we incorporate a domain classification branch after feature extraction and adopt domain adversarial training to encourage feature alignment between real and synthetic domains. Experimental results on five standard tracking benchmarks demonstrate that our method significantly enhances tracking accuracy compared to existing self-supervised approaches without introducing any additional labeling cost. The proposed framework not only ensures complete target coverage during training but also shows strong scalability to deeper network architectures, offering a practical and effective solution for real-world tracking applications.

摘要

随着传感器在自动驾驶和智能安全等应用中的广泛使用，从各种传感器数据中进行稳定高效的目标跟踪变得越来越重要。自监督视觉跟踪因其有潜力消除对昂贵人工标注的依赖而受到越来越多的关注；然而，现有方法通常在不完整的对象表示上进行训练，导致推理过程中的定位不准确。此外，当前方法在应用于深度网络时通常会遇到困难。为了解决这些限制，我们提出了一种基于图像合成和域对抗学习的新型自监督跟踪框架。我们首先构建一个大规模的真实世界目标对象数据库，然后通过将这些目标随机插入背景帧中，同时应用几何和外观变换来模拟现实变化，从而合成训练视频对。为了减少合成内容引入的域转移，我们在特征提取后加入一个域分类分支，并采用域对抗训练来促进真实域和合成域之间的特征对齐。在五个标准跟踪基准上的实验结果表明，与现有的自监督方法相比，我们的方法在不引入任何额外标注成本的情况下显著提高了跟踪精度。所提出的框架不仅在训练期间确保了目标的完整覆盖，而且对更深的网络架构显示出强大的可扩展性，为实际的跟踪应用提供了一种实用有效的解决方案。

https://cdn.ncbi.nlm.nih.gov/pmc/blobs/ebe5/12349480/46b4b1c5337d/sensors-25-04621-g001.jpg

相似文献

Self-Supervised Visual Tracking via Image Synthesis and Domain Adversarial Learning.

Sensors (Basel). 2025 Jul 25;25(15):4621. doi: 10.3390/s25154621.

Prescription of Controlled Substances: Benefits and Risks

A medical image classification method based on self-regularized adversarial learning.

Med Phys. 2024 Nov;51(11):8232-8246. doi: 10.1002/mp.17320. Epub 2024 Jul 30.

Integrated neural network framework for multi-object detection and recognition using UAV imagery.

Front Neurorobot. 2025 Jul 30;19:1643011. doi: 10.3389/fnbot.2025.1643011. eCollection 2025.

Short-Term Memory Impairment

Comparison of self-administered survey questionnaire responses collected using mobile apps versus other methods.

Cochrane Database Syst Rev. 2015 Jul 27;2015(7):MR000042. doi: 10.1002/14651858.MR000042.pub2.

Leveraging a foundation model zoo for cell similarity search in oncological microscopy across devices.

Front Oncol. 2025 Jun 18;15:1480384. doi: 10.3389/fonc.2025.1480384. eCollection 2025.

Boundary-aware information maximization for self-supervised medical image segmentation.

Med Image Anal. 2024 May;94:103150. doi: 10.1016/j.media.2024.103150. Epub 2024 Mar 28.

An open-source deep learning framework for respiratory motion monitoring and volumetric imaging during radiation therapy.

Med Phys. 2025 Jul;52(7):e18015. doi: 10.1002/mp.18015.

The Lived Experience of Autistic Adults in Employment: A Systematic Search and Synthesis.

Autism Adulthood. 2024 Dec 2;6(4):495-509. doi: 10.1089/aut.2022.0114. eCollection 2024 Dec.

本文引用的文献

A Target Tracking Method Based on a Pyramid Channel Attention Mechanism.

Sensors (Basel). 2025 May 20;25(10):3214. doi: 10.3390/s25103214.

Enhanced YOLOv5: An Efficient Road Object Detection Method.

Sensors (Basel). 2023 Oct 10;23(20):8355. doi: 10.3390/s23208355.

CoSOV1Net: A Cone- and Spatial-Opponent Primary Visual Cortex-Inspired Neural Network for Lightweight Salient Object Detection.

Sensors (Basel). 2023 Jul 17;23(14):6450. doi: 10.3390/s23146450.

Self-Supervised Tracking via Target-Aware Data Synthesis.

IEEE Trans Neural Netw Learn Syst. 2024 Jul;35(7):9186-9197. doi: 10.1109/TNNLS.2022.3231537. Epub 2024 Jul 10.

A Video Target Tracking and Correction Model with Blockchain and Robust Feature Location.

Sensors (Basel). 2023 Feb 22;23(5):2408. doi: 10.3390/s23052408.

Self-Supervised Deep Correlation Tracking.

IEEE Trans Image Process. 2021;30:976-985. doi: 10.1109/TIP.2020.3037518. Epub 2020 Dec 9.

GOT-10k: A Large High-Diversity Benchmark for Generic Object Tracking in the Wild.

IEEE Trans Pattern Anal Mach Intell. 2021 May;43(5):1562-1577. doi: 10.1109/TPAMI.2019.2957464. Epub 2021 Apr 1.

t-Distributed Stochastic Neighbor Embedding (t-SNE): A tool for eco-physiological transcriptomic analysis.

Mar Genomics. 2020 Jun;51:100723. doi: 10.1016/j.margen.2019.100723. Epub 2019 Nov 26.

Synthetic Data Generation for End-to-End Thermal Infrared Tracking.

IEEE Trans Image Process. 2019 Apr;28(4):1837-1850. doi: 10.1109/TIP.2018.2879249. Epub 2018 Nov 2.

Object Tracking Benchmark.

IEEE Trans Pattern Anal Mach Intell. 2015 Sep;37(9):1834-48. doi: 10.1109/TPAMI.2014.2388226.

文献AI研究员

20分钟写一篇综述，助力文献阅读效率提升50倍。

立即体验

用中文搜PubMed

大模型驱动的PubMed中文搜索引擎

马上搜索

文档翻译

学术文献翻译模型，支持多种主流文档格式。

立即体验

通过图像合成和域对抗学习实现自监督视觉跟踪

Self-Supervised Visual Tracking via Image Synthesis and Domain Adversarial Learning.

作者信息

机构信息

出版信息

相似文献

本文引用的文献

文献AI研究员

用中文搜PubMed

文档翻译

Suppr 超能文献

相似文献

本文引用的文献