首页 /研究 /Bootstrapped Self-Supervised Training with Monocular Video for Semantic Segmentation and Depth Estimation

PERCEPTION

Bootstrapped Self-Supervised Training with Monocular Video for Semantic Segmentation and Depth Estimation

Yihao Zhang, John J. Leonard

发表年份: 2021
引用次数: 5

摘要

For a robot deployed in the world, it is desirable to have the ability of autonomous learning to improve its initial pre-set knowledge. We formalize this as a bootstrapped self-supervised learning problem where a system is initially bootstrapped with supervised training on a labeled dataset and we look for a self-supervised training method that can subsequently improve the system over the supervised training baseline using only unlabeled data. In this work, we leverage temporal consistency between frames in monocular video to per-form this bootstrapped self-supervised training. We show that a well-trained state-of-the-art semantic segmentation network can be further improved through our method. In addition, we show that the bootstrapped self-supervised training framework can help a network learn depth estimation better than pure supervised training or self-supervised training.

关键词

Leverage (statistics)Artificial intelligenceComputer scienceMachine learningSemi-supervised learningSupervised learningSegmentationMonocularTraining setConsistency (knowledge bases)

Bootstrapped Self-Supervised Training with Monocular Video for Semantic Segmentation and Depth Estimation

摘要

关键词

相关论文

Statistical Learning Theory

Artificial intelligence: a modern approach

Applied Nonlinear Control

A new optimizer using particle swarm theory