A Unified Framework for Tracking Based Text Detection and Recognition from Web Videos
Tian, Shu; Yin, Xu-Cheng; Su, Ya; Hao, Hong-Wei
刊名IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE
2018-03-01
卷号40期号:3页码:542-554
关键词Video Text Extraction Text Tracking Tracking Based Text Detection Tracking Based Text Recognition Embedded Captions
DOI10.1109/TPAMI.2017.2692763
文献子类Article
英文摘要Video text extraction plays an important role for multimedia understanding and retrieval. Most previous research efforts are conducted within individual frames. A few of recent methods, which pay attention to text tracking using multiple frames, however, do not effectively mine the relations among text detection, tracking and recognition. In this paper, we propose a generic Bayesian-based framework of Tracking based Text Detection And Recognition (T(2)DAR) from web videos for embedded captions, which is composed of three major components, i.e., text tracking, tracking based text detection, and tracking based text recognition. In this unified framework, text tracking is first conducted by tracking-by-detection. Tracking trajectories are then revised and refined with detection or recognition results. Text detection or recognition is finally improved with multi-frame integration. Moreover, a challenging video text (embedded caption text) database (USTB-VidTEXT) is constructed and publicly available. A variety of experiments on this dataset verify that our proposed approach largely improves the performance of text detection and recognition from web videos.
WOS关键词NATURAL SCENE IMAGES ; READING TEXT ; SEGMENTATION ; EXTRACTION
WOS研究方向Computer Science ; Engineering
语种英语
WOS记录号WOS:000424465900003
内容类型期刊论文
源URL[http://ir.ia.ac.cn/handle/173211/40787]  
专题数字内容技术与服务研究中心_听觉模型与认知计算
推荐引用方式
GB/T 7714
Tian, Shu,Yin, Xu-Cheng,Su, Ya,et al. A Unified Framework for Tracking Based Text Detection and Recognition from Web Videos[J]. IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE,2018,40(3):542-554.
APA Tian, Shu,Yin, Xu-Cheng,Su, Ya,&Hao, Hong-Wei.(2018).A Unified Framework for Tracking Based Text Detection and Recognition from Web Videos.IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE,40(3),542-554.
MLA Tian, Shu,et al."A Unified Framework for Tracking Based Text Detection and Recognition from Web Videos".IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE 40.3(2018):542-554.
个性服务
查看访问统计
相关权益政策
暂无数据
收藏/分享
所有评论 (0)
暂无评论
 

除非特别说明,本系统中所有内容都受版权保护,并保留所有权利。


©版权所有 ©2017 CSpace - Powered by CSpace