《電子技術應用》
您所在的位置:首頁 > 其他 > 设计应用 > 基于分割的自然场景下文本检测方法与应用
基于分割的自然场景下文本检测方法与应用
2021年电子技术应用第2期
陈小顺,王良君
江苏大学 计算机科学与通信工程学院,江苏 镇江212013
摘要: 自然场景文本检测识别在智能设备中应用广泛,而对文本识别的第一步则是对文本进行精确的定位检测。对于现有像素分割方法PixelLink中存在的弯曲文本定位包含过多背景信息、检测图像后处理不足两个主要问题提出改进。引入特征通道注意力机制,关注生成特征图中特征通道间的权重关系,提升检测方法的鲁棒性。接着改变公开数据集标注形式,将坐标点表示为一串带有方向的序列形式,在LSTM模型中进行多边形框的学习与框定。最后在公开数据集和自建数据集上进行文本检测测试。实验表明,改进的检测方法在各数据集中表现优于原方法,与当前领先方法精度相近,能够在各个环境中完成对文本的检测功能。
中圖分類號: TN911.73;TP391.4
文獻標識碼: A
DOI:10.16157/j.issn.0258-7998.200316
中文引用格式: 陳小順,王良君. 基于分割的自然場景下文本檢測方法與應用[J].電子技術應用,2021,47(2):54-57.
英文引用格式: Chen Xiaoshun,Wang Liangjun. Text detection and application in natural scene based on segmentation[J]. Application of Electronic Technique,2021,47(2):54-57.
Text detection and application in natural scene based on segmentation
Chen Xiaoshun,Wang Liangjun
School of Computer Science and Telecommunication Engineering, Jiangsu University,Zhenjiang 212013,China
Abstract: Text recognition in nature scene is currently applied in various intelligence equipment. The first step of text recognition is to precisely locate the text. In the Pixel Link text location methods, there are mainly two problems: too much background information is incorporated in the text region, and the test accuracy is insufficient. Aiming at these issues, an improved text location method was proposed to precisely locate the text in the natural scene. At first, an attention mechanism was incorporated into the original network. By focusing on the weight relationship between feature channels in the generated feature map, one can improve the weight coefficient of effective feature channels, and suppress the weight of inefficient or invalid feature channels. In the second, by changing the form of data set annotation, the coordinate points can be expressed as a series of sequence forms, so that the text lines can be framed adaptively in the LSTM model. At last, the located object is rotated according to the angle between a pair of vertexes in the polygon frame, and is subsequently fed to the text recognition interface to obtain the final character. Finally, the text detection test is carried out on the open data set and self-built data set. The experimental results show that the improved detection method is superior to the original method on different dataset, and the accuracy is similar to the current leading method.
Key words : pixel segmentation;attention mechanism;LSTM;natural scene text detection

0 引言

    視覺圖像是人們獲取外界信息的主要來源,文本則是對事物的一種凝練描述,人通過眼睛捕獲文本獲取信息,機器設備的眼睛則是冰冷的攝像頭。如何讓機器設備從拍照獲取的圖像中準確檢測識別文本信息逐漸為各界學者關注。

    現(xiàn)代文本檢測方法多為基于深度學習的方法,主要分為基于候選框和基于像素分割的兩種形式。本文選擇基于像素分割的深度學習模型作為文本檢測識別的主要研究方向,能夠同時滿足對自然場景文本的精確檢測,又能保證后續(xù)設備功能(如語義分析等功能)的拓展。




本文詳細內(nèi)容請下載:http://m.tom3567.com/resource/share/2000003385




作者信息:

陳小順,王良君

(江蘇大學 計算機科學與通信工程學院,江蘇 鎮(zhèn)江212013)

此內(nèi)容為AET網(wǎng)站原創(chuàng),未經(jīng)授權(quán)禁止轉(zhuǎn)載。
主站蜘蛛池模板: 97成人在线免费视频| 国产日韩久久| 久久精品国产成人精品| 久久天堂国产精品| 国产高清自拍99| 亚洲二区自拍| 九九九九免费视频| 国产精品自产拍在线观看中文| 久久久久久久久久久视频| 日韩欧美视频一区二区三区四区| 一区二区欧美日韩| 中文字幕久久一区| 国产日韩欧美影视| 亚洲欧美日韩在线综合| 日韩一级在线免费观看| 久久久视频精品| 亚洲中文字幕无码不卡电影| 久久精品人人做人人爽| 福利视频久久| 日本欧美精品久久久| 韩国视频理论视频久久| 成人a在线观看| 久久国产精品久久久久V| 欧洲中文字幕国产精品| 国产免费亚洲高清| 亚洲91精品在线观看| 九九精品在线观看| 国产精品久久国产精品99gif| 中文字幕欧美日韩一区二区三区| 91精品视频专区| 国产精品视频yy9099| 国产精品亚洲精品| 国产精品com| 日韩在线观看精品| 青青青在线观看视频| 国产精品免费在线播放| 国产精国产精品| 日本一区二区三区在线视频| 久久国产色av免费观看| 欧美激情 国产精品| 国外色69视频在线观看|