| 講演抄録/キーワード |
| 講演名 |
2018-12-13 10:15
情景内文字のCNNによる拡大 ○中村俊貴・内田誠一(九大) PRMU2018-76 |
| 抄録 |
(和) |
本研究は,畳み込みニューラルネットワーク (CNN) を用い,End-to-Endで情景画像内の文字を拡大することを目的とする.
教師画像として重心を変えずに文字を拡大した画像を用い,畳み込み・逆畳み込み層から構成されるEncoder-Decoder型CNNにより学習を行った.
さらに,文字拡大のタスクを文字隠蔽・文字抽出・文字拡大・画像合成の4つに分割し,それぞれに対する4つのCNNを学習しそれらを結合することで文字の拡大を行う手法を試みた.
加えて,これら4つのCNNを結合し,重みを再利用するFine tuningを行ったCNNでも拡大を行った.
また,これらの結果に対して,文字検出を用いた文字拡大手法との比較を行い,定量的に評価した. |
| (英) |
The purpose of this research is magnifying scene texts using convolutional neural network (CNN) with end-to-end.
We trained Encoder-Decoder type CNN composed of convolution/deconvolution using images magnified scene text without changing the center of gravity as teacher image.
Also, we tried a method of magnifying scene texts by dividing the task into 4 of text erasing, text extraction, text magnifying, and image merging, and learning 4 CNNs for each, and combining them.
Furthermore, we merging 1 CNN these 4 CNNs and fine tuned by reusing weights.
In addition, we compare these results with text magnifying method using character detection and evaluated it quantitatively. |
| キーワード |
(和) |
深層学習 / 情景内文字 / 文字拡大 / / / / / |
| (英) |
/ / / / / / / |
| 文献情報 |
信学技報, vol. 118, no. 362, PRMU2018-76, pp. 7-12, 2018年12月. |
| 資料番号 |
PRMU2018-76 |
| 発行日 |
2018-12-06 (PRMU) |
| ISSN |
Online edition: ISSN 2432-6380 |
著作権に ついて |
技術研究報告に掲載された論文の著作権は電子情報通信学会に帰属します.(許諾番号:10GA0019/12GB0052/13GB0056/17GB0034/18GB0034) |
| PDFダウンロード |
PRMU2018-76 |