| 講演抄録/キーワード |
| 講演名 |
2016-10-20 14:30
[招待講演]くずし字・古文書の自動解読にむけて ○寺沢憲吾(公立はこだて未来大) PRMU2016-92 |
| 抄録 |
(和) |
講演者はこれまで,毛筆手書き文書や古活字文書などの,通常のOCR(光学文字認識)の適用が困難であるような文書に対し,OCRとは異なるアプローチで人間の理解を促進するためのツールを開発する研究を行ってきた.毛筆手書き文字を対象としたワードスポッティングの研究,未翻刻の古活字を対象とした頻出語・重要語の抽出の研究を中心として,日本語古典籍を対象とした自動文字認識も現在は視野に入れている.本稿では,講演者らのこうした研究の中から,最も基本となるワードスポッティングの手法について詳しく解説するとともに,現在と今後の研究の展望についても述べる. |
| (英) |
We have been studying the ways to facilitate utilities of information technologies for the analysis and understanding of historical documents for which the ordinary OCR method does not work well. Our main study topics include word spotting for cursive scripts, frequent or important word extraction from machine-unreadable scripts, automatic character recognition for ancient documents, and so on. In this article, we will give an explanation of our current method for word spotting. In addition, our prospects for further studies will also be described. |
| キーワード |
(和) |
くずし字 / 古文書 / 歴史的文書画像処理 / 文書画像認識 / ワードスポッティング / / / |
| (英) |
Historical or Cursive Document Processing / Document Image Recognition / Word Spotting / / / / / |
| 文献情報 |
信学技報, vol. 116, no. 259, PRMU2016-92, pp. 9-14, 2016年10月. |
| 資料番号 |
PRMU2016-92 |
| 発行日 |
2016-10-13 (PRMU) |
| ISSN |
Print edition: ISSN 0913-5685 Online edition: ISSN 2432-6380 |
著作権に ついて |
技術研究報告に掲載された論文の著作権は電子情報通信学会に帰属します.(許諾番号:10GA0019/12GB0052/13GB0056/17GB0034/18GB0034) |
| PDFダウンロード |
PRMU2016-92 |