| 講演抄録/キーワード |
| 講演名 |
2013-06-10 13:30
映像データにおける局所特徴のバースト性を考慮したトピックモデリング ○謝 洋・江口浩二(神戸大) PRMU2013-20 |
| 抄録 |
(和) |
映像データを構成するキーフレーム画像における視覚単語とそれに対応する発話単語に着目し,それらを統合するトピックモデルCorr-DCMLDAを提案する.Corr-DCMLDAは2つの特徴を持つ. 一つは,発話単語の潜在トピックは視覚単語における潜在トピックに依存して決まる点である.もう一つは,同一の潜在トピックに関して映像ごとに異なる視覚単語や発話単語が用いられる傾向を考慮する点である.ネット映像データを用いたジャンル分類に関してCorr-DCMLDAが従来手法よりも有効であることを示す. |
| (英) |
In this paper we propose a topic model, Corr-DCMLDA, which can integrate visual words and the corresponding speech transcript words for key-frame images in each video data. Corr-DCMLDA has two distinguishing characteristics. One is the generative process that it first generates topics for visual words and then only the topics associated with the visual words are used to generate speech transcript words. The other is the assumption that the same latent topics are represented with different visual words and speech transcript words in different videos. We demonstrate through experiments on genre classification with online videos that Corr-DCMLDA works more effectively than some prior models. |
| キーワード |
(和) |
映像トピックモデリング / Dirichlet compound multinomialモデル / 周辺化ギブスサンプリング / / / / / |
| (英) |
Video topic modeling / Dirichlet compound multinomial models / collapsed Gibbs sampling / / / / / |
| 文献情報 |
信学技報, vol. 113, no. 75, PRMU2013-20, pp. 5-10, 2013年6月. |
| 資料番号 |
PRMU2013-20 |
| 発行日 |
2013-06-03 (PRMU) |
| ISSN |
Print edition: ISSN 0913-5685 Online edition: ISSN 2432-6380 |
著作権に ついて |
技術研究報告に掲載された論文の著作権は電子情報通信学会に帰属します.(許諾番号:10GA0019/12GB0052/13GB0056/17GB0034/18GB0034) |
| PDFダウンロード |
PRMU2013-20 |