Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
文献紹介:A Latent Variable Recurrent Neural Network...
Search
Sponsored
·
Your Podcast. Everywhere. Effortlessly.
Share. Educate. Inspire. Entertain. You do you. We'll handle the rest.
→
Atom
October 21, 2019
86
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
文献紹介:A Latent Variable Recurrent Neural Network for Discourse Relation Language Models
Atom
October 21, 2019
More Decks by Atom
See All by Atom
Tauri v2で作る⼀括画像ブラー / Bulk image blur created with Tauri v2
roraidolaurent
0
16
YouTubeのチャット欄の配置変更 / Changing the layout of the YouTube chat field
roraidolaurent
0
43
文献紹介 / Structure-based Knowledge Tracing: An Influence Propagation View
roraidolaurent
0
130
文献紹介 / Knowledge Tracing with GNN
roraidolaurent
0
110
文献紹介 / Non-Intrusive Parametric Reduced Order Models withHigh-Dimensional Inputs via Gradient-Free Active Subspace
roraidolaurent
0
79
ニューラルネットワークのベイズ推論 / Bayesian inference of neural networks
roraidolaurent
2
2.9k
Graph Convolutional Networks
roraidolaurent
0
270
文献紹介 / A Probabilistic Annotation Model for Crowdsourcing Coreference
roraidolaurent
0
100
文献紹介Deep Temporal-Recurrent-Replicated-Softmax for Topical Trends over Time
roraidolaurent
0
150
Featured
See All Featured
Designing Powerful Visuals for Engaging Learning
tmiket
1
550
The SEO Collaboration Effect
kristinabergwall1
1
560
Groundhog Day: Seeking Process in Gaming for Health
codingconduct
0
350
XXLCSS - How to scale CSS and keep your sanity
sugarenia
250
1.3M
Large-scale JavaScript Application Architecture
addyosmani
515
110k
Leveraging LLMs for student feedback in introductory data science courses - posit::conf(2025)
minecr
1
390
世界の人気アプリ100個を分析して見えたペイウォール設計の心得
akihiro_kokubo
PRO
74
42k
Prompt Engineering for Job Search
mfonobong
0
450
The Spectacular Lies of Maps
axbom
PRO
1
990
Impact Scores and Hybrid Strategies: The future of link building
tamaranovitovic
0
440
Designing for humans not robots
tammielis
254
26k
Bridging the Design Gap: How Collaborative Modelling removes blockers to flow between stakeholders and teams @FastFlow conf
baasie
0
700
Transcript
A Latent Variable Recurrent Neural Network for Discourse Relation Language
Models 文献紹介 2019/10/21 長岡技術科学大学 自然言語処理研究室 吉澤 亜斗武
Abstract ・単語のシーケンスや隣接する文の潜在的な談話関係を モデル化する潜在変数RNN(LVRNN)を提案 ・談話関係を潜在変数で表し,タスクに応じて予測または 周辺化することが可能 ・談話関係の分類,対話行為の分類,談話における言語モデルの タスクで先行研究よりも優れていることを示した. 2
1. Introduction ・ニューラルモデルは確率的グラフィカルモデルと比べ, 柔軟性がない. ・先行研究では,きれいに複数の言語を扱うモデルを扱えている. ・確率的グラフィカルモデルは層が多すぎるとtrainが困難 ・RNN言語モデルと談話関係を表す潜在変数モデルを 組み合わせたハイブリッドモデルを提案 3
1. Introduction ・また,提案モデルはVAE を必要とするRNNの複雑なモデル でなく,実装及びトレーニングが簡単である. ・提案モデルでは浅い談話関係に焦点を当てており, 談話全体の内容を補足していない. ・先行研究より談話関係分類,対話行為分類においては有効 ・提案モデルは当時のSotAよりも優れている. 4
2. Background 5 RNNLM token in a sentence by ,
∈ 1 … and = , ∈ 1…
2. Background 6 RNNLMの欠点の一つは文間の情報を伝搬できない. Document Context Language Model (DCLM) −1
:前の文の最後の隠れ状態
3.1 Discourse Relation Language Models 7 浅い談話関係をもつ潜在変数 を導入
3.1 Discourse Relation Language Models 8 潜在変数 はコンテキスト情報のベクトルの要約
3.2 Inference 9 談話関係は少数なので推論が簡単に
3.3 Learning 10 Joint likelihood objective : 言語モデルと談話関係予測のタスク Conditional objective:談話関係予測のタスク
4.1 Data 11 ・Penn Discourse Treebank (PDTB) annotated on a
corpus of Wall Street Journal acticles ・ Switchboard dialogue act corpus (SWDA) annotated on a collections of phone conversations 両方とも談話関係と対話関係の注釈が含まれている.
4.2 Implementation 12 詳細は論文で ・単層LSTM ・初期化:ランダム(ただし, は別途設定) ・学習:AdaGrad 初期学習率λ=0.1, ドロップアウトτ=0.5
・ハイパーパラメータ:次元数などはグリッドサーチ
5.1 Implicit discourse relation prediction on the PDTB 13 両方の提案手法が既存の
手法よりも優れた結果に. 二項検定の結果も良い
5.2 Dialogue Act tagging 14 精度が既存のものよりもよ く,二項検定も良い結果に (F1非公開)
5.3 Discourse-aware language modeling 15
5.3 Discourse-aware language modeling 16 ・ベースラインに談話関連情報を追加することで, 談話関係の曖昧さを解消がおき,優れた結果になった. ・トレーニングに談話注釈が必要なため大規模なデータセットに 対応した言語モデリングではない. ・談話関係は周辺化しているので,
もっと良いトレーニング方法があるのではと考察
7 Conclusion 17 ・隣接するシーケンス間の浅い談話の関係に関する 確率的ニューラルモデルを提案 ・確率的表現を維持しながら識別訓練されたベクトル表現を学習 ・2つの談話関係検出タスクでStoAよりも優れており, 言語モデルとしても適用できることがわかった. ・モデルのスケールアップが今後の課題