Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
Autoencoding Variational Inference for Topic Mo...
Search
Kento Nozawa
June 15, 2017
Research
30k
3
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
Autoencoding Variational Inference for Topic Modelsの解説スライド
ICLR2017読み会のスライド
https://connpass.com/event/57631/
Kento Nozawa
June 15, 2017
More Decks by Kento Nozawa
See All by Kento Nozawa
[最先端NLP勉強会2026] Checklists Are Better Than Reward Models For Aligning Language Models
nzw0301
1
320
Analysis on Negative Sample Size in Contrastive Unsupervised Representation Learning
nzw0301
0
240
[IJCAI-ECAI 2022] Evaluation Methods for Representation Learning: A Survey
nzw0301
0
710
[NeurIPS Japan meetup 2021 talk] Understanding Negative Samples in Instance Discriminative Self-supervised Representation Learning
nzw0301
0
280
[IBIS2021] 対照的自己教師付き表現学習おける負例数の解析
nzw0301
0
230
Understanding Negative Samples in Instance Discriminative Self-supervised Representation Learning
nzw0301
0
590
Introduction of PAC-Bayes and its Application for Contrastive Unsupervised Representation Learning
nzw0301
2
920
NLP Tutorial; word representation learning
nzw0301
0
270
Analyzing Centralities of Embedded Nodes
nzw0301
0
240
Other Decks in Research
See All in Research
Kaggle|AI Agent Security — しくじり先生、俺みたいになるな
pomcho555
1
120
PHTalks Bengaluru - SSRF When All Else Fails
dk999
0
1.2k
Anthropic が提案する LLM の内部状態を自然言語で説明可能にした Natural Language Autoencoders / Natural Language Autoencoders Produce Unsupervised Explanations of LLM Activations
shunk031
0
320
学術バーQ AI研究最前線:自己教師あり学習による画像モデルの事前学習
naok615
0
130
Karkada さんの論文 × 2 の紹介: (1) Closed-Form Training Dynamics Reveal Learned Features and Linear Structure in Word2Vec-like Models, (2) Symmetry in language statistics shapes the geometry of model representations
eumesy
PRO
1
800
[IR Reading 2026春 論文紹介] LLM-based Listwise Reranking under the Effect of Positional Bias (ECIR 2026) /IR-Reading-2026-Spring
koheishinden
PRO
0
450
超効率化への挑戦:1bit LLMの現状と展望
yumaichikawa
0
780
La génomique au service de la fromageabilité du lait grâce aux spectres MIR
institudelelevage
PRO
0
140
論文紹介: Understanding Epistemic Language with a Language-augmented Bayesian Theory of Mind
hisaokatsumi
0
170
2026年 オープンキャンパス 研究室紹介
junkurihara
0
220
[CV勉強会@関東 CVPR2026] PSDesigner: Automated Graphic Design with a Human-Like Creative Workflow / kantocv 67th CVPR 2026
shunk031
0
350
第64回CV・PRML勉強会 論文紹介:Linguistic Priors for Visual Decoupling: Towards Symmetric Vision-Brain Alignment
sokikatayama
0
230
Featured
See All Featured
Claude Code のすすめ
schroneko
67
230k
Creating an realtime collaboration tool: Agile Flush - .NET Oxford
marcduiker
35
2.6k
HDC tutorial
michielstock
2
930
How to build a perfect <img>
jonoalderson
1
6.1k
Embracing the Ebb and Flow
colly
88
5.2k
A Soul's Torment
seathinner
8
3.7k
Documentation Writing (for coders)
carmenintech
77
5.6k
How to Build an AI Search Optimization Roadmap - Criteria and Steps to Take #SEOIRL
aleyda
1
2.2k
Practical Tips for Bootstrapping Information Extraction Pipelines
honnibal
25
2.1k
Conquering PDFs: document understanding beyond plain text
inesmontani
PRO
4
3.1k
Utilizing Notion as your number one productivity tool
mfonobong
4
610
The Curious Case for Waylosing
cassininazir
1
550
Transcript
Autoencoding Variational Inference For Topic Models Akash Srivastava and Charles
Sutton ICLR2017ಡΈձ ಡΉਓ: @nzw0301
֓ཁ 1. Latent Dirichlet Allocation (LDA) ΛNeural Variational Inference (NVI)
Ͱ • Dirichlet ͷ reparameterization trick 2. ৽ϞσϧͷఏҊ 3. ѱ͍ہॴղʹϋϚΔͷΛ༧ 2
ࣄલࣝɿLDAͱVAEͷ֓ཁ 3
LDA จॻͷ֬తੜϞσϧ [Blei et al., 2003]
จॻͷτϐοΫQ [cВ ݚڀ ՝ ࣝ Պֶऀ ʜ ػցֶश ਓೳ Ϟσϧ αϯϓϧ ʜ τϐοΫͷ୯ޠ p(w|β) Ќ Ќ ػցֶश ػցֶशݚڀ ਓೳ՝ Ϟσϧ-%" Պֶֶण࢘ ίʔύε 4
VAE: Encoder • NNΛͬͨੜϞσϧ • Encoder: • σʔλ͔Β֬ͷύϥϝʔλͷม • ֬જࡏมΛੜ
• Decoder: • જࡏม͔Βσʔλੜ • Reparameterization trick • BPʹαϯϓϧΛؚΊΔ • ඪ४ਖ਼نͷαϯϓϧͱͷ ύϥϝʔλ͔ΒαϯϓϧΛߏ 5
VAE: Decoder • NNΛͬͨੜϞσϧ • Encoder: • σʔλ͔Β֬ͷύϥϝʔλͷม • ֬જࡏมΛੜ
• Decoder: • જࡏม͔Βσʔλੜ • Reparameterization trick • BPʹαϯϓϧΛؚΊΔ • ඪ४ਖ਼نͷαϯϓϧͱͷ ύϥϝʔλ͔ΒαϯϓϧΛߏ 6
VAE: Reparameterization trick • NNΛͬͨੜϞσϧ • Encoder: • σʔλ͔Β֬ͷύϥϝʔλͷม •
֬જࡏมΛੜ • Decoder: • જࡏม͔Βσʔλੜ • Reparameterization trick • BPʹαϯϓϧΛؚΊΔ • ඪ४ਖ਼نͷαϯϓϧͱͷ ύϥϝʔλ͔ΒαϯϓϧΛߏ 7
VAE: ϩεؔ 8 L (⇥) = D X d=1 (
1 2 ⇣ tr (⌃0) + µT 0 µ0 K log | ⌃0 | ⌘ + E ✏⇠N (0,1) ⇣ log p xd |f ( µ0 + ⌃ 1/2 0 ✏ ) ⌘ ) (Ⅰ) ࣄલͱͷKLμΠόʔδΣϯε (Ⅱ) ର ࣜશମ: Evidence Lower Bound (I) (Ⅱ)
ຊ 9
Reparameterization trick for Dirichlet Distribution • LDAͷθ: Dirichlet͔Βαϯϓϧ • Scale
family DistributionͰͳ͍ͨΊɼߏͰ͖ͳ͍ 10 จॻͷτϐοΫQ [cВ
Reparameterization trick for Dirichlet Distribution • LDAͷθ: Dirichlet͔Βαϯϓϧ • Scale
family DistributionͰͳ͍ͨΊɼߏͰ͖ͳ͍ • Laplace approximation • ਖ਼نͷαϯϓϧʹsoftmaxؔΛద༻ͯ͠༻ • ࣄલͷύϥϝʔλɿ µk = log( ↵k) 1 K K X i=1 log ↵i ⌃k,k = 1 ↵k (1 2 K ) + 1 K2 K X i=1 1 ↵k 11
ωοτϫʔΫͱϩεؔ 12 X encoder µ( X ) ⌃ ( X
) KL {N( z ; µ( X ) , ⌃ ( X ))||N( z ; µ1, ⌃1)} ✏ ⇠ N(✏; 0, I ) + decoder: f ( Z ) loss ( x, f ( Z )) • σ: softmaxؔ • β : DecoderͷॏΈʢunnormalizedʣ • σ(β): ୯ޠͷDiriclet͔ΒͷαϯϓϧʹରԠ L ( ⇥ ) = D X d=1 ( 1 2 ⇣ tr ( ⌃ 1 1 ⌃0) + ( µ1 µ0) T ⌃ 1 1 ( µ1 µ0) K + log |⌃1 | |⌃0 | ⌘ + E ✏⇠N (0,1) wt d log ⇣ ( µ0 + ⌃1/2 0 ✏ ) ⌘ !) θ සϕΫτϧ
prodLDA: ఏҊϞσϧ • Products of Experts • βͱθͷੵʹsoftmaxؔ 13 L
( ⇥ ) = D X d=1 ( 1 2 ⇣ tr ( ⌃ 1 1 ⌃0) + ( µ1 µ0) T ⌃ 1 1 ( µ1 µ0) K + log |⌃1 | |⌃0 | ⌘ + E ✏⇠N (0,1) wt d log ⇣ ( µ0 + ⌃1/2 0 ✏ ) ⌘ !) ( ✓)
࠷దԽͱωοτϫʔΫͷ NVIͷɿ ֶशͷॳظஈ֊Ͱlocal optimumʹߦ͖͍͢ • AdamͷύϥϝʔλΛௐ • ηͱβ1 ͷͷߴΊʹઃఆ •
Batch NormalizationͱDropoutΛ༻ 14
࣮ݧ 1. CoherenceͱPerplexity • ޙड़ 2. ֶशͱࣄલΛม͑ͨͱ͖ͷޮՌ • ߴֶ͍श &
Dirichlet͕ϕλʔ 3. ςετσʔλʹର͢Δ࠷దԽͷ༗ແ • ͠ͳ͍͍ͯ͘ 4. p(w|β)ͷϦετ • লུ 15
Coherence 16 දจ͔ΒҾ༻ • LDA VAE: ఏҊਪ๏ • prodLDA: ఏҊਪ๏+ఏҊϞσϧ
• LDA DMFVI: Online Mean-Field Variational Inference • NVDM: VAEϕʔεͷจॻϞσϦϯά දͷ: 40ճ࣮ߦͯ͠ࢉग़
Perplexity 17 දจ͔ΒҾ༻
ϨϏϡʔ: ؾʹͳͬͨͷΛ͍͔ͭ͘ Q1. NVDMͰadamͷֶशΛม͑ͨํ͕ެฏ A1. จʹө Q2. ϋΠύʔύϥϝʔλ࠷దԽ͔ͨ͠ A2. ൺֱख๏͍ͯ͠ΔɼఏҊख๏BO
Rating: 6-7-6-5 18
ͦͷଞ • ஶऀ࣮: TensorFlow • NVDMͷஶऀΒͷ৽Ϟσϧ͕ICML2017ʹ࠾ 19