Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
Autoencoding Variational Inference for Topic Mo...
Search
Sponsored
·
Your Podcast. Everywhere. Effortlessly.
Share. Educate. Inspire. Entertain. You do you. We'll handle the rest.
→
Kento Nozawa
June 15, 2017
Research
30k
3
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
Autoencoding Variational Inference for Topic Modelsの解説スライド
ICLR2017読み会のスライド
https://connpass.com/event/57631/
Kento Nozawa
June 15, 2017
More Decks by Kento Nozawa
See All by Kento Nozawa
[最先端NLP勉強会2026] Checklists Are Better Than Reward Models For Aligning Language Models
nzw0301
1
310
Analysis on Negative Sample Size in Contrastive Unsupervised Representation Learning
nzw0301
0
230
[IJCAI-ECAI 2022] Evaluation Methods for Representation Learning: A Survey
nzw0301
0
700
[NeurIPS Japan meetup 2021 talk] Understanding Negative Samples in Instance Discriminative Self-supervised Representation Learning
nzw0301
0
270
[IBIS2021] 対照的自己教師付き表現学習おける負例数の解析
nzw0301
0
230
Understanding Negative Samples in Instance Discriminative Self-supervised Representation Learning
nzw0301
0
590
Introduction of PAC-Bayes and its Application for Contrastive Unsupervised Representation Learning
nzw0301
2
910
NLP Tutorial; word representation learning
nzw0301
0
270
Analyzing Centralities of Embedded Nodes
nzw0301
0
230
Other Decks in Research
See All in Research
全国町字単位空き家率推定データver1.0データ仕様
microbaseinc
0
210
論文紹介:Doc-to-LoRA: Learning to Instantly Internalize Contexts
yukako_nakano
0
150
typst の使い方:言語学を研究する学生のために
gitomochang
0
570
XDPerf: A High-Performance Traffic Generator Built with WASM and eBPF
takehaya
1
290
重要だけど測れていないもの:高齢者ケアの見えない課題
theoriatec2024
0
500
ふとした出会いで生まれたSkillが、 社内利用1位になるまで
mikimhk
21
23k
最先端NLP 2026 論文紹介: Wait, Wait, Wait... Why Do Reasoning Models Loop? / SNLP Paper Review: Wait, Wait, Wait... Why Do Reasoning Models Loop?
tkng
0
210
【ローカルAIに向き合う展示会vol.2】液体時間定数型モジュールを用いた オリジナルの双方向エンコーダーモデルNexteraBERT 推論速度向上検討並びにダウンストリーム評価
rikkabotan7
0
190
COMETAを用いたデータ民主化運動の歴史
sazimai
0
240
HackSick vol.7 LT資料【LLMアーキテクチャ入門・事前学習時の躓き所解説】 スパースなAttention・状態空間モデル
rikkabotan7
0
170
Data Visualization Tools in the Age of AI
flekschas
0
200
Model Discovery and Graph Simulation: A Lightweight Gateway to Chaos Engineering
anatolykr
0
290
Featured
See All Featured
Leveraging Curiosity to Care for An Aging Population
cassininazir
1
490
A Modern Web Designer's Workflow
chriscoyier
698
190k
Principles of Awesome APIs and How to Build Them.
keavy
128
18k
The Psychology of Web Performance [Beyond Tellerrand 2023]
tammyeverts
49
3.5k
For a Future-Friendly Web
brad_frost
183
10k
Keith and Marios Guide to Fast Websites
keithpitt
413
23k
Code Reviewing Like a Champion
maltzj
528
40k
Discover your Explorer Soul
emna__ayadi
2
1.3k
Scaling GitHub
holman
464
140k
Building a Modern Day E-commerce SEO Strategy
aleyda
45
9.2k
Automating Front-end Workflow
addyosmani
1369
210k
Have SEOs Ruined the Internet? - User Awareness of SEO in 2025
akashhashmi
0
490
Transcript
Autoencoding Variational Inference For Topic Models Akash Srivastava and Charles
Sutton ICLR2017ಡΈձ ಡΉਓ: @nzw0301
֓ཁ 1. Latent Dirichlet Allocation (LDA) ΛNeural Variational Inference (NVI)
Ͱ • Dirichlet ͷ reparameterization trick 2. ৽ϞσϧͷఏҊ 3. ѱ͍ہॴղʹϋϚΔͷΛ༧ 2
ࣄલࣝɿLDAͱVAEͷ֓ཁ 3
LDA จॻͷ֬తੜϞσϧ [Blei et al., 2003]
จॻͷτϐοΫQ [cВ ݚڀ ՝ ࣝ Պֶऀ ʜ ػցֶश ਓೳ Ϟσϧ αϯϓϧ ʜ τϐοΫͷ୯ޠ p(w|β) Ќ Ќ ػցֶश ػցֶशݚڀ ਓೳ՝ Ϟσϧ-%" Պֶֶण࢘ ίʔύε 4
VAE: Encoder • NNΛͬͨੜϞσϧ • Encoder: • σʔλ͔Β֬ͷύϥϝʔλͷม • ֬જࡏมΛੜ
• Decoder: • જࡏม͔Βσʔλੜ • Reparameterization trick • BPʹαϯϓϧΛؚΊΔ • ඪ४ਖ਼نͷαϯϓϧͱͷ ύϥϝʔλ͔ΒαϯϓϧΛߏ 5
VAE: Decoder • NNΛͬͨੜϞσϧ • Encoder: • σʔλ͔Β֬ͷύϥϝʔλͷม • ֬જࡏมΛੜ
• Decoder: • જࡏม͔Βσʔλੜ • Reparameterization trick • BPʹαϯϓϧΛؚΊΔ • ඪ४ਖ਼نͷαϯϓϧͱͷ ύϥϝʔλ͔ΒαϯϓϧΛߏ 6
VAE: Reparameterization trick • NNΛͬͨੜϞσϧ • Encoder: • σʔλ͔Β֬ͷύϥϝʔλͷม •
֬જࡏมΛੜ • Decoder: • જࡏม͔Βσʔλੜ • Reparameterization trick • BPʹαϯϓϧΛؚΊΔ • ඪ४ਖ਼نͷαϯϓϧͱͷ ύϥϝʔλ͔ΒαϯϓϧΛߏ 7
VAE: ϩεؔ 8 L (⇥) = D X d=1 (
1 2 ⇣ tr (⌃0) + µT 0 µ0 K log | ⌃0 | ⌘ + E ✏⇠N (0,1) ⇣ log p xd |f ( µ0 + ⌃ 1/2 0 ✏ ) ⌘ ) (Ⅰ) ࣄલͱͷKLμΠόʔδΣϯε (Ⅱ) ର ࣜશମ: Evidence Lower Bound (I) (Ⅱ)
ຊ 9
Reparameterization trick for Dirichlet Distribution • LDAͷθ: Dirichlet͔Βαϯϓϧ • Scale
family DistributionͰͳ͍ͨΊɼߏͰ͖ͳ͍ 10 จॻͷτϐοΫQ [cВ
Reparameterization trick for Dirichlet Distribution • LDAͷθ: Dirichlet͔Βαϯϓϧ • Scale
family DistributionͰͳ͍ͨΊɼߏͰ͖ͳ͍ • Laplace approximation • ਖ਼نͷαϯϓϧʹsoftmaxؔΛద༻ͯ͠༻ • ࣄલͷύϥϝʔλɿ µk = log( ↵k) 1 K K X i=1 log ↵i ⌃k,k = 1 ↵k (1 2 K ) + 1 K2 K X i=1 1 ↵k 11
ωοτϫʔΫͱϩεؔ 12 X encoder µ( X ) ⌃ ( X
) KL {N( z ; µ( X ) , ⌃ ( X ))||N( z ; µ1, ⌃1)} ✏ ⇠ N(✏; 0, I ) + decoder: f ( Z ) loss ( x, f ( Z )) • σ: softmaxؔ • β : DecoderͷॏΈʢunnormalizedʣ • σ(β): ୯ޠͷDiriclet͔ΒͷαϯϓϧʹରԠ L ( ⇥ ) = D X d=1 ( 1 2 ⇣ tr ( ⌃ 1 1 ⌃0) + ( µ1 µ0) T ⌃ 1 1 ( µ1 µ0) K + log |⌃1 | |⌃0 | ⌘ + E ✏⇠N (0,1) wt d log ⇣ ( µ0 + ⌃1/2 0 ✏ ) ⌘ !) θ සϕΫτϧ
prodLDA: ఏҊϞσϧ • Products of Experts • βͱθͷੵʹsoftmaxؔ 13 L
( ⇥ ) = D X d=1 ( 1 2 ⇣ tr ( ⌃ 1 1 ⌃0) + ( µ1 µ0) T ⌃ 1 1 ( µ1 µ0) K + log |⌃1 | |⌃0 | ⌘ + E ✏⇠N (0,1) wt d log ⇣ ( µ0 + ⌃1/2 0 ✏ ) ⌘ !) ( ✓)
࠷దԽͱωοτϫʔΫͷ NVIͷɿ ֶशͷॳظஈ֊Ͱlocal optimumʹߦ͖͍͢ • AdamͷύϥϝʔλΛௐ • ηͱβ1 ͷͷߴΊʹઃఆ •
Batch NormalizationͱDropoutΛ༻ 14
࣮ݧ 1. CoherenceͱPerplexity • ޙड़ 2. ֶशͱࣄલΛม͑ͨͱ͖ͷޮՌ • ߴֶ͍श &
Dirichlet͕ϕλʔ 3. ςετσʔλʹର͢Δ࠷దԽͷ༗ແ • ͠ͳ͍͍ͯ͘ 4. p(w|β)ͷϦετ • লུ 15
Coherence 16 දจ͔ΒҾ༻ • LDA VAE: ఏҊਪ๏ • prodLDA: ఏҊਪ๏+ఏҊϞσϧ
• LDA DMFVI: Online Mean-Field Variational Inference • NVDM: VAEϕʔεͷจॻϞσϦϯά දͷ: 40ճ࣮ߦͯ͠ࢉग़
Perplexity 17 දจ͔ΒҾ༻
ϨϏϡʔ: ؾʹͳͬͨͷΛ͍͔ͭ͘ Q1. NVDMͰadamͷֶशΛม͑ͨํ͕ެฏ A1. จʹө Q2. ϋΠύʔύϥϝʔλ࠷దԽ͔ͨ͠ A2. ൺֱख๏͍ͯ͠ΔɼఏҊख๏BO
Rating: 6-7-6-5 18
ͦͷଞ • ஶऀ࣮: TensorFlow • NVDMͷஶऀΒͷ৽Ϟσϧ͕ICML2017ʹ࠾ 19