Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
文献紹介: Attention is not Explanation
Search
Yumeto Inaoka
March 19, 2019
Research
530
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
文献紹介: Attention is not Explanation
Yumeto Inaoka
March 19, 2019
More Decks by Yumeto Inaoka
See All by Yumeto Inaoka
文献紹介: Quantity doesn’t buy quality syntax with neural language models
yumeto
1
220
文献紹介: Open Domain Web Keyphrase Extraction Beyond Language Modeling
yumeto
0
290
文献紹介: Self-Supervised_Neural_Machine_Translation
yumeto
0
200
文献紹介: Comparing and Developing Tools to Measure the Readability of Domain-Specific Texts
yumeto
0
210
文献紹介: PAWS: Paraphrase Adversaries from Word Scrambling
yumeto
0
220
文献紹介: Beyond BLEU: Training Neural Machine Translation with Semantic Similarity
yumeto
0
330
文献紹介: EditNTS: An Neural Programmer-Interpreter Model for Sentence Simplification through Explicit Editing
yumeto
0
440
文献紹介: Decomposable Neural Paraphrase Generation
yumeto
0
260
文献紹介: Analyzing the Limitations of Cross-lingual Word Embedding Mappings
yumeto
0
300
Other Decks in Research
See All in Research
敵対生成プロンプト同時探索による内省型プロンプト最適化
kinoue_smarthr
0
400
クラウド・AI 時代の研究開発 DX / R&D Digital Transformation
hariby
0
130
RS-Agent: Automating Remote Sensing Tasks through Intelligent Agent
satai
3
530
ハードウェア研究で国際トップ会議を目指す!IROS 2027での論文採択を目指して
ayatokanada
6
3.5k
Cross-Media Human-Information Interaction
signer
PRO
0
220
進学校の生徒にはア行の苗字が多いのか
ozekinote
0
560
最先端NLP勉強会2026 論文紹介:Reasoning with Sampling: Your Base Model is Smarter Than You Think (ICLR 2026 paper)
kogoro
4
580
重要だけど測れていないもの:高齢者ケアの見えない課題
theoriatec2024
0
510
Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance
satai
3
130
20260624 NLP colloquium: 単一のhubテキストがCLIPを壊す:hubnessによる埋め込みの脆弱性特定
de9uch1
2
230
最先端NLP 2026 論文紹介: Wait, Wait, Wait... Why Do Reasoning Models Loop? / SNLP Paper Review: Wait, Wait, Wait... Why Do Reasoning Models Loop?
tkng
0
210
2026年版中小企業白書・小規模企業白書の概要
ozekinote
0
210
Featured
See All Featured
Designing for Timeless Needs
cassininazir
1
470
Stop Working from a Prison Cell
hatefulcrawdad
274
21k
How GitHub (no longer) Works
holman
316
150k
Testing 201, or: Great Expectations
jmmastey
46
8.3k
The Cost Of JavaScript in 2023
addyosmani
55
10k
Claude Code どこまでも/ Claude Code Everywhere
nwiizo
67
58k
The Limits of Empathy - UXLibs8
cassininazir
1
640
KATA
mclloyd
PRO
35
15k
Digital Projects Gone Horribly Wrong (And the UX Pros Who Still Save the Day) - Dean Schuster
uxyall
1
2.7k
New Earth Scene 8
popppiees
3
2.5k
[Rails World 2023 - Day 1 Closing Keynote] - The Magic of Rails
eileencodes
38
3k
Applied NLP in the Age of Generative AI
inesmontani
PRO
4
2.4k
Transcript
Attention is not Explanation 文献紹介 2019/03/19 長岡技術科学大学 自然言語処理研究室 稲岡 夢人
Literature 2 Title Attention is not Explanation Author Sarthak Jain,
Byron C. Wallace Conference NAACL-HLT 2019 Paper https://arxiv.org/abs/1902.10186
Abstract Attentionは入力単位で重みを示す その分布が重要性を示すものとして扱われることがある → 重みと出力の間の関係は明らかにされていない 標準的なAttentionは意味のある説明を提供しておらず、 そのように扱われるべきでないことを示す
3
調査方法 1. 重みが重要度と相関しているか 2. 事実と反する重み設定が予測を変化させるか 4
Tasks Binary Text Classification Question Answering (QA)
Natural Language Inference (NLI) 5
Datasets 6
Results 7
Definitions 出力結果の比較に使用する距離 Attentionの比較に使用する距離 8
Results Attentionの変化を大きくしても結果が変化しない → 出力への影響が小さい 9
Results DiabetesはPositiveのクラスにおいては影響が大きい → 高精度で糖尿病を示すトークンが存在するため 10
Adversarial Attention 出力を大きく変化させるようにAttentionを変化させる Attentionが少し変化しただけで出力が大きく変化するか ← Attentionの挙動を確認 12
Results 少しのAttentionの変化で出力が大きく変化している 13
Conclusions 重要度とAttentionの重みは相関が弱い 事実に反する重みは必ずしも出力を変化させない Seq2seqについては今後の課題とする 14