Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
文献紹介: Bag-of-Words as Target for Neural Machine...
Search
Yumeto Inaoka
January 22, 2019
Research
220
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
文献紹介: Bag-of-Words as Target for Neural Machine Translation
2019/1/22の文献紹介で発表
Yumeto Inaoka
January 22, 2019
More Decks by Yumeto Inaoka
See All by Yumeto Inaoka
文献紹介: Quantity doesn’t buy quality syntax with neural language models
yumeto
1
220
文献紹介: Open Domain Web Keyphrase Extraction Beyond Language Modeling
yumeto
0
290
文献紹介: Self-Supervised_Neural_Machine_Translation
yumeto
0
200
文献紹介: Comparing and Developing Tools to Measure the Readability of Domain-Specific Texts
yumeto
0
210
文献紹介: PAWS: Paraphrase Adversaries from Word Scrambling
yumeto
0
220
文献紹介: Beyond BLEU: Training Neural Machine Translation with Semantic Similarity
yumeto
0
330
文献紹介: EditNTS: An Neural Programmer-Interpreter Model for Sentence Simplification through Explicit Editing
yumeto
0
440
文献紹介: Decomposable Neural Paraphrase Generation
yumeto
0
260
文献紹介: Analyzing the Limitations of Cross-lingual Word Embedding Mappings
yumeto
0
300
Other Decks in Research
See All in Research
高性能計算機クラスタを用いた大規模点群処理による森林の単木抽出と構造解析
kentaitakura
1
120
研究室単位での自律的 IPv6接続性確立に向けたAS共同運用モデルの提案と実証
reokashiwa
PRO
0
210
Using our influence and power for patient safety
helenbevan
0
410
HackSick vol.7 LT資料【LLMアーキテクチャ入門・事前学習時の躓き所解説】 スパースなAttention・状態空間モデル
rikkabotan7
0
170
EIRによる不正端末のブロッキング 5G時代におけるデバイス識別と不正対策の進化
stellarcraft
0
130
セマンティック通信勉強会 6Gに向けたデバイス間効率的な通信の技術紹介・課題・今後展望
satai
3
320
AIエージェント時代のLLM-jpモデルのあるべき姿
k141303
0
620
東京大学工学部計数工学科、計数工学特別講義の説明資料
kikuzo
0
650
ハードウェア研究で国際トップ会議を目指す!IROS 2027での論文採択を目指して
ayatokanada
6
3.5k
Ghost in the 7‑Zip: The Shadow of Residential Proxies Creeping into Your Life
nttcom
0
2.1k
MM-OVSeg: Multimodal Optical–SAR Fusion for Open-Vocabulary Segmentation in Remote Sensing
satai
3
120
量子サマースクール2026「量子計算機アーキテクチャ分野の概観」
youten622
1
860
Featured
See All Featured
Amusing Abliteration
ianozsvald
1
290
How To Speak Unicorn (iThemes Webinar)
marktimemedia
1
570
Lessons Learnt from Crawling 1000+ Websites
charlesmeaden
PRO
1
1.6k
Avoiding the “Bad Training, Faster” Trap in the Age of AI
tmiket
0
230
Ecommerce SEO: The Keys for Success Now & Beyond - #SERPConf2024
aleyda
1
2.1k
Leveraging Curiosity to Care for An Aging Population
cassininazir
1
490
The Anti-SEO Checklist Checklist. Pubcon Cyber Week
ryanjones
0
230
The Illustrated Guide to Node.js - THAT Conference 2024
reverentgeek
1
470
Claude Code どこまでも/ Claude Code Everywhere
nwiizo
67
58k
What does AI have to do with Human Rights?
axbom
PRO
1
2.4k
The Curse of the Amulet
leimatthew05
2
14k
Kristin Tynski - Automating Marketing Tasks With AI
techseoconnect
PRO
0
510
Transcript
1 Bag-of-Words as Target for Neural Machine Translation 文献紹介 2019/1/22
長岡技術科学大学 自然言語処理研究室 稲岡 夢人
Literature • Bag-of-Words as Target for Neural Machine Translation •
Shuming Ma, Xu SUN, Yizhong Wang, Junyang Lin • Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), pages 332-338, 2018. 2
Abstract 翻訳において正解はひとつじゃない 既存のNMTではひとつのみを正解として使用 → 他の正解は誤りとして学習される 正解同士は似たBag-of-Words (BoW)
を共有する → BoWによって正解とそれ以外を区別できる 学習セットにない正解を考慮するためにBoWを利用 → 中国語-英語の翻訳において優位性を確認 3
Introduction NMTは首尾一貫の妥当な翻訳の生成ができる 現在のNeural Machine Translation (NMT)の 多くはSequence-to-Sequence モデル(Seq2Seq)に
基づいている 4
Seq2Seq (Overview) 5 私 は 元気だ <BOS> I am fine
<EOS> 入力文 出力文 Encoder Decoder
Seq2Seq (Encoder) 6 私 は 元気だ One-hot vector Embedding layer
Recurrent layer 入力文
Seq2Seq (Decoder) 7 I <BOS> I am fine am fine
<EOS> One-hot vector Embedding layer Recurrent layer One-hot vector Output layer 出力文
Introduction NMTではひとつの正解のみを 学習に用いる 他の正解は誤った翻訳と学習 → 悪影響を与える可能性 8
Introduction 正しい翻訳は似たBoWを共有 → 正しい翻訳と誤った翻訳は BoWで区別できる 文とBoWの両方を対象とする 手法を提案 →
T.2よりT.1を優遇 9
Bag-of-Words Generation マルチラベル分類問題のようにBoWを生成 Decoderの出力である単語レベルのスコアベクトル を 合計して、文レベルのスコアベクトルを得る 文レベルのスコアベクトルは、文中の任意の位置に
対応する単語が出現する確率を表す 10
Notation データセットに含まれるサンプル数:N i番目のサンプル:(, ) (x: source, y: target)
= 1 , 2 , … , = 1 , 2 , … , = 1 , 2 , … , はのBoWを表す 11
Bag-of-Words Generation 12 = softmax = �
Targets and Loss Function 文の翻訳とBoWの生成でそれぞれ損失関数(1 , 2 )を定義
重み で2つの損失を足し合わせる() (𝑖𝑖 : epoch , k, : fixed-value) 1 = − � =1 log l2 = − � =1 log = 1 + 2 = min(, + 𝛼𝛼) 13 𝑖𝑖
Experiments LDCコーパス(1.25M)で学習、NIST翻訳タスクで評価 語彙サイズを英中それぞれ5万語に設定 BLEUで評価 14
Results 15 4.55 BLEU points↑
Results 16 4.55 BLEU points↑
Results 17
Conclusions 正解訳とBoWの両方を考慮する手法を提案 提案手法が強力なベースラインに対して優位である結果 Morphologically-rich language*や低資源言語において どのように適用するかについて今後の課題とする *
文法的関係が相対位置や助詞ではなく単語の変化で 決まるような言語 18