Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
The Natural Language Decathlon: Multitask Learn...
Search
Sponsored
·
Your Podcast. Everywhere. Effortlessly.
Share. Educate. Inspire. Entertain. You do you. We'll handle the rest.
→
Scatter Lab Inc.
July 10, 2019
Research
930
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
The Natural Language Decathlon: Multitask Learning as Question Answering
Scatter Lab Inc.
July 10, 2019
More Decks by Scatter Lab Inc.
See All by Scatter Lab Inc.
zeta introduction
scatterlab
0
1.9k
SimCLR: A Simple Framework for Contrastive Learning of Visual Representations
scatterlab
0
4.4k
Adversarial Filters of Dataset Biases
scatterlab
0
2.3k
Sparse, Dense, and Attentional Representations for Text Retrieval
scatterlab
0
2.3k
Weight Poisoning Attacks on Pre-trained Models
scatterlab
0
2.2k
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
scatterlab
0
2.6k
Beyond Accuracy: Behavioral Testing of NLP Models with CheckList
scatterlab
0
2.3k
Open-Retrieval Conversational Question Answering
scatterlab
0
2.3k
What Can Neural Networks Reason About?
scatterlab
0
2.3k
Other Decks in Research
See All in Research
LLM の Attention 機構まとめ — 数式・計算量・メモリ
puwaer
8
2.3k
論文紹介:HalluCitation Matters
wasyro
0
130
Language and AI
ayaniwa
0
170
データセンター事業者を取り巻く近年の状況とその中での研究開発動向、テストベッドへの貢献の可能性
kikuzo
1
260
ScoreMatchingRiesz for Automatic Debiased Machine Learning and Policy Path Estimation with an Application to Japanese Monetary Policy Evaluation
masakat0
0
300
2026年度 生成AI を活用した論文執筆ガイド/ワークショップ / 2026 Academic Year Guide to Writing Papers Using Generative AI - Workshop
ks91
PRO
0
190
Cross-Media Human-Information Interaction
signer
PRO
0
120
YOLO26_ Key Architectural Enhancements and Performance Benchmarking for Real-Time Object Detection
satai
3
880
適応的スパムフィルタのための軽量な類似メッセージカウンタ / jsai2026-adaptive-spam-filter
monochromegane
0
4.3k
Apache Gravitinoで実現する Icebergカタログ統合とアクセスの一元化
matsumooon
0
350
AIを叩き台として、 「検証」から「共創」へと進化するリサーチ
mela_dayo
0
310
第64回CV・PRML勉強会 論文紹介:Linguistic Priors for Visual Decoupling: Towards Symmetric Vision-Brain Alignment
sokikatayama
0
140
Featured
See All Featured
Self-Hosted WebAssembly Runtime for Runtime-Neutral Checkpoint/Restore in Edge–Cloud Continuum
chikuwait
0
650
Fireside Chat
paigeccino
42
4k
Testing 201, or: Great Expectations
jmmastey
46
8.2k
The Invisible Side of Design
smashingmag
301
52k
How to optimise 3,500 product descriptions for ecommerce in one day using ChatGPT
katarinadahlin
PRO
1
3.7k
More Than Pixels: Becoming A User Experience Designer
marktimemedia
3
460
A Soul's Torment
seathinner
6
3.1k
JAMstack: Web Apps at Ludicrous Speed - All Things Open 2022
reverentgeek
1
490
DBのスキルで生き残る技術 - AI時代におけるテーブル設計の勘所
soudai
PRO
67
56k
BBQ
matthewcrist
89
10k
Designing Dashboards & Data Visualisations in Web Apps
destraynor
231
55k
Digital Ethics as a Driver of Design Innovation
axbom
PRO
1
340
Transcript
스캐터랩(ScatterLab) ੌ࢚ച ੋҕמ Scatterlab ML Technical Seminar Session 2 (QA):
백영민 The Natural Language Decathlon: Multitask Learning as Question Answering McCann et al. Salesforce Research Machine Learning Engineer
#1. Concept
!3 Multitask Learning ৈ۞ о taskܳ э ण೧ࠁ!
• Method: ౠ objectiveܳ оҊ णػ model/representationਸ ܲ downstream taskী
ਊೞח Ѫ • ੌ߈ਵ۽ Language Modeling١ Natural Language ߈ੋ ౠࢿਸ णೡ ࣻ ח objective ਊ • : • Random initializeীࢲ दೞח Ѫ ࠁ જ Ѿҗܳ ࠁҊ, ࡅܲ ࣻ۴ਸ оמೞѱ ೧ષ • ߑध • Representation: Word2Vec, Glove ١ fixed representation, ULMFit, ELMO ١ intermediate representation(context aware)ਸ downstream taskীࢲ ਊೞח Ѫ(߹ب model ઓ) • Model: BERT, GPT١ pre-trainingী ਊ೮؍ modelਸ downstream taskীࢲ “fine-tuning” !4 #1 Concept Transfer Learning
• Method: ৈ۞ taskٜਸ ೞա ݽ؛۽ زदী णदఃח Ѫ •
Chunking, POS tagging, NER, SRL, dependency parsing, NLI ١ NLP taskٜਸ زੌೠ ݽ؛۽ زदী ण • : • ৈ۞ taskܳ زदী modelingೡ ࣻ • ੜ णغݶ ౠ taskী ೠػ Ѫ ইצ “General Representation”ਸ ਸ ࣻ • Zero-shot Learning, Meta-Learning ١ ਊ оמࢿ • ୭Ӕ ഝߊ োҳо ܖযҊ ח ࠙ঠ • ই singletask learningী ࠺Үؼ݅ೠ ࢿמਸ ࠁৈҊ ঋ݅, challengingೠ োҳٜ ݆ ܖযҊ . • Image classification + NLP • ࣗѐೡ ֤ޙ “MQAN” singletask learningী Ӓա݃ ࠺Үؼ݅ೠ ࢿמਸ ࠁ(BERT ֤ޙ ੑפ) !5 #1 Concept Multitask Learning
#2. Method
!7 Approach ݽٚ taskܳ QAഋधী ݏࠁ!
• য questionী ೠ ਸ “contextղীࢲ” ח ޙઁܳ ಿ (
Context ղࠗী Ҋ о) • Ex) SQuAD, RACE… • ࠁా द index৬ indexܳ ח ߑध !8 #2 Approach Question Answering
!9 #2 Approach Question Answering Idea: ݽٚ taskٜਸ “QAഋध”ਵ۽ ٜ݅যࢲ
“Multitask Learning”ਸ ೧ࠁ! QA Translation Summary NLI Sentiment Analysis
• ୨ 10ѐ taskܳ Multitask Learning !10 #2 Approach Decathlon
!11 Model Architecture Multitask Question Answering Network(MQAN)
!12 #2 Architecture Overview
!13 #2 Architecture I/O • Input: • Q: Question Sentences
• C: Context Sentences • A: Answer Sentences(for generation - autoregressive) • Output: • General QA: Contextীࢲ द, indexܳ • Q(Question Sentence) + C(Context Sentences) + Outer Vocabulary(Generation) ী ࢶఖ
!14 #2 Architecture Feature - Input Representation
!15 #2 Architecture Feature - Alignment <dummy dataܳ ֍ח ਬ>
!16 #2 Architecture Feature - Dual Coattention
!17 #2 Architecture Feature - Compression & Self-Attention
!18 #2 Architecture Feature - Answer Representation
!19 #2 Architecture Feature - Answer Representation
!20 #2 Architecture Feature - Answer
!21 Training Strategy Curriculum learning
!22 #3 Traning Strategy Multitask Learning Strategy • Round-robin Algorithm
• CPU scheduling ߑߨ ೞա۽ ஹೊఠ ਗਸ ࢎਊೡ ࣻ ח ӝഥܳ “۽ࣁझٜীѱ ҕ”ೞѱ ࠗৈ • п ۽ࣁझী ੌदрਸ ೡ, ೡػ दр աݶ Ӓ ۽ࣁझ ਫ਼द ࠁܨ, ܲ ۽ࣁझীѱ ӝഥܳ ષ • п Taskী ੌߓܳ ೡ, ೡػ ߓо աݶ Ӓ Taskਸ ਫ਼द ࠁܨ, ܲ Taskীѱ ӝഥܳ ષ • Fully Joint • п Task ࣽࢲܳ ҊೞҊ, round-robin algorithmਸ ా೧ batchܳ sampling • Single-task trainingীࢲ iterationਵ۽ب ࣻ۴೮؍ taskٜ ੜ غ݅ աݠח single-task݅ ण೮ ਸ ٸ݅ఀ ࢿמਸ ࠁৈ ޅೣ
!23 #3 Traning Strategy Multitask Learning Strategy • Curriculum Strategy
• Curriculumਸ ٜ݅যࢲ learningदெࠁ! • First Phase:࠺Ү ए taskٜਸ ݢ ण -> Second Phase:য۰ taskٜਸ ण • First Phase(SST, QA-SRL, QA-ZRE, WOZ, WikiSQL, MWSC) -> Second Phase(Others)
!24 #3 Traning Strategy Multitask Learning Strategy • Curriculum Strategy
• Curriculumਸ ٜ݅যࢲ learningदெࠁ! • First Phase:࠺Ү ए taskٜਸ ݢ ण -> Second Phase:য۰ taskٜਸ ण • First Phase(SST, QA-SRL, QA-ZRE, WOZ, WikiSQL, MWSC) -> Second Phase(Others) • Anti-Curriculum Strategy • Curriculumী “߈(Anti)ೞח” ۚ - Curriculum Learning Bengio et al. [2009] • ए taskח ܲ taskٜী بਸ ࣻ ח ਬਊೠ representationਸ णೡ ࣻ হ! • First Phase:য۰ taskٜਸ ݢ ण -> Second Phase:ए taskٜਸ ण • First Phase(SQuAD, IWSLT, CNN/DM, MNLI) -> Second Phase(Others)
#3. Result
!26 Result ־о־о ੜ೮ա?
!27 #1single vs multitask Single vs Multitask Training
!28 #2 curriculum Training Strategy
!29 #2 curriculum Pointer weight distribution
!30 Q & A хࢎפ