Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
Dialogue Natural Language Inference
Search
Sponsored
·
Your Podcast. Everywhere. Effortlessly.
Share. Educate. Inspire. Entertain. You do you. We'll handle the rest.
→
Scatter Lab Inc.
April 03, 2020
Research
2.4k
0
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
Dialogue Natural Language Inference
Scatter Lab Inc.
April 03, 2020
More Decks by Scatter Lab Inc.
See All by Scatter Lab Inc.
zeta introduction
scatterlab
0
2k
SimCLR: A Simple Framework for Contrastive Learning of Visual Representations
scatterlab
0
4.5k
Adversarial Filters of Dataset Biases
scatterlab
0
2.3k
Sparse, Dense, and Attentional Representations for Text Retrieval
scatterlab
0
2.4k
Weight Poisoning Attacks on Pre-trained Models
scatterlab
0
2.2k
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
scatterlab
0
2.6k
Beyond Accuracy: Behavioral Testing of NLP Models with CheckList
scatterlab
0
2.4k
Open-Retrieval Conversational Question Answering
scatterlab
0
2.4k
What Can Neural Networks Reason About?
scatterlab
0
2.3k
Other Decks in Research
See All in Research
Visual SLAM未来予測 / Future Prediction in Visual SLAM
koide3
1
1k
SoftMatcha 2: 1兆語規模コーパスの超高速かつ柔らかい検索
e869120_sub
7
3.8k
長時間動画QAにおけるマルチエージェント推論 ・SVAgent: Storyline-Guided Long Video Understanding via Cross-Modal Multi-Agent Collaboration
murakawatakuya
1
200
SAM3を用いたコマ・吹き出しの 領域検出と分割構造からの読み順推定
kzmssk
0
120
RS-Agent: Automating Remote Sensing Tasks through Intelligent Agent
satai
3
540
横浜市長(山中氏)の言動にかかる第三者による調査報告書
y150saya
0
200
Vector Map as Language: Toward Unified Remote Sensing Vector Mapping
satai
3
270
OWASP AISVS - C7
shiell
2
700
Anthropic が提案する LLM の内部状態を自然言語で説明可能にした Natural Language Autoencoders / Natural Language Autoencoders Produce Unsupervised Explanations of LLM Activations
shunk031
0
210
コーディングエージェントとABNを再考
hf149
2
860
EIRによる不正端末のブロッキング 5G時代におけるデバイス識別と不正対策の進化
stellarcraft
0
130
Claude Code × autoresearch 実践
mathbullet
0
270
Featured
See All Featured
The Invisible Side of Design
smashingmag
301
52k
Practical Orchestrator
shlominoach
191
12k
Groundhog Day: Seeking Process in Gaming for Health
codingconduct
0
340
A Guide to Academic Writing Using Generative AI - A Workshop
ks91
PRO
1
440
B2B Lead Gen: Tactics, Traps & Triumph
marketingsoph
0
230
Let's Do A Bunch of Simple Stuff to Make Websites Faster
chriscoyier
508
140k
The browser strikes back
jonoalderson
0
1.7k
SEO Brein meetup: CTRL+C is not how to scale international SEO
lindahogenes
1
2.9k
Facilitating Awesome Meetings
lara
57
7.1k
The Success of Rails: Ensuring Growth for the Next 100 Years
eileencodes
47
8.3k
Color Theory Basics | Prateek | Gurzu
gurzu
0
460
Information Architects: The Missing Link in Design Systems
soysaucechin
1
1.1k
Transcript
Dialogue Natural Language Inference Sean Welleck et al., ACL’19 ࢿࠁ
(ML Research Scientist, Pingpong)
ݾର ݾର 1. Dialogue Consistency and NLI 2. Dialogue NLI
Dataset 1. Triple Generation 2. Triple Annotation 3. Re-ranking with NLI 4. Evaluation 1. On Dialogue NLI 2. On Consistency in Dialogue
Dialogue Consistency and NLI Dialogue Consistency and NLI
Dialogue Consistency and NLI Dialogue Consistency and NLI • ചীࢲ
࠺ੌҙࢿ • ࢚ਵ۽ ൞ӈೞա ೠߣ ߊࢤೞݶ ఋѺ ఀ • Semanticೠ ޙਸ ݅٘ח ֢۱݅ਵ۽ח ೧Ѿ ࠛо • Natural Language Inference (NLI) • NLU, sentence representation ١ NLP ߈ਸ ੜೞӝ ਤೠ ࣻױਵ۽ॄ જ • NLI ݽ؛ downstream task ࢿמ ೱ࢚ী ӝৈ
Dialogue Consistency and NLI Dialogue Consistency and NLI • ಕܰࣗա:
ޙ ഋక۽ അ. • ചীࢲ ੌҙࢿ • ӝࠄਵ۽ Persona consistency • ֤ܻਵ۽ ߓغח ݈ ইפۄب э ࢎۈ ݈ೡ Ѫ э ঋ ޙ • : ച ղীࢲ ೠ ࢎۈ ೠ ف ݈ ߓغח • : Ӓ ࢎۈ ಕܰࣗա৬ ߓغח P = {p1 , …, pm } (uA i , uA j ) (uA i , pA k )
Dialogue Consistency and NLI
Dialogue NLI Dataset Dialogue NLI Dataset
Dialogue NLI Dataset Dialogue NLI Dataset • ߊച-ಕܰࣗա , ಕܰࣗա-ಕܰࣗա
हਵ۽ ܖয • ߊച-ߊച हب ನೣغয ਵա प ೞ ঋ (ui , pj ) (pi , pj ) (ui , uj )
Triple Generation Dialogue NLI Dataset • Triple • PersonaChatীࢲ ಕܰࣗա
ޙҗ ߊച ੌࠗ۽ Triple ۨ࠶ • ՙܻ Entailment, Neutral, Contradiction కӦ • Tripleਸ ӝળਵ۽ E, N, Cܳ ݅ٚ! • Entailment: э Tripleী ࣘೞח ف ޙՙܻ • Neutral, Contradiction: 3о ߑߨ (e1 , r, e2 ) (u, p), (p, p)
Neutral Pairs Dialogue NLI Dataset • Miscellaneous utterance যו Tripleীب
ࣘೞ ঋח ߊച ৬ ಕܰࣗա ޙ ҙ҅ח Neutral • Persona pairing Ground truth ಕܰࣗաՙܻח ࠂغѢա ݽࣽغ ঋחח ઁ ೞী э Tripleਸ ҕਬೞ ঋ ח ಕܰࣗաՙܻ Ҋ, ೞਤ ޙٜՙܻب • Relation swap ࢲ۽ ة݀ੋ ࢎपਸ աఋղח ҙ҅ ী ࣘೞח ޙٜՙܻ u p (r, r′ )
Contradiction Pairs Dialogue NLI Dataset • Relation swap ࢲ۽ ݽࣽغח
ҙ҅ ী ࣘೞח ޙٜՙܻ • Entity swap Triple ীࢲ ೧ࢲ о عਸ ٸ ݽࣽغח ҃ ف Tripleী ࣘೞח ޙٜ ՙܻ • Numeric Tripleী ನೣػ ंܳ ܲ ं۽ ߄Լࢲ ٜ݅য ޙҗ ਗې Tripleী ؍ ޙٜਸ (r, r′ ) (e1 , r, e2 ) e2 → e′ 2 (e1 , r, e′ 2 )
Triple Annotation Dialogue NLI Dataset • ಕܰࣗա ޙ →
<category> <relation> <category> ex) <person> have_pet <animal> relation , entity ա, entityח schemaী হਵݶ ੑ۱ • ٜ݅য Triple۽ ࠙ܨ ف ઑѤ ೞաܳ ݅ೞݶ 1. о sub-string 2. (e1 , r, e2 ) ∈ ℛ ∈ ℰ u ∈ U u ∈ (e1 , r, e2 ) e2 u sim(u, p) ≥ τ
Statistics Dialogue NLI Dataset • Gold-standard test set: test set
ۨ࠶ ݏҊ ೠ ࢎۈ 3ݺ 2ݺ ࢚ੋ ࢠ݅ ݽ Ѫ
Dialogue NLI Dataset
Re-ranking with NLI Re-ranking with NLI
Consistent Dialogue Agents via NLI Re-ranking with NLI •
ߊച ஏী NLI ݽ؛ ஏ Ѿҗ ഝਊ NLI ݽ؛ Contradictionۄ ౸ױೠ റࠁח confidence݅ఀ ಕօ౭ܳ ષ ࢜۽ ࣻ۽ Re-ranking
Evaluation Evaluation
On Dialogue NLI Evaluation • InferSent, ESIM ف ݽ؛ ࢎਊ
On Consistency in Dialogue Evaluation • ݽ؛ • ച ݽ؛:
Key-value memory networkܳ PersonaChatਵ۽ ण • NLI ݽ؛: ESIMਸ Dialogue NLI۽ ण • ಣо ࣇ • PersonaChatীࢲ Triple ী ೧ೞח ߊച ܳ Ҋ agent ಕܰࣗաী ী ࣘೞח ޙ ਵݶ ܳ ਵ۽ р • Entailment ޙ 10ѐ, Contradiction ޙ 10ѐ, ޙ 10ѐܳ റࠁ۽ م • ݫܼ • Hits@k, Entail@k, Contradict@k (e1 , r, e2 ) u (e1 , r, e2 ) u
Evaluation
Result Evaluation
Human Evaluation Evaluation • ParlAIܳ ా೧ w/o re-rankingҗ w/ re-rankingਸ
࠺Ү • ಣо ୋب • ݽ؛ ݃ա ಕܰࣗաܳ ੜ ߸೮חо? (1~5) • ݽ؛ п ߊചо ಕܰࣗա৬ ੌҙغחо? (0, 1) • ݽ؛ п ߊചо ݽ؛ ߊച, ݽ؛ ಕܰࣗա৬ ݽࣽغחо? (0, 1)
хࢎפ✌ ୶о ޙ ژח ҾӘೠ ݶ ઁٚ ইې োۅ۽
োۅ ࣁਃ! ࢿࠁ (ML Research Scientist, Pingpong)
[email protected]