Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
Open-Retrieval Conversational Question Answering
Search
Scatter Lab Inc.
July 24, 2020
Research
0
2.3k
Open-Retrieval Conversational Question Answering
Scatter Lab Inc.
July 24, 2020
Tweet
Share
More Decks by Scatter Lab Inc.
See All by Scatter Lab Inc.
zeta introduction
scatterlab
0
1.8k
SimCLR: A Simple Framework for Contrastive Learning of Visual Representations
scatterlab
0
4.2k
Adversarial Filters of Dataset Biases
scatterlab
0
2.3k
Sparse, Dense, and Attentional Representations for Text Retrieval
scatterlab
0
2.3k
Weight Poisoning Attacks on Pre-trained Models
scatterlab
0
2.2k
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
scatterlab
0
2.5k
Beyond Accuracy: Behavioral Testing of NLP Models with CheckList
scatterlab
0
2.3k
What Can Neural Networks Reason About?
scatterlab
0
2.3k
Exploring the Limits of Transfer Learning with Unified Text-to-Text Transformer
scatterlab
0
2.2k
Other Decks in Research
See All in Research
Time to Cash: The Full Stack Breakdown of Modern ATM Attacks
ratatata
0
180
競合や要望に流されない─B2B SaaSでミニマム要件を決めるリアルな取り組み / Don't be swayed by competitors or requests - A real effort to determine minimum requirements for B2B SaaS
kaminashi
0
430
J-RAGBench: 日本語RAGにおける Generator評価ベンチマークの構築
koki_itai
0
1.1k
Learning to (Learn at Test Time): RNNs with Expressive Hidden States
kurita
1
300
Attaques quantiques sur Bitcoin : comment se protéger ?
rlifchitz
0
110
Tiaccoon: Unified Access Control with Multiple Transports in Container Networks
hiroyaonoe
0
260
Mamba-in-Mamba: Centralized Mamba-Cross-Scan in Tokenized Mamba Model for Hyperspectral Image Classification
satai
3
400
データサイエンティストの業務変化
datascientistsociety
PRO
0
110
空間音響処理における物理法則に基づく機械学習
skoyamalab
0
150
Satellites Reveal Mobility: A Commuting Origin-destination Flow Generator for Global Cities
satai
3
300
[IBIS 2025] 深層基盤モデルのための強化学習驚きから理論にもとづく納得へ
akifumi_wachi
19
9.1k
長期・短期メモリを活用したエージェントの個別最適化
isidaitc
0
370
Featured
See All Featured
Automating Front-end Workflow
addyosmani
1371
200k
The Mindset for Success: Future Career Progression
greggifford
PRO
0
200
StorybookのUI Testing Handbookを読んだ
zakiyama
31
6.5k
Beyond borders and beyond the search box: How to win the global "messy middle" with AI-driven SEO
davidcarrasco
0
26
Highjacked: Video Game Concept Design
rkendrick25
PRO
0
260
From π to Pie charts
rasagy
0
97
The Limits of Empathy - UXLibs8
cassininazir
1
200
Primal Persuasion: How to Engage the Brain for Learning That Lasts
tmiket
0
190
Designing Experiences People Love
moore
143
24k
Navigating Algorithm Shifts & AI Overviews - #SMXNext
aleyda
0
1k
Code Reviewing Like a Champion
maltzj
527
40k
Color Theory Basics | Prateek | Gurzu
gurzu
0
160
Transcript
Open-Retrieval Conversational Question Answering ࢲ࢚ (ܻࢲ ࢎ౭झ, ೝಯ)
ѐਃ Open-Retrieval Conversational Question Answering
ѐਃ ѐਃ • SIGIR 20 • Chen Qu, Liu Yang,
Cen Chen, Minghui Qiu, W. Bruce Croft, Mohit Iyyer • University of Massachusetts Amherst, Ant Financial, Alibaba Group • Conversational searchਸ ਤ೧ ConvQAܳ open retrieval settingਵ۽ ഛೞח Ѫ ਃ োҳ ਃ
ѐਃ ѐਃ • Conversational search information retrieval Ҿӓੋ ݾী ೞա
• ୭Ӕ োҳٜ conversational searchܳ response rankingҗ conversational question answering۽ ೧Ѿ • ױࣽ ߸ਸ য candidate setীࢲ ҊܰѢա য passageীࢲ spanਸ ࢶఖ • ח conversational searchীࢲ retrieval ӝୡੋ ഝਸ ޖदೞח ߑध • ࠄ ֤ޙ open-retrieval conversational question answering(ORConvQA) settingਸ ઁউೞৈ ޙઁܳ ೧Ѿ
ѐਃ ѐਃ • ORConvQAী ೠ োҳܳ ਤ೧ OR-QuAC ؘఠ ࣇਸ
ٜ݅ਵݴ ORConvQAܳ ਤೠ end-to-end दझమਸ ҳ୷ೞݴ ےझನݠ ӝ߈ retriever, reranker ৬ reader ١ਸ ನೣ • OR-QuACܳ ࢚ਵ۽ ೠ ֤ޙ प learnable retriever ਃࢿਸ ૐݺ • ژೠ ݽٚ दझమ ҳࢿ ਃࣗ(retriever, reranker ৬ reader)ীࢲ history modelingਸ ࢎਊೞݶ दझమ ѱ ѐࢶ ؼ ࣻ ਸ ࠁ
Dataset Open-Retrieval Conversational Question Answering
ORConvQA? Dataset • conversational search systemsਸ ҳ୷ೞӝ ਤೠ ୶о ױ҅۽ࢲ
߸ਸ Ҋܰ ӝ ী retrieve evidenceܳ large collection۽ ࠗఠ Ѩ࢝ 1. ࠁܳ ҳೞח ചܳ ઁҕ(information seeker৬ information provider)৬ ೞח QuAC dataset 2. QuAC ޙਸ context-independentೞѱ द ࢿೠ CANARD dataset 3. Wikipedia passage
Dataset
CANARD? Dataset • QuAC dialogsח self-containedೞ ঋח ড חؘ ח
ࠛ৮ೠ ୡӝ ޙਵ۽ ੋ೧ ߊࢤ • ܳ ٜয seekerীѱ a Chinese polymathic scientistੋ Zhang Hengী ೧ ߓۄҊ ೮חؘ ޙ "җҗ ӝࣿҗ যڃ ҙ ҅о णפө?” • ۞ೠ ࠛౠೞҊ ݽഐೠ ୡӝ ޙ ചܳ ೧ࢳೞӝ য۵ѱ ೞӝ ٸޙী ҕѐ Ѩ࢝ ജ҃ীࢲ ޙઁܳ ঠӝ • CANARD ؘఠ ࣁীࢲ ઁҕೞח context-independent rewritesਵ۽ ೞৈ ޙઁܳ ೧Ѿ, Ӓۢ "Zhang Heng җ ӝ ࣿҗ যڃ ҙ҅о णפө?"۽ ޙ
CANARD? Dataset • ߣ૩ ޙী ೧ࢲ݅ Үܳ ࣻ೯ೞݶ ച
ղীࢲ history dependenciesਸ Ӓ۽ ਬೞݶࢲ ചо self-contained • QuAC test set ҕѐغয ঋӝ ٸޙী QuAC dev setਸ ਊೞৈ CANARD test setਸ ݅ٞ • ژೠ QuAC train set 10%ܳ dev۽ ഝਊ. • CANARDী হח QuAC ޙ ತӝ೮ਵݴ ܳ ਊೠ ࢤ ؘఠ ੋ OR-QuAC ؘఠ ా҅ח җ э.
Model Open-Retrieval Conversational Question Answering
ݽ؛ Retriever, Reranker, Reader۽ ա Model
ݽ؛ Retriever, Reranker, Reader۽ ա Model
Passage Retriever Dataset • Passage Encoder • Question Encoder •
Retrieval Score
Retrieval score ӝળਵ۽ ࢚ਤ top-Kѐ ޙࢲܳ rerank৬ reader۽ ׳ Model
ݽ؛ Retriever, Reranker, Reader۽ ա Model
Reranker& Reader Encoding Dataset • Input • Contextualized Representations •
sequence representation
Reranker& Reader Dataset • Sequence Representation • Reranker (W_rr is
vector) • Reader (span prediction)
Training Open-Retrieval Conversational Question Answering
Retriever pretraining Training • retrieval scores for the batch •
to maximize the probability of the gold passage for each question • Pretraining loss Pretraning റী passage encoderח offlineਵ۽ ك. Faissܳ ࢎਊ೧ࢲ Ѿҗܳ ࡳই১.
Concurrent Learning Training • Retriever loss • Reranker loss •
Reader loss
Inference Training • Retrieval Ѿҗ Top-K ޙࢲܳ ݽف ੋಌ۠झ ೞৈ
п ޙࢲ߹ spanਸ ஏ • Retriever loss + Reranker loss + Reader lossо ઁੌ ޙࢲ spanਸ ୭ઙ ਵ۽ ஏ
RESULTS Open-Retrieval Conversational Question Answering
Competing Method RESULTS • DrQA : TF-IDF + RNN based
reader • BERTserini : BM25 + BERT reader • ORConvQA without history : our method + window size 0 • ORConvQA : our method • Evaluation Metric : word level F1, human equivalence score (HEQ), Mean Reciprocal Rank(MRR), Recall
DrQA < BERTserini < Ours w/o hist < Ours RESULTS
Ablation study RESULTS
History windows size ઑ RESULTS
хࢎפ✌ ୶о ޙ ژח ҾӘೠ ݶ ઁٚ ইې োۅ۽
োۅ ࣁਃ! ࢲ࢚ (ܻࢲ ࢎ౭झ, ೝಯ)
[email protected]
Linked in. @pingpong