A White Box Analysis of ColBERT - Sorbonne Université Access content directly
Preprints, Working Papers, ... Year : 2020

A White Box Analysis of ColBERT

Abstract

Transformer-based models are nowadays state-of-the-art in ad-hoc Information Retrieval, but their behavior is far from being understood. Recent work has claimed that BERT does not satisfy the classical IR axioms. However, we propose to dissect the matching process of ColBERT, through the analysis of term importance and exact/soft matching patterns. Even if the traditional axioms are not formally verified, our analysis reveals that ColBERT: (i) is able to capture a notion of term importance; (ii) relies on exact matches for important terms.

Dates and versions

hal-03084279 , version 1 (21-12-2020)

Identifiers

Cite

Thibault Formal, Benjamin Piwowarski, Stéphane Clinchant. A White Box Analysis of ColBERT. 2020. ⟨hal-03084279⟩
21 View
0 Download

Altmetric

Share

Gmail Mastodon Facebook X LinkedIn More