Contextual embedding models encode whole documents before chunk pooling
AIRetrieval systems that split long documents into chunks lose surrounding context. Contextual embedding models address this by encoding the entire document once and pooling chunk vectors afterward. They are usually trained using one gold chunk per query.