Skip to content
Read the original: Perplexity· Published 25/100AI score25/100

Contextual embedding models encode whole documents before chunk pooling

Original titleRetrieval systems often split long documents into chunks, but this strips away the surrounding context.

AISummary

Retrieval systems that split long documents into chunks lose surrounding context. Contextual embedding models address this by encoding the entire document once and pooling chunk vectors afterward. They are usually trained using one gold chunk per query.

Read the original x.com

Source: Perplexity · x.comPublished · added here