Perplexity's 0.6B and 9B embedding models share one embedding space
Original titleBoth sizes are distilled token by token from one 18B teacher, so they share one embedding space.
AISummary
Perplexity's 0.6B and 9B models are both distilled token by token from one 18B teacher, so they share a single embedding space. A corpus indexed with the 9B model can be searched using 0.6B queries, raising ViDoRe v3 from 62.3% to 63.5% with no added query cost.
Source: Perplexity · x.comPublished · added here