OZZZER · AI NEWS1 of 3 free stories opened
← Back to AI News

Models & tools · 1 Oct 2026 · 05:23 CEST

Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence

MarkTechPost · 1 Oct 2026 · 05:23 CESTRead original at MarkTechPost ↗
Share
LinkedInXFacebookWhatsApp

Publisher preview · OZZZER analysis pending editorial review.

PUBLISHER ARTICLE PREVIEW

From the original article

Perplexity Research and turbopuffer have released pplx-embed-v2-context-9b-preview, a contextual embedding model for RAG pipelines. Each chunk is embedded with the full document in view. The real change is the training signal. The model learns to retrieve the answer along with the context needed to verify it, not one ‘gold passage.’ P

Is it deployable? Yes, as a self-hosted preview. Weights are on Hugging Face under the MIT license. Loading requires transformers>=5.4.0 with trust_remote_code=True. It is not yet on the Perplexity API. The model card notes that weights and interface may change without backward compatibility.

RAG systems split long documents into chunks. A chunk often depends on an entity, heading, or definition stated elsewhere. Contextual models address this with late chunking. The document is encoded in one pass, then pooled per chunk.

Training, however, usually marks one gold chunk per query. Every other chunk becomes a negative, including the sentences that make the

Source

MarkTechPost · 1 Oct 2026 · 05:23 CEST

Open the original at MarkTechPost ↗