perplexity-ai / pplx-embed-v2-context-9b-preview

huggingface.co
Total runs: 305
24-hour runs: 160
7-day runs: 301
30-day runs: 301
Model's Last Updated: September 30 2026
feature-extraction

Introduction of pplx-embed-v2-context-9b-preview

Model Details of pplx-embed-v2-context-9b-preview

pplx-embed-v2-context-9b-preview

pplx-embed-v2-context-9b-preview is a contextual embedding model for document chunks in RAG systems. A document is passed as a list of chunks; the chunks are encoded together, so each chunk's embedding reflects its surrounding context, and one embedding is returned per chunk.

This is a preview release, not a final model. Weights, embeddings, and the interface may change in later versions without backward compatibility, so embeddings produced with this preview should not be mixed with embeddings from a future release.

Queries and documents are encoded with different methods : use encode_queries for queries and encode for document chunks. The model is trained with separate query and document prefixes, and encoding queries with encode silently degrades retrieval quality.

Like pplx-embed-context-v1 , the model natively produces unnormalized int8-quantized embeddings. Compare embeddings with cosine similarity , or pass normalize_embeddings=True and use the dot product.

Model
Model Dimensions MRL Quantization Instruction Pooling
pplx-embed-v2-context-9b-preview 2048 1024, 2048 INT8 No (fixed query/document prefixes) Mean
Usage

Requires transformers>=5.4.0 , torch , numpy , safetensors , and tqdm . The model uses custom code, so load it with trust_remote_code=True .

from transformers import AutoModel

model = AutoModel.from_pretrained(
    "perplexity-ai/pplx-embed-v2-context-9b-preview",
    trust_remote_code=True,
).to("cuda")

doc_chunks = [
    [
        "Curiosity begins in childhood with endless questions about the world.",
        "As we grow, curiosity drives us to explore new ideas.",
        "Scientific breakthroughs often start with a curious question.",
    ],
    [
        "The curiosity rover explores Mars searching for ancient life.",
        "Each discovery on Mars sparks new questions about the universe.",
    ],
]

# One (chunk_count, 2048) array per document:
# doc_embeddings[0].shape == (3, 2048), doc_embeddings[1].shape == (2, 2048)
doc_embeddings = model.encode(doc_chunks, normalize_embeddings=True)

# Each query is a single-chunk row.
queries = [["What drives scientific breakthroughs?"]]
query_embeddings = model.encode_queries(queries, normalize_embeddings=True)

scores = doc_embeddings[0] @ query_embeddings[0][0]
Options

encode(documents, ...) and encode_queries(queries, ...) accept:

Argument Default Description
batch_size 32 Documents (or queries) per forward pass
normalize_embeddings False L2-normalize outputs
convert_to_numpy True Return NumPy arrays; False returns CPU tensors
show_progress_bar False Show a progress bar
device None Move the model to this device before encoding
Matryoshka dimensions

The model was trained with Matryoshka losses at 1024 and 2048 dimensions. To use 1024-dimensional embeddings, take the first 1024 values of each unnormalized embedding and normalize afterwards:

import numpy as np

emb = model.encode(doc_chunks)  # unnormalized int8 values
emb_1024 = [e[:, :1024] / np.linalg.norm(e[:, :1024], axis=-1, keepdims=True) for e in emb]

Other truncation sizes were not trained.

Runs of perplexity-ai pplx-embed-v2-context-9b-preview on huggingface.co

305
Total runs
160
24-hour runs
301
3-day runs
301
7-day runs
301
30-day runs

More Information About pplx-embed-v2-context-9b-preview huggingface.co Model

More pplx-embed-v2-context-9b-preview license Visit here:

https://choosealicense.com/licenses/mit

pplx-embed-v2-context-9b-preview huggingface.co

pplx-embed-v2-context-9b-preview huggingface.co is an AI model on huggingface.co that provides pplx-embed-v2-context-9b-preview's model effect (), which can be used instantly with this perplexity-ai pplx-embed-v2-context-9b-preview model. huggingface.co supports a free trial of the pplx-embed-v2-context-9b-preview model, and also provides paid use of the pplx-embed-v2-context-9b-preview. Support call pplx-embed-v2-context-9b-preview model through api, including Node.js, Python, http.

pplx-embed-v2-context-9b-preview huggingface.co Url

https://huggingface.co/perplexity-ai/pplx-embed-v2-context-9b-preview

perplexity-ai pplx-embed-v2-context-9b-preview online free

pplx-embed-v2-context-9b-preview huggingface.co is an online trial and call api platform, which integrates pplx-embed-v2-context-9b-preview's modeling effects, including api services, and provides a free online trial of pplx-embed-v2-context-9b-preview, you can try pplx-embed-v2-context-9b-preview online for free by clicking the link below.

perplexity-ai pplx-embed-v2-context-9b-preview online free url in huggingface.co:

https://huggingface.co/perplexity-ai/pplx-embed-v2-context-9b-preview

pplx-embed-v2-context-9b-preview install

pplx-embed-v2-context-9b-preview is an open source model from GitHub that offers a free installation service, and any user can find pplx-embed-v2-context-9b-preview on GitHub to install. At the same time, huggingface.co provides the effect of pplx-embed-v2-context-9b-preview install, users can directly use pplx-embed-v2-context-9b-preview installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

pplx-embed-v2-context-9b-preview install url in huggingface.co:

https://huggingface.co/perplexity-ai/pplx-embed-v2-context-9b-preview

Url of pplx-embed-v2-context-9b-preview

pplx-embed-v2-context-9b-preview huggingface.co Url

Provider of pplx-embed-v2-context-9b-preview huggingface.co

perplexity-ai
ORGANIZATIONS

Other API from perplexity-ai

huggingface.co

Total runs: 831
Run Growth: 275
Growth Rate: 33.09%
Updated:February 27 2025
huggingface.co

Total runs: 110
Run Growth: 91
Growth Rate: 82.73%
Updated:February 07 2026