Skip to main content
Pinecone Docs

Search documentation

Type to search this documentation.

all-mpnet-base-v2

Use the all-mpnet-base-v2 embedding or reranking model with Pinecone: specs and index setup. all-mpnet-base-v2 is a sentence and short paragraph encoder.

This model has the best quality of the sbert all family of models.

### Using the model

#### Installation:

```python theme={null}

!pip install -U sentence-transformers pinecone

```

### Create Index

```python theme={null}

from pinecone import Pinecone, ServerlessSpec

pc = Pinecone(api_key="API_KEY")

index_name = "all-mpnet-base-v2"

if not pc.has_index(index_name):
    pc.create_index(
        name=index_name,
        dimension=768,
        metric="cosine",
        spec=ServerlessSpec(
            cloud='aws',
            region='us-east-1'
        )
    )

index = pc.Index(index_name)

```

### Embed & Upsert

```python theme={null}

from sentence_transformers import SentenceTransformer
import torch
device = 'cuda' if torch.cuda.is_available() else 'cpu'

model = SentenceTransformer('sentence-transformers/all-mpnet-base-v2').to(device)

data = [
    {"id": "vec1", "text": "Apple is a popular fruit known for its sweetness and crisp texture."},
    {"id": "vec2", "text": "The tech company Apple is known for its innovative products like the iPhone."},
    {"id": "vec3", "text": "Many people enjoy eating apples as a healthy snack."},
    {"id": "vec4", "text": "Apple Inc. has revolutionized the tech industry with its sleek designs and user-friendly interfaces."},
    {"id": "vec5", "text": "An apple a day keeps the doctor away, as the saying goes."},
]
sentences = [x["text"] for x in data]
embeddings =  model.encode(sentences)  

vectors = []
for d, e in zip(data, embeddings):
    vectors.append({
        "id": d['id'],
        "values": e,
        "metadata": {'text': d['text']}
    })

index.upsert(
    vectors=vectors,
    namespace="ns1"
)

```

### Query

```python theme={null}

query = "Tell me about the tech company known as Apple"

query_embedding = model.encode(query).tolist()
print(query_embedding)

results = index.query(
    namespace="ns1",
    vector=query_embedding,
    top_k=3,
    include_values=False,
    include_metadata=True
)

print(results)

```

Lorem Ipsum

Suggest an edit

Propose a replacement for this page. The site team reviews it before applying any changes.

Export
Documentation menu