Integrated Inference for Embeddings
Introducing Pinecone Inference to streamline your AI workflow

Introducing Pinecone Inference to streamline your AI workflow

7/9/2024 · Gibbs Cullen

What this post added

This post introduces Pinecone Inference, a new API that provides easy and low-latency access to embedding and reranking models hosted on Pinecone's infrastructure. It highlights the `multilingual-e5-large` model as the initial offering and demonstrates its usage with a Python code snippet for generating embeddings. The post also mentions the flexibility to use other embedding models and providers, and the upcoming addition of more reranking models.

Read the original post ↗