7/9/2024 · Gibbs Cullen
What this post added
This post introduces Pinecone Inference, a new API that provides easy and low-latency access to embedding and reranking models hosted on Pinecone's infrastructure. It highlights the `multilingual-e5-large` model as the initial offering and demonstrates its usage with a Python code snippet for generating embeddings. The post also mentions the flexibility to use other embedding models and providers, and the upcoming addition of more reranking models.