⚠️ Pending Update · 2026-08-29 Verification · Content may be outdated, please refer to official docs Updated: 2026-08-29 · Status: Pending Verification
Introduction
Retrieval-Augmented Generation (RAG) needs a vector store to hold embeddings and answer similarity queries. Free options fall into three deployment shapes: managed cloud free tiers, serverless starter tiers, and self-hosted open-source libraries. The right choice depends less on feature checklists than on your data scale and operational appetite — a 50K-vector prototype and a 10M-vector production index point to very different tools, and over-engineering the former or under-engineering the latter are equally common mistakes. Hybrid filtering is another dimension worth knowing: Qdrant, Weaviate, and pgvector all support pre-filtering by metadata before vector search, which lets you restrict results to a tenant, a date range, or a document set without a second query. Pinecone's serverless tier also supports this, but the syntax differs, so choosing a store with familiar filter semantics can save real engineering time when your schema grows complex.
Mainstream Options
| Option | Free Quota | Deployment |
|---|---|---|
| Supabase pgvector | 500MB database | Managed Postgres |
| Qdrant Cloud | 1GB cluster (~50K vectors) | Managed |
| Pinecone Starter | 1 index, 100K vectors, 1536 dim | Serverless |
| Weaviate WCS | 14-day sandbox | Managed |
| Chroma | Unlimited (local resources) | Self-hosted |
| Milvus Lite | Unlimited (local resources) | Embedded |
Call Example (Qdrant Cloud)
from qdrant_client import QdrantClient
from qdrant_client.models import Distance, VectorParams, PointStruct
client = QdrantClient(url="https://xxx.aws.cloud.qdrant.io", api_key="...")
client.recreate_collection(
collection_name="docs",
vectors_config=VectorParams(size=1536, distance=Distance.COSINE))
client.upsert(collection_name="docs",
points=[PointStruct(id=1, vector=[0.1]*1536, payload={"text": "hello"})])
hits = client.search(collection_name="docs", query_vector=[0.1]*1536, limit=5)
Selection Guide
For fewer than 100K vectors, Supabase pgvector is the least-friction choice — you reuse a managed Postgres instance and query with standard SQL, avoiding a separate component to operate, and most teams already have Postgres skills. Between 100K and 1M vectors, Qdrant Cloud or Pinecone Starter hit the sweet spot of managed convenience and adequate capacity. Above 1M vectors or when throughput matters, self-host Chroma or Milvus, or upgrade to a paid tier. To avoid lock-in, access your store through a standard interface like LangChain's VectorStore abstraction, so you can swap the backend without rewriting application code when scale or pricing forces a migration. Finally, measure recall against a brute-force baseline before trusting any vector store for production. Approximate nearest neighbor search trades recall for speed, and the default tuning on managed tiers prioritizes latency, which can silently drop relevant results. A small test set with known ground truth, compared against exact search, tells you whether your store's defaults are acceptable or need tightening — and this check is free regardless of which tier you use.
🚀 Get Started: One-Click Free API Access
Want to call all the free models above with a single API key, no need to sign up for each provider? Apishare.cc provides a unified API Key — one key, 100+ models, free models at zero cost.
👉 Register on Apishare.cc → Get your unified API Key
📊 Want to see more free model rankings? Check out the Sep 2026 Free LLM API Rankings →
Start on APIShare in three steps
Ready to try the options above? Three steps get you running:
- Create an account - open the APIShare free API registration page. An email address is all you need; no credit card required.
- Browse the free API catalog - head to the complete free API list and filter by text, image, audio, embedding, or multimodal. Each entry shows its free quota, rate limit, and availability status.
- Grab a key and integrate - generate an API key in your dashboard and paste it into your application. Every plan includes actively-updated APIs gateways covering every provider mentioned in this guide.
Already have an account? Use the APIShare login page, or visit the APIShare homepage for a full platform overview. Registration is free, and you can stop at any time.
Every outbound link in this guide carries a UTM parameter (utm_source=apishare_devto&utm_medium=referral&utm_campaign=free_api_article) for clean campaign attribution.
About the Free API Aggregator
The models covered in this guide are all served through the APIShare free API aggregator, which gives you one key for the whole catalog.
- Full model catalog: APIShare free API directory
- Sign up for a free trial key: Register and claim your API key