Memory your agents can afford. Predictable pricing.
Entity-scoped memory, semantic recall, source-backed retrieval, and storage in a single API. Stop paying separately for S3 + Pinecone + embedding APIs.
Free
Solo NodeEverything you need to give one prototype agent durable memory.
- Remembers up to 500 facts or notes
- 100K indexed memory chunks
- 10K recalls / month
- 1 GB memory storage
- Python, TypeScript, Go, .NET SDKs
- Community support
Pro
MeshEntity-scoped memory for teams shipping AI-native products.
Start on the Free plan — upgrade to Pro anytime.
- Remembers up to 5,000 facts or notes
- 1M indexed memory chunks
- 100K recalls / month
- 25 entity partitions
- 15 GB memory storage
- Replication & sharding
Scale
ClusterHigh-volume memory for distributed agent fleets.
Start on the Free plan — upgrade to Scale anytime.
- Remembers up to 25,000 facts or notes
- 5M indexed memory chunks
- 500K recalls / month
- 1,000 entity partitions
- Backup snapshots
- Isolation proof for tenants
Business
ConstellationMemory platform for teams that need security, compliance, and priority support.
- Remembers up to 100,000 facts or notes
- 15M indexed memory chunks
- 2M recalls / month
- 10,000 entity partitions
- 150 GB memory storage
- SSO, RBAC & audit log
Replace 3 services
Storage + vectors + embeddings in one API. No S3, no Pinecone, no embedding API bills.
Your data, isolated
Encrypted, single-tenant storage on Aether's infrastructure — export or delete your data anytime.
Scales with you
From your first request to high throughput, scaling and redundancy are handled for you.
A consolidated cost model
Compare one plan with a four-service stack
A self-assembled reference stack combines vector reads, storage egress, embedding tokens, and per-search reranking. The comparisons below use dated 2026 list prices and each Aether plan's included capacity. Paid Aether plans publish separate overage units for usage beyond that capacity.
Pro
Save 81%Same workload on a metered Pinecone + S3 + embeddings + rerank stack.
Scale
Save 81%Same workload on a metered Pinecone + S3 + embeddings + rerank stack.
Business
Save 88%Same workload on a metered Pinecone + S3 + embeddings + rerank stack.
Compare the published pricing inputs
Estimate a representative self-assembled stack from dated list prices, then compare it with the Aether plan selected for the same workload. The calculator is a planning model, not a quote.
Compare every feature
Predictable baseline pricing with clear usage-based overages beyond included capacity. Every plan includes all SDKs, all LLM integrations, and local ONNX embeddings.
Feature comparison
Free
Everything you need to start building RAG pipelines locally.
Capacity
- Documents
- 500
- Vectors
- 100K
- Queries / month
- 10K
- Storage
- 1 GB
- Replication factor
- 1x
- API keys
- 2
- Archive bytes / month
- Not included
- Partitions
- 3
Infrastructure
- Multi-node replication
- No
- Data sharding
- No
- Federated search
- No
- Prometheus metrics
- No
- Backup snapshots
- No
- Custom embedding models
- No
- Permanent archive (Arweave)
- No
- Metadata filtering
- No
- Recency-weighted ranking
- No
- Freshness ranking
- No
Security & Compliance
- RBAC
- No
- SSO (SAML / OIDC)
- No
- Audit log
- No
- Provable tenant isolation
- No
- Uptime SLA
- None
Support
- Community (GitHub / Discord)
- Yes
- Email support
- No
Pro
Replicated clusters for teams shipping AI-native products.
Capacity
- Documents
- 5,000
- Vectors
- 1M
- Queries / month
- 100K
- Storage
- 15 GB
- Replication factor
- 2x
- API keys
- 10
- Archive bytes / month
- 1 GB
- Partitions
- 25
Infrastructure
- Multi-node replication
- Yes
- Data sharding
- Yes
- Federated search
- Yes
- Prometheus metrics
- Yes
- Backup snapshots
- No
- Custom embedding models
- No
- Permanent archive (Arweave)
- No
- Metadata filtering
- No
- Recency-weighted ranking
- No
- Freshness ranking
- No
Security & Compliance
- RBAC
- No
- SSO (SAML / OIDC)
- No
- Audit log
- No
- Provable tenant isolation
- No
- Uptime SLA
- None
Support
- Community (GitHub / Discord)
- Yes
- Email support
- (48hr SLA)
Scale
Expanded capacity for globally distributed agent fleets.
Capacity
- Documents
- 25,000
- Vectors
- 5M
- Queries / month
- 500K
- Storage
- 50 GB
- Replication factor
- 2x
- API keys
- 50
- Archive bytes / month
- 50 GB
- Partitions
- 1,000
Infrastructure
- Multi-node replication
- Yes
- Data sharding
- Yes
- Federated search
- Yes
- Prometheus metrics
- Yes
- Backup snapshots
- Yes
- Custom embedding models
- No
- Permanent archive (Arweave)
- No
- Metadata filtering
- Yes
- Recency-weighted ranking
- Yes
- Freshness ranking
- Yes
Security & Compliance
- RBAC
- No
- SSO (SAML / OIDC)
- No
- Audit log
- No
- Provable tenant isolation
- Yes
- Uptime SLA
- 99.9%
Support
- Community (GitHub / Discord)
- Yes
- Email support
- (24hr SLA)
Business
Full platform with security, compliance, and priority support.
Capacity
- Documents
- 100,000
- Vectors
- 15M
- Queries / month
- 2M
- Storage
- 150 GB
- Replication factor
- 3x
- API keys
- Unlimited
- Archive bytes / month
- 100 GB
- Partitions
- 10,000
Infrastructure
- Multi-node replication
- Yes
- Data sharding
- Yes
- Federated search
- Yes
- Prometheus metrics
- Yes
- Backup snapshots
- Yes
- Custom embedding models
- Yes
- Permanent archive (Arweave)
- No
- Metadata filtering
- Yes
- Recency-weighted ranking
- Yes
- Freshness ranking
- Yes
Security & Compliance
- RBAC
- Yes
- SSO (SAML / OIDC)
- Yes
- Audit log
- Yes
- Provable tenant isolation
- Yes
- Uptime SLA
- 99.95%
Support
- Community (GitHub / Discord)
- Yes
- Email support
- (4hr SLA)
Feature comparison
Capacity
| Feature | Free tier | Pro tier | Scale tier | Business tier |
|---|---|---|---|---|
| Documents | 500 | 5,000 | 25,000 | 100,000 |
| Vectors | 100K | 1M | 5M | 15M |
| Queries / month | 10K | 100K | 500K | 2M |
| Storage | 1 GB | 15 GB | 50 GB | 150 GB |
| Replication factor | 1x | 2x | 2x | 3x |
| API keys | 2 | 10 | 50 | Unlimited |
| Archive bytes / month | Not included | 1 GB | 50 GB | 100 GB |
| Partitions | 3 | 25 | 1,000 | 10,000 |
Infrastructure
| Feature | Free tier | Pro tier | Scale tier | Business tier |
|---|---|---|---|---|
| Multi-node replication | No | Yes | Yes | Yes |
| Data sharding | No | Yes | Yes | Yes |
| Federated search | No | Yes | Yes | Yes |
| Prometheus metrics | No | Yes | Yes | Yes |
| Backup snapshots | No | No | Yes | Yes |
| Custom embedding models | No | No | No | Yes |
| Permanent archive (Arweave) | No | No | No | No |
| Metadata filtering | No | No | Yes | Yes |
| Recency-weighted ranking | No | No | Yes | Yes |
| Freshness ranking | No | No | Yes | Yes |
Security & Compliance
| Feature | Free tier | Pro tier | Scale tier | Business tier |
|---|---|---|---|---|
| RBAC | No | No | No | Yes |
| SSO (SAML / OIDC) | No | No | No | Yes |
| Audit log | No | No | No | Yes |
| Provable tenant isolation | No | No | Yes | Yes |
| Uptime SLA | None | None | 99.9% | 99.95% |
Support
| Feature | Free tier | Pro tier | Scale tier | Business tier |
|---|---|---|---|---|
| Community (GitHub / Discord) | Yes | Yes | Yes | Yes |
| Email support | No | (48hr SLA) | (24hr SLA) | (4hr SLA) |
Need hyperscale or enterprise controls?
Our Enterprise Sovereign tier offers unlimited scaling, strict tenant isolation, and custom embedding models.
Published plan data
Plan names, prices, and included units
Monthly queries are measured per month. Annual prices are effective monthly rates billed annually. Published overage units appear only where the current plan data defines them.
| Plan | Monthly | Annual billing | Documents | Stored vectors | Monthly queries | Storage | Partitions | Published overages |
|---|---|---|---|---|---|---|---|---|
| Free | $0/month | $0/month, billed annually | 500 | 100,000 | 10,000 | 1 GB | 3 | No published overage units |
| Pro | $49/month | $39/month, billed annually | 5,000 | 1,000,000 | 100,000 | 15 GB | 25 | $2 per 1K queries · $5 per 500 documents · $1 per 100K vectors · $3 per 1 GB storage |
| Scale | $199/month | $159/month, billed annually | 25,000 | 5,000,000 | 500,000 | 50 GB | 1,000 | $1.5 per 1K queries · $4 per 1K documents · $0.8 per 100K vectors · $2.5 per 1 GB storage |
| Business | $499/month | $399/month, billed annually | 100,000 | 15,000,000 | 2,000,000 | 150 GB | 10,000 | $1 per 1K queries · $3 per 1K documents · $0.5 per 100K vectors · $2 per 1 GB storage |
| Enterprise | Custom | Custom | Unlimited | Unlimited | Unlimited | Unlimited | Unlimited | No published overage units |
Pricing questions
The same plan source powers pricing, billing, and dashboard limits, so the numbers stay aligned across Aether.
What happens when I hit a plan limit?
Free-tier hard limits return 402 with a plan-limit error code. Paid plans include a grace buffer for annual billing and then bill overages where the plan supports them. Rate limits return 429 with a retry-after header so clients can back off cleanly.
How do upgrades and downgrades work?
You can change plans from the billing page in the Aether dashboard. Upgrades take effect immediately. Downgrades apply at the next billing period and remain available only when current usage fits inside the target plan's limits.
What is the refund policy?
Monthly and annual subscriptions are cancellable from the dashboard. Refund requests are reviewed case by case for duplicate charges, accidental upgrades, and service-impacting incidents.
When should I talk to sales?
Talk to sales for Enterprise, dedicated environments, compliance reviews, custom embedding models, high-volume migration support, or limits beyond the published plans.
Start building with Aether today.
Free forever on the Solo Node plan. No credit card required. Upgrade when you need replication, redundancy, and enterprise features.
© 2026 Quintessence Group, Inc. All rights reserved.