Gate 1
Retrieval contract
Define queries, relevant items, hard negatives, top-k, recall or relevance threshold, latency objective, freshness, filters, empty-result behavior, and agent fallback.
Retrieval evaluation before migration
Describe your agent memory or RAG corpus, embedding, dimensions, access patterns, partition candidates, filters, quality labels, latency, freshness, security, and budget. Generate a shadow-test plan instead of treating a service benchmark as application proof.
Gate 1
Define queries, relevant items, hard negatives, top-k, recall or relevance threshold, latency objective, freshness, filters, empty-result behavior, and agent fallback.
Gate 2
Record embedding model and version, dimension, normalization, item size, vector-index partition key, distribution, filter fields, write rate, updates, deletes, and backfill.
Gate 3
Test tenant isolation, IAM, encryption, sensitive-data controls, retention, deletion proof, backups, logs, model migration, feature flag, and rollback.
Gate 4
Measure concurrency, throttling, retries, index and table operations, storage, embedding generation, transfer, observability, and cost per accepted agent answer.
Freeze the corpus, embedding version, dimensions, partition and filter scheme, relevance labels, query set, top-k, and thresholds. Replay expected and burst traffic beside the current retriever; compare quality, latency, freshness, throttling, delete behavior, full cost, and final answer safety. Keep shadow results hidden from users and exercise the fallback.
Quality, latency, freshness, isolation, operations, and cost all pass on representative data.
Partition skew, filters, dimensions, relevance labels, deletion, or model migration is unresolved.
The new path does not improve accepted answers or operations enough to justify migration risk.
Official facts checked August 15, 2026. Recheck current API, quotas, dimensions, consistency, pricing, and Regions.