real embeddingsThis is the anatomy of a retrieval-augmented answer. Pick a question, then move the levers on the left — the same ones tuned in production. Document scores and the retrieval-quality metrics are measured: the corpus is genuinely embedded (three models, three enrichment levels, computed offline) and the cosine math runs in your browser. Latency, tokens, and cost are modeled from published pricing. The LLM answer is pre-authored per quality tier — no API calls leave this page.