Maxwell Grody

Ask the archive

Ask a question of eight years of Mindscape transcripts and watch the answer being built. A model drafts an answer from the retrieved passages, a second pass checks each claim against them, and the page shows every step. Hosted models are faster. The model on my own machine can take a minute.

How an answer is built

A retrieval step finds the passages and slips in one extra, connected passage the question did not ask for. A model drafts an answer from the passages without being told which one that is. A second model pass reviews every claim against the passages and removes what it judges unsupported. That review is itself a model's judgment and can be wrong in either direction. The answer is then hashed, and only after that is the extra passage revealed. The retriever runs on one machine in Rockville. The drafting and review model is your choice above, either the Qwen checkpoint on that machine or a hosted model.

Why the wall. If the drafting model knew which passage was the surplus it could favor it, and "the surplus made it into the answer" would measure the prompt rather than the passage. The graph is built so the reveal tool is unreachable before the answer is frozen; a test suite checks that topology on every change.

Passages are short quotations from the transcripts, linked to the episode. The model is the one you picked above, a Qwen checkpoint served locally or a hosted GLM or DeepSeek model; the retriever is Heart of Gold.