arborist/docs/_source/arborist-one-pager.rst
russell@unturf.com 75ae470581
docs: rewrite arborist-one-pager + arborist-two-pager for external readers — drop ticket refs, schema versions, and internal vocabulary; reframe around user value (verified answers, fabricated-citation prevention, replay)
Old drafts read as internal substrate notes. Rewrites lead with what the system does for a consumer or evaluator and what it costs to run, with no references to internal tickets, table names, schema-version strings, governance hash dimensions, or per-record audit-mode tokens. Appendix diagrams updated in lockstep: friendly labels ("grounded / partly grounded / not grounded") replace the schema-column trichotomy, layer names paraphrased away from SURFACE/CORE/PROVIDENCE.
2026-05-14 10:02:08 -04:00

61 lines
2.9 KiB
ReStructuredText
Raw Blame History

This file contains ambiguous Unicode characters

This file contains Unicode characters that might be confused with other characters. If you think that this is intentional, you can safely ignore this warning. Use the Escape button to reveal them.

arborist — answers your AI can prove
=====================================
.. class:: center
*one-page summary · permacomputer.com · AGPL-3.0*
----
Most retrieval-augmented question-answering systems hand a language
model some context and ship whatever the model says. There is no way to
tell whether the answer faithfully reflects the source or whether the
model embroidered it. **arborist closes that gap.** Every answer is
verified against its source *after* the language model finishes, by a
mechanical checker — not by another AI grading the first one. The
checker labels each answer **grounded** (every claim was found in the
source), **partly grounded** (some claims, not all), or **not
grounded**. The label travels with the answer and is stored in a
tamper-evident chain back to the source bytes.
**Fabricated citations become impossible.** When the model answers, it
never types the quoted text. arborist tags each candidate source
chunk with a short label — ``E1``, ``E2`` — and asks the model to
answer using those labels. The model might write *"Jupiter is the
largest planet [E1], with a radius of about 70 000 km [E2]"*; arborist
renders the actual chunk text at display time. A model cannot fabricate
a quote it never types.
**The proof path is cheap and mechanical.** Verification is text
comparison, not embeddings, not similarity, not another model. It runs
on a laptop. Optional smart-ranking layers — cross-encoder rerankers,
entailment models — exist on the side; they help arborist find better
evidence, they do not influence whether an answer is certified.
**Same question, same document, same answer.** Answers are content-
addressed: the cache key folds in the document, the question, the
model identity, and the policy under which the answer was checked.
Two users asking the same question of the same document under the same
policy hit the same record. Reproducible. Replayable. Shareable.
**What it has measured on real traffic.**
- **100% misattribution catch at 0% false positives** — every answer
where the cited source was unrelated to the claim was flagged,
without a single grounded answer wrongly demoted, across the full
pooled test bed.
- **5565% topic-deflection catch at 00.4% false positives** — picks
off-topic answers out of the stream while leaving on-topic answers
untouched.
- **100% citation coverage on the curated textbook corpus** — every
cited claim resolves to a chain of evidence ending at a public-domain
or open-licensed source.
**What it costs.** Python 3.12. SQLite, one file (~2 GB for a
Wikipedia-sized corpus). No GPU for the proof path. Use any
OpenAI-compatible inference endpoint; the free reference endpoint is
`hermes.ai.unturf.com <https://hermes.ai.unturf.com>`_. Source:
`git.unturf.com/engineering/unturf/arborist
<https://git.unturf.com/engineering/unturf/arborist>`_. Full
whitepaper: `unfirehose.com/merkle-providence-reverse-rag.html
<https://unfirehose.com/merkle-providence-reverse-rag.html>`_.