How it works

From upload to answer.

Five steps. The slow part, reading, is left to your LLM, and it reads very little.

The pipeline

Five steps. No vectors.

  1. 01

    Upload

    Docs, PDFs, Office files, e-books, zips, audio or a GitHub repo.

    Any text
  2. 02

    Index

    A trigram index of every line, and a brief of every file.

    0.1 s
  3. 03

    Ask

    A question in plain words. Grape picks the search terms.

    Plain words
  4. 04

    Search

    Paragraphs ranked by how much of the question they cover.

    10 ms
  5. 05

    Answer

    Your LLM reads only those few passages.

    1 call
Two shortcuts

Less work than RAG, twice.

0LLM calls for an off-topic question. The search alone refuses it.
0.1 sUntil a changed document is searchable. Vector RAG re-embeds it for 98 to 140 s.