TL;DR
- The empty cell is real. No product combines {a genuine prose editor} × {your private corpus surfaced proactively as you write} × {on-device / EU by default} × {pay-once}. NotebookLM has the RAG but no editor and is cloud / EU-Enterprise-only; Craft has the editor + on-device but no corpus RAG; the only stack that hits all four — Obsidian + Smart Connections + Copilot-on-Ollama — exists only as a hand-configured hobbyist rig. Polish it and unify it, and the cell is ours.
- We're one product away, because we already own both halves. ollwrite (the editor) + oll-memory (ingest / retrieve / extract, on-device via Ollama) = the substrate. The MVP is a "proactive memory sidecar": drop a folder → relevant docs surface beside the paragraph you're writing → cite / chat, all local.
- The buyer writes from prior work, and the ROI is billable. Beachhead = solo / boutique consultants + analysts (bill $190–510/hr, save 10+ hrs/wk). Undercut the all-subscription field with honest pay-once — genuinely honest because on-device inference is ~$0/run: the one architecture where "pay once, run forever" isn't bait.
Third in a line: Beyond the Wrapper (the platform thesis) → Platform Extensions build ledger (the primitives, built) → this (the first product those primitives compose into). This doc is the product-and-platform thesis for ollwrite + oll-memory specifically.
1 · The product — three panes, retrieval as a background sense
Your docs, auto-surfaced
Proactive, semantic. Relevant to the current paragraph — no query typed.
The intelligent editor we already have
ollwrite: inline ops, word-diff, accept / reject. The prose surface, unchanged.
…the indemnity dispute ⚠ churn figure — no source found reshaped how we…
Chat with your corpus
Jump-to-chunk citations. Ask it what you said, when, where.
The disciplines that separate magic from gimmick
Precision over recall
3 dead-on memories beat 15 fuzzy ones. An empty pane is a feature — surfacing junk is worse than surfacing nothing, because it burns the trust the whole product runs on.
Provenance you trust instantly
Every card is who / when; one click lands on the real source at the real span. Trust is the product; groundedness is the mechanism.
It surprises you with your own life
The forgotten thread you didn't know to search for. The "how did it know that" moment is the retention hook — and only proactive retrieval produces it.
Retrieval output is never inert text — it is always an actionable object: drag-into-draft-as-citation, click-to-expand, or dismiss. And dismissals train relevance — the corpus learns what you consider noise, so precision compounds per-user without a subscription treadmill.
2 · The two flagship workflows
A · "Write a book from 15 years of email"
Ingest is a live ledger, not a spinner. Drop the archive → watch it resolve: 48,213 emails → 31,904 after dedup → embedding on-device. Then a passive "What I found" surface turns a scary blob into something browsable — recurring people, a timeline, topics by year. The corpus stops being a wall and becomes a map.
As you write, the left pane silently repopulates (~600ms debounce, keyed on the current paragraph) with memory cards — provenance + snippet + date — draggable straight into prose as citations. The right pane interrogates: "what did I say about the Zurich deal that year?" → cited answers that jump to the exact chunk. You are writing from your life, with your life doing the remembering.
B · "Draft a report from my research folder + bullet points"
Ingest the folder and learn the report skeleton — structured-extract the section order, headings, typical length, and voice from the user's own past reports. Type bullets → hit "Draft from bullets" → each bullet expands into prose in the learned house structure + voice, pulling cited evidence, delivered as accept / reject diffs. The blank page never happens.
Gap detection is the differentiator. Any claim it can't ground gets an inline amber ⚠ churn figure — no source found. Chips aggregate into a right-pane "Before this is done, I need:" checklist; resolving inline or in chat resolves both. Plus "Expand with sources" on any thin paragraph. The tool that tells you what it doesn't know is the tool you trust with billable work.
3 · The four "can't-write-without-this" moments
1 · It remembered for you
The forgotten memory surfaces itself. You didn't search — it knew. The single most defensible feeling: no reactive tool can produce it.
2 · Bullets → a cited draft
In your voice, grounded in your evidence. The blank page — the writer's most expensive minute — simply never happens.
3 · It admits what it doesn't know
Gap-chips. The anti-hallucination, trust-earning moment — the opposite of a confident wrong answer, which is what kills reliance on AI writing.
4 · One click to the source
Every claim is one click from where it came from. Groundedness becomes reliance; reliance becomes a paid habit.
4 · The competitor map — the intersection nobody owns
| Product | Corpus RAG | Proactive surface-while-writing | Real long-form editor | On-device | Non-Enterprise EU | Pricing |
|---|---|---|---|---|---|---|
| NotebookLM best RAG, but siloed notebooks · ~13% hallucination on failure | Yes | No | No editor | No | No | Free / $4.99 / $19.99 / $99.99+ |
| Notion AI / Q&A shallow at scale · lags >3–5k words | Yes (shallow) | No (pull) | Weak | No | Enterprise-only | Full AI needs $20 Business + $10/1k credits |
| Saga conceptually closest — but cloud-only | Yes (mention-triggered) | Partial | Good / light | No | No | $6 / $12 (cheapest) |
| Craft editor + on-device — but no retrieval · export lock-in is gripe #1 | No corpus RAG | No | Best editor | Yes (on-device models) | CHF-priced | Free / ~CHF-tier subs |
| Sudowrite best fiction model (Muse) · Story Bible is manual, not your corpus | Manual Story Bible | No | Strong | No | No | $10 / $22 / $44 · credits |
| Lex / Type | None / manual-attach | No | Strong | No | No | ~$8–16/mo |
| Scrivener the pay-once precedent | None (no AI) | No | Strong | Yes (local) | — | $59.99 one-time |
| Obsidian + Smart Connections + Copilot (Ollama) the real feature-competitor — a 2-plugin science project | Yes | Yes (SC = only proactive surfacer) | Markdown-hacker | Yes (local) | Yes (local) | Copilot $14.99/mo or $349.99 one-time |
| Local-LLM tools (GPT4All / AnythingLLM / Msty) | Reactive RAG | No | No editor | Yes | Yes | Free / Msty $349 lifetime |
5 · Threat ranking — and how we beat each
1 NotebookLM — free + Google's gravity
Beat with a real editor, cross-corpus context, and on-device / EU. One line: "NotebookLM you can actually write in, over ALL your files, that never leaves your machine." Its notebooks are silos; ours is your whole knowledge. Its answers fabricate on failure; ours admit gaps.
2 Obsidian + plugins — the real feature-competitor
Beat with zero-config polish. Collapse the science project into one app: local by default, a clean prose editor instead of a markdown pane, no CORS / model-picker / RAM ritual. Same capability, none of the assembly. This is the head-to-head match, and product craft is the whole edge.
3 Notion AI — incumbent surface area
Beat with proactive not pull, a long-form editor that doesn't lag past 3–5k words, and non-Enterprise EU / on-device. Notion makes you ask; we surface. Notion gates residency behind Enterprise; we make it the default.
4 Craft — best editor, no retrieval
Beat by adding the retrieval layer it lacks entirely + fixing export lock-in with open formats. Craft proves on-device editing sells; it just never connected to your corpus. We start where it stops.
6 · Platform primitives to add — the P1–P7 ladder
| Primitive | What it is | Features it unlocks | Effort | Reuse |
|---|---|---|---|---|
| P1 · Context-assembly / proactive retrieval | Embed the current paragraph → hybrid-retrieve + rerank → the side rail | The whole proactive-surface magic; the left pane | S | Every writing surface |
| P2 · Folder / bulk on-device ingestion | Drop a folder → parse PDF / docx / md via LlamaIndex readers → embed locally | The corpus itself; the ingest ledger | S–M | Foundational — everything |
| P3 · Incremental / watch sync | Hash-diff, re-embed only deltas | Always-fresh corpus without re-ingest | M | Retention layer |
| P4 · Style / voice profile | Extract a style rulebook from the corpus → inject as system rules | "Draft in your voice"; the ghostwriter wedge | M | The differentiator |
| P5 · Pattern / structure extraction | The recurring skeleton across past docs → a reusable template | "Draft from bullets" in house format; the free SEO tool | M | Highest B2B value + top-of-funnel |
| P6 · Citation / provenance tracking | Every span carries its source chunk | Click-to-source; gap-chips; the trust layer | S–M | Required for legal / academic |
| P7 · Per-user private corpus + ACL | Namespace by oll-core JWT | Teams, shared corpora, seat pricing | M | Unlocks B2B / teams |
7 · Who pays — and pricing's undercut lane
| Segment | Pain | ROI | Willingness to pay |
|---|---|---|---|
| Consultants / analysts ← beachhead | Re-reading old decks / notes to write the next one | Billable $190–510/hr, save 10+ hrs/wk | $30–80/mo or $99–199 pay-once |
| Grant writers | Recomposing boilerplate across proposals | Faster submissions, more grants out | Comps $50–150/mo (Grantable) → undercut $99–149 pay-once |
| Lawyers | Provenance + privilege; cloud AI is a no | Groundedness for citeable work; on-device for privilege | $30–60/mo private tier — white space vs Harvey / Spellbook $100–350/seat |
| Authors / ghostwriters | Blank page; matching a client's voice | $15–50k/book; voice-profile is the wedge | Sudowrite's turf — prosumer pay-once |
| Researchers / academics | Synthesizing a reading pile into prose | Cited drafts from their own library | Price-sensitive: $9–29 |
| Agencies | House voice + house format at scale | P4 + P5 + P7 = consistent output across a team | $20–40/seat |
The honest pay-once point
The entire AI-writing category is subscription or credits; pay-once is nearly unclaimed — only Scrivener ($60) and Msty / Copilot ($349 lifetime) hold it. Our points: personal pay-once 29–49 · pro / consultant 99–149 · team 39/seat/mo.
And here's why it's defensible, not a gimmick: 2026 lifetime-deals now cap credits because cloud inference costs money on every single run — so "lifetime" quietly means "lifetime of a small credit bucket." On-device / self-hosted-Ollama is the one architecture where "pay once, run forever on your hardware" is genuinely true (~$0/run), not bait. That is an ethical and a competitive differentiator, and it is only ours because the Model gateway already routes to on-device.
8 · Exploitable gripes
Credit / usage anxiety
Sudowrite, Cursor's June-2025 20× backlash → public apology + refunds, Coda, Notion's $10/1k credits. → Flat / pay-once wins the trust.
Shallow retrieval
NotebookLM fabricates + silos; Notion misses at scale. → Precision-first, cited, whole-corpus.
Laggy editors
Notion degrades past 3–5k words. → A real long-form surface.
Lock-in / export
Craft export called "a disaster." → Open formats, your data stays yours.
Privacy opacity
EU residency paywalled behind Enterprise. → On-device / EU by default.
Plugin jank
Obsidian's proactive rig is a config marathon. → Zero-config polish is the moat.
9 · Co-optimized distribution
P5 → a free public "Templatize your report" tool. Drop 3 past reports → get your reusable house-format skeleton, no login (via the guest-checkout pattern). That's an SEO front door → an anonymized template / pattern library as evergreen long-tail SEO → profession communities (r/consulting, analyst Slacks, grant-writer forums, r/writing) via the honest draft-only reply pipeline → and crucially, on-device / private as the headline, which is the one thing that lets us post in confidential-work communities (legal, corporate strategy) where "cloud AI" is an instant no.
10 · Sequenced roadmap + recommendation
- Night 1–2 — First paying product: P2 + P1 + P6 into ollwrite → "private writing memory," pay-once 29–49 on existing Stripe rails. Ingest → proactive surface → cited. Shippable on day two.
- Night 3–4 — P4 voice profile → the 99 Pro tier; ghostwriters + authors. The differentiator lands.
- Night 5–6 — P5 pattern-extraction shipped as a premium feature AND the free "templatize" tool → first B2B + SEO on.
- Week 2 — P7 multi-tenant ACL → teams 39/seat; legal + agencies.
- Ongoing — P3 watch-sync → retention; the corpus stays fresh without re-ingest.
Recommendation (propose-only)
The empty cell is real and we own both halves — this is the highest-conviction net-new product in the portfolio, and the roadmap ships a chargeable slice by night two. P2 + P1 + P6 into ollwrite is the move, with P4 as the premium wedge right behind it.
⚠ Ship-vs-build flag (the honest part)
This is a net-new build while the humaniz.me first-franc path sits at the LIVE-Stripe finish line. The guardrail says: close the first stranger dollar first — or run this strictly as a parallel overnight track — not jump off a finished asset to start a new one. This doc is the roadmap for after / alongside the first franc, not instead of it. It feeds Strategy + Backlog; they decide, this informs.
Follows Beyond the Wrapper · executes on the Platform Extensions build ledger. Propose-only · 2026-07-05 · reference links are honest and directional; public build-in-public figures are not audited.