Product thesis · propose-only · 2026-07-05

Write From Your Own Knowledge — ollwrite × oll-memory

We already have the editor (ollwrite) and now the memory (oll-memory). Welding them is a product no one ships: a real writing surface with your private document corpus as live, proactive context — on-device, pay-once. Here's the thesis, the competitors, the primitives, and who pays. A follow-up to Beyond the Wrapper; it feeds Strategy + Backlog, it does not decide them.

TL;DR

  1. The empty cell is real. No product combines {a genuine prose editor} × {your private corpus surfaced proactively as you write} × {on-device / EU by default} × {pay-once}. NotebookLM has the RAG but no editor and is cloud / EU-Enterprise-only; Craft has the editor + on-device but no corpus RAG; the only stack that hits all four — Obsidian + Smart Connections + Copilot-on-Ollama — exists only as a hand-configured hobbyist rig. Polish it and unify it, and the cell is ours.
  2. We're one product away, because we already own both halves. ollwrite (the editor) + oll-memory (ingest / retrieve / extract, on-device via Ollama) = the substrate. The MVP is a "proactive memory sidecar": drop a folder → relevant docs surface beside the paragraph you're writing → cite / chat, all local.
  3. The buyer writes from prior work, and the ROI is billable. Beachhead = solo / boutique consultants + analysts (bill $190–510/hr, save 10+ hrs/wk). Undercut the all-subscription field with honest pay-once — genuinely honest because on-device inference is ~$0/run: the one architecture where "pay once, run forever" isn't bait.

Third in a line: Beyond the Wrapper (the platform thesis) → Platform Extensions build ledger (the primitives, built) → this (the first product those primitives compose into). This doc is the product-and-platform thesis for ollwrite + oll-memory specifically.

1 · The product — three panes, retrieval as a background sense

The spine of the whole thesis: retrieval is a background sense, not a button. You don't go and search your corpus — it surfaces itself, keyed on the sentence under your cursor. That single inversion is what separates this from every "chat your docs" tool.
◀ Left · Memory

Your docs, auto-surfaced

Proactive, semantic. Relevant to the current paragraph — no query typed.

email · A. Roth · 2019-03"…we walked away from the Zurich deal over the indemnity clause…"
report.pdf · p.12 · 2021"churn fell to 4.1% after onboarding rework."
● Middle · Editor

The intelligent editor we already have

ollwrite: inline ops, word-diff, accept / reject. The prose surface, unchanged.

…the indemnity dispute ⚠ churn figure — no source found reshaped how we…

▶ Right · Chat

Chat with your corpus

Jump-to-chunk citations. Ask it what you said, when, where.

answer · 3 sources"You raised the indemnity concern in 3 threads, Mar–Jun 2019 → open source ↗"

The disciplines that separate magic from gimmick

Precision over recall

3 dead-on memories beat 15 fuzzy ones. An empty pane is a feature — surfacing junk is worse than surfacing nothing, because it burns the trust the whole product runs on.

Provenance you trust instantly

Every card is who / when; one click lands on the real source at the real span. Trust is the product; groundedness is the mechanism.

It surprises you with your own life

The forgotten thread you didn't know to search for. The "how did it know that" moment is the retention hook — and only proactive retrieval produces it.

Retrieval output is never inert text — it is always an actionable object: drag-into-draft-as-citation, click-to-expand, or dismiss. And dismissals train relevance — the corpus learns what you consider noise, so precision compounds per-user without a subscription treadmill.

2 · The two flagship workflows

Concrete, because a thesis you can't picture is a thesis you can't build. Both run entirely on the ollwrite editor + oll-memory primitives — nothing here needs a model we don't already route.

A · "Write a book from 15 years of email"

Ingest is a live ledger, not a spinner. Drop the archive → watch it resolve: 48,213 emails → 31,904 after dedup → embedding on-device. Then a passive "What I found" surface turns a scary blob into something browsable — recurring people, a timeline, topics by year. The corpus stops being a wall and becomes a map.

As you write, the left pane silently repopulates (~600ms debounce, keyed on the current paragraph) with memory cards — provenance + snippet + date — draggable straight into prose as citations. The right pane interrogates: "what did I say about the Zurich deal that year?" → cited answers that jump to the exact chunk. You are writing from your life, with your life doing the remembering.

B · "Draft a report from my research folder + bullet points"

Ingest the folder and learn the report skeleton — structured-extract the section order, headings, typical length, and voice from the user's own past reports. Type bullets → hit "Draft from bullets" → each bullet expands into prose in the learned house structure + voice, pulling cited evidence, delivered as accept / reject diffs. The blank page never happens.

Gap detection is the differentiator. Any claim it can't ground gets an inline amber ⚠ churn figure — no source found. Chips aggregate into a right-pane "Before this is done, I need:" checklist; resolving inline or in chat resolves both. Plus "Expand with sources" on any thin paragraph. The tool that tells you what it doesn't know is the tool you trust with billable work.

3 · The four "can't-write-without-this" moments

If one of these lands in the first session, the tool becomes a habit. These are the emotional hooks the whole design serves.

1 · It remembered for you

The forgotten memory surfaces itself. You didn't search — it knew. The single most defensible feeling: no reactive tool can produce it.

2 · Bullets → a cited draft

In your voice, grounded in your evidence. The blank page — the writer's most expensive minute — simply never happens.

3 · It admits what it doesn't know

Gap-chips. The anti-hallucination, trust-earning moment — the opposite of a confident wrong answer, which is what kills reliance on AI writing.

4 · One click to the source

Every claim is one click from where it came from. Groundedness becomes reliance; reliance becomes a paid habit.

4 · The competitor map — the intersection nobody owns

Scored on the four axes that define the empty cell. The point isn't that competitors are bad — several are excellent — it's that none holds all four corners, and the corner we can hold is proactive-corpus × real-editor × on-device.
ProductCorpus RAGProactive
surface-while-writing
Real long-form editorOn-deviceNon-Enterprise EUPricing
NotebookLM
best RAG, but siloed notebooks · ~13% hallucination on failure
YesNoNo editorNoNoFree / $4.99 / $19.99 / $99.99+
Notion AI / Q&A
shallow at scale · lags >3–5k words
Yes (shallow)No (pull)WeakNoEnterprise-onlyFull AI needs $20 Business + $10/1k credits
Saga
conceptually closest — but cloud-only
Yes (mention-triggered)PartialGood / lightNoNo$6 / $12 (cheapest)
Craft
editor + on-device — but no retrieval · export lock-in is gripe #1
No corpus RAGNoBest editorYes (on-device models)CHF-pricedFree / ~CHF-tier subs
Sudowrite
best fiction model (Muse) · Story Bible is manual, not your corpus
Manual Story BibleNoStrongNoNo$10 / $22 / $44 · credits
Lex / TypeNone / manual-attachNoStrongNoNo~$8–16/mo
Scrivener
the pay-once precedent
None (no AI)NoStrongYes (local)$59.99 one-time
Obsidian + Smart Connections + Copilot (Ollama)
the real feature-competitor — a 2-plugin science project
YesYes (SC = only proactive surfacer)Markdown-hackerYes (local)Yes (local)Copilot $14.99/mo or $349.99 one-time
Local-LLM tools (GPT4All / AnythingLLM / Msty)Reactive RAGNoNo editorYesYesFree / Msty $349 lifetime
The finding: the only stack hitting editor + proactive-corpus + on-device is Obsidian + plugins — a hobbyist rig (CORS wrangling, model-picker, RAM tuning, index corruption), not a product. Unify it into one beautiful app and polish it, and the cell is unclaimed.

5 · Threat ranking — and how we beat each

Ranked by how directly each threatens the cell we want. Beating each is a positioning choice, not a feature race.

1 NotebookLM — free + Google's gravity

Beat with a real editor, cross-corpus context, and on-device / EU. One line: "NotebookLM you can actually write in, over ALL your files, that never leaves your machine." Its notebooks are silos; ours is your whole knowledge. Its answers fabricate on failure; ours admit gaps.

2 Obsidian + plugins — the real feature-competitor

Beat with zero-config polish. Collapse the science project into one app: local by default, a clean prose editor instead of a markdown pane, no CORS / model-picker / RAM ritual. Same capability, none of the assembly. This is the head-to-head match, and product craft is the whole edge.

3 Notion AI — incumbent surface area

Beat with proactive not pull, a long-form editor that doesn't lag past 3–5k words, and non-Enterprise EU / on-device. Notion makes you ask; we surface. Notion gates residency behind Enterprise; we make it the default.

4 Craft — best editor, no retrieval

Beat by adding the retrieval layer it lacks entirely + fixing export lock-in with open formats. Craft proves on-device editing sells; it just never connected to your corpus. We start where it stops.

Cross-cutting levers (each threat is beaten by some subset): proactive retrieval (only Smart Connections does it today) · on-device / EU by default · whole-corpus grounding · honest pay-once / flat pricing.

6 · Platform primitives to add — the P1–P7 ladder

The product decomposes into reusable primitives on Core + Model. Each is small, each is teachable, and each — once built — powers every future writing surface, not just this one. That's the compounding thesis from Beyond the Wrapper, made concrete.
PrimitiveWhat it isFeatures it unlocksEffortReuse
P1 · Context-assembly / proactive retrievalEmbed the current paragraph → hybrid-retrieve + rerank → the side railThe whole proactive-surface magic; the left paneSEvery writing surface
P2 · Folder / bulk on-device ingestionDrop a folder → parse PDF / docx / md via LlamaIndex readers → embed locallyThe corpus itself; the ingest ledgerS–MFoundational — everything
P3 · Incremental / watch syncHash-diff, re-embed only deltasAlways-fresh corpus without re-ingestMRetention layer
P4 · Style / voice profileExtract a style rulebook from the corpus → inject as system rules"Draft in your voice"; the ghostwriter wedgeMThe differentiator
P5 · Pattern / structure extractionThe recurring skeleton across past docs → a reusable template"Draft from bullets" in house format; the free SEO toolMHighest B2B value + top-of-funnel
P6 · Citation / provenance trackingEvery span carries its source chunkClick-to-source; gap-chips; the trust layerS–MRequired for legal / academic
P7 · Per-user private corpus + ACLNamespace by oll-core JWTTeams, shared corpora, seat pricingMUnlocks B2B / teams
Build order by leverage: P2 → P1 → P6 → P4 → P5 → P7 → P3. P2 + P1 + P6 is already a shippable product (ingest → proactive surface → cited). P4 / P5 are the premium wedge. P7 unlocks teams. P3 is retention. And the compounding line: one primitive stack → many pay-once products, each new vertical a weekend not a rebuild — the Core-client architecture doing exactly what it was designed to.

7 · Who pays — and pricing's undercut lane

The buyer profile is specific: someone who writes from a body of prior work and whose time is billable. That's where proactive-corpus retrieval turns straight into money.
SegmentPainROIWillingness to pay
Consultants / analysts ← beachheadRe-reading old decks / notes to write the next oneBillable $190–510/hr, save 10+ hrs/wk$30–80/mo or $99–199 pay-once
Grant writersRecomposing boilerplate across proposalsFaster submissions, more grants outComps $50–150/mo (Grantable) → undercut $99–149 pay-once
LawyersProvenance + privilege; cloud AI is a noGroundedness for citeable work; on-device for privilege$30–60/mo private tier — white space vs Harvey / Spellbook $100–350/seat
Authors / ghostwritersBlank page; matching a client's voice$15–50k/book; voice-profile is the wedgeSudowrite's turf — prosumer pay-once
Researchers / academicsSynthesizing a reading pile into proseCited drafts from their own libraryPrice-sensitive: $9–29
AgenciesHouse voice + house format at scaleP4 + P5 + P7 = consistent output across a team$20–40/seat

The honest pay-once point

The entire AI-writing category is subscription or credits; pay-once is nearly unclaimed — only Scrivener ($60) and Msty / Copilot ($349 lifetime) hold it. Our points: personal pay-once 29–49 · pro / consultant 99–149 · team 39/seat/mo.

And here's why it's defensible, not a gimmick: 2026 lifetime-deals now cap credits because cloud inference costs money on every single run — so "lifetime" quietly means "lifetime of a small credit bucket." On-device / self-hosted-Ollama is the one architecture where "pay once, run forever on your hardware" is genuinely true (~$0/run), not bait. That is an ethical and a competitive differentiator, and it is only ours because the Model gateway already routes to on-device.

8 · Exploitable gripes

Where the incumbents bleed. Each gripe is a positioning line we can honestly claim.

Credit / usage anxiety

Sudowrite, Cursor's June-2025 20× backlash → public apology + refunds, Coda, Notion's $10/1k credits. → Flat / pay-once wins the trust.

Shallow retrieval

NotebookLM fabricates + silos; Notion misses at scale. → Precision-first, cited, whole-corpus.

Laggy editors

Notion degrades past 3–5k words. → A real long-form surface.

Lock-in / export

Craft export called "a disaster." → Open formats, your data stays yours.

Privacy opacity

EU residency paywalled behind Enterprise. → On-device / EU by default.

Plugin jank

Obsidian's proactive rig is a config marathon. → Zero-config polish is the moat.

9 · Co-optimized distribution

Distribution is the documented weakness (strong engineer, thin distribution) — so it's baked into the build, not bolted on. The trick: one primitive is simultaneously a premium feature and the top of funnel.

P5 → a free public "Templatize your report" tool. Drop 3 past reports → get your reusable house-format skeleton, no login (via the guest-checkout pattern). That's an SEO front door → an anonymized template / pattern library as evergreen long-tail SEO → profession communities (r/consulting, analyst Slacks, grant-writer forums, r/writing) via the honest draft-only reply pipeline → and crucially, on-device / private as the headline, which is the one thing that lets us post in confidential-work communities (legal, corporate strategy) where "cloud AI" is an instant no.

The co-optimization win: P5 (pattern-extraction) is a paid premium feature and the free SEO tool and the reason we can enter privacy-sensitive communities — one primitive, three distribution jobs. Engineering IS the distribution.

10 · Sequenced roadmap + recommendation

Overnight cadence — each stage ships something chargeable on the existing Stripe rails. Propose-only: this is the shape, not a committed plan.
  1. Night 1–2 — First paying product: P2 + P1 + P6 into ollwrite → "private writing memory," pay-once 29–49 on existing Stripe rails. Ingest → proactive surface → cited. Shippable on day two.
  2. Night 3–4 — P4 voice profile → the 99 Pro tier; ghostwriters + authors. The differentiator lands.
  3. Night 5–6 — P5 pattern-extraction shipped as a premium feature AND the free "templatize" tool → first B2B + SEO on.
  4. Week 2 — P7 multi-tenant ACL → teams 39/seat; legal + agencies.
  5. Ongoing — P3 watch-sync → retention; the corpus stays fresh without re-ingest.

Recommendation (propose-only)

The empty cell is real and we own both halves — this is the highest-conviction net-new product in the portfolio, and the roadmap ships a chargeable slice by night two. P2 + P1 + P6 into ollwrite is the move, with P4 as the premium wedge right behind it.

⚠ Ship-vs-build flag (the honest part)

This is a net-new build while the humaniz.me first-franc path sits at the LIVE-Stripe finish line. The guardrail says: close the first stranger dollar first — or run this strictly as a parallel overnight track — not jump off a finished asset to start a new one. This doc is the roadmap for after / alongside the first franc, not instead of it. It feeds Strategy + Backlog; they decide, this informs.

Competitors notebooklm.google/plans · felloai (NotebookLM analysis) · sudowrite.com/pricing · lex.page/pricing · type.ai/pricing · notion.so (pricing / data-residency) · craft.do/pricing · saga.so · literatureandlatte.com (Scrivener) · smartconnections.app · obsidiancopilot.com · msty.ai · cursor.com (pricing / June-2025 backlash)
Business / buyers winder.ai (consulting AI ROI) · grantable.co · spineauthors + reedsy (ghostwriting economics) · bindlegal + hyperstart (legal AI) · saastools.blog + affinityally (lifetime-deal economics) · markaicode (on-device RAG)

Follows Beyond the Wrapper · executes on the Platform Extensions build ledger. Propose-only · 2026-07-05 · reference links are honest and directional; public build-in-public figures are not audited.