From 1b958f48c85afea2ce4d5da1d8867aea46ec2412 Mon Sep 17 00:00:00 2001 From: Dhravya Shah Date: Tue, 29 Sep 2026 17:40:50 -0700 Subject: [PATCH] docs: explain advanced PDF extraction availability (#1728) --- apps/docs/concepts/content-types.mdx | 10 ++++++++++ apps/docs/ingestion/add-memories.mdx | 4 ++++ apps/docs/overview/billing.mdx | 4 ++++ 3 files changed, 18 insertions(+) diff --git a/apps/docs/concepts/content-types.mdx b/apps/docs/concepts/content-types.mdx index 164da1ff..a9a4a342 100644 --- a/apps/docs/concepts/content-types.mdx +++ b/apps/docs/concepts/content-types.mdx @@ -55,6 +55,16 @@ await client.documents.uploadFile({ **Extracts:** Text, tables, headers. OCR for scanned documents. +#### Advanced document extraction + +**Scale and Enterprise** include advanced extraction for PDFs: page-by-page OCR with descriptions of figures and diagrams merged into the extracted text. This helps make the visual content in reports, research papers, and technical documents searchable alongside their text. + +**Free, Pro, and Max** continue to support standard PDF extraction, including OCR for scanned documents. The advanced extraction tier applies to PDFs; it does not change support for other file types. + +Use the same `uploadFile` call shown above. Supermemory selects the extraction tier automatically from your organization's plan; no additional request parameter is needed. If advanced extraction cannot process a PDF, Supermemory falls back to standard extraction, so figure and diagram descriptions are not guaranteed for every document. + +See [Billing & usage](/overview/billing#feature-availability) for plan availability. + ### Microsoft Office Word, Excel, and PowerPoint files upload the same way — Supermemory detects the type from the file itself: diff --git a/apps/docs/ingestion/add-memories.mdx b/apps/docs/ingestion/add-memories.mdx index bf6489ff..fc0dce32 100644 --- a/apps/docs/ingestion/add-memories.mdx +++ b/apps/docs/ingestion/add-memories.mdx @@ -189,6 +189,10 @@ Upload PDFs, images, and documents directly. **Limits:** 50MB max file size + +Scale and Enterprise include [advanced document extraction for PDFs](/concepts/content-types#advanced-document-extraction), with page-by-page OCR and descriptions of figures and diagrams. It is selected automatically from your organization's plan; no extra upload parameter is needed. Free, Pro, and Max keep standard PDF extraction, including OCR for scans. + + --- ## Parameters diff --git a/apps/docs/overview/billing.mdx b/apps/docs/overview/billing.mdx index dda9d86f..936d90a9 100644 --- a/apps/docs/overview/billing.mdx +++ b/apps/docs/overview/billing.mdx @@ -183,6 +183,8 @@ Usage meters are shared; **feature gates** differ by tier. | Feature | Free | Pro | Max | Scale | Enterprise | |---|---|---|---|---|---| | Memory API, search, profiles | Yes | Yes | Yes | Yes | Yes | +| Standard PDF extraction, including OCR for scans | Yes | Yes | Yes | Yes | Yes | +| [Advanced document extraction for PDFs](/concepts/content-types#advanced-document-extraction) (figure and diagram descriptions) | — | — | — | Yes | Yes | | Diff / delta token billing | Yes | Yes | Yes | Yes | Yes | | Coding plugins (Claude Code, Codex, Cursor, OpenCode, OpenClaw, Hermes) | Yes | Yes | Yes | Yes | Yes | | Team management | — | Yes | Yes | Yes | Yes | @@ -197,6 +199,8 @@ Usage meters are shared; **feature gates** differ by tier. Coding plugins are available on every plan, including Free. Their API usage still consumes your plan's credits at the applicable rates. +PDF uploads use the extraction tier available to your organization automatically; no extra API parameter is required. On plans without advanced extraction, uploads still use standard extraction rather than returning a feature-gate error. Advanced extraction can also fall back to standard extraction if needed. This tier distinction applies to PDFs, not all supported file types. + Overrides can be applied per org in metadata (`featureOverrides`) for enterprise deals. ### HTTP behavior when gated