docs: explain advanced PDF extraction availability (#1728)

This commit is contained in:
Dhravya Shah 2026-09-29 17:40:50 -07:00 • committed by GitHub
parent 640eeaa8e8
commit 1b958f48c8
No known key found for this signature in database
GPG key ID: B5690EEEBB952194
3 changed files with 18 additions and 0 deletions

View file

@ -55,6 +55,16 @@ await client.documents.uploadFile({
**Extracts:** Text, tables, headers. OCR for scanned documents.
#### Advanced document extraction
**Scale and Enterprise** include advanced extraction for PDFs: page-by-page OCR with descriptions of figures and diagrams merged into the extracted text. This helps make the visual content in reports, research papers, and technical documents searchable alongside their text.
**Free, Pro, and Max** continue to support standard PDF extraction, including OCR for scanned documents. The advanced extraction tier applies to PDFs; it does not change support for other file types.
Use the same `uploadFile` call shown above. Supermemory selects the extraction tier automatically from your organization's plan; no additional request parameter is needed. If advanced extraction cannot process a PDF, Supermemory falls back to standard extraction, so figure and diagram descriptions are not guaranteed for every document.
See [Billing & usage](/overview/billing#feature-availability) for plan availability.
### Microsoft Office
Word, Excel, and PowerPoint files upload the same way — Supermemory detects the type from the file itself:

View file

@ -189,6 +189,10 @@ Upload PDFs, images, and documents directly.
**Limits:** 50MB max file size
<Note>
Scale and Enterprise include [advanced document extraction for PDFs](/concepts/content-types#advanced-document-extraction), with page-by-page OCR and descriptions of figures and diagrams. It is selected automatically from your organization's plan; no extra upload parameter is needed. Free, Pro, and Max keep standard PDF extraction, including OCR for scans.
</Note>
---
## Parameters

View file

@ -183,6 +183,8 @@ Usage meters are shared; **feature gates** differ by tier.
| Feature | Free | Pro | Max | Scale | Enterprise |
|---|---|---|---|---|---|
| Memory API, search, profiles | Yes | Yes | Yes | Yes | Yes |
| Standard PDF extraction, including OCR for scans | Yes | Yes | Yes | Yes | Yes |
| [Advanced document extraction for PDFs](/concepts/content-types#advanced-document-extraction) (figure and diagram descriptions) | — | — | — | Yes | Yes |
| Diff / delta token billing | Yes | Yes | Yes | Yes | Yes |
| Coding plugins (Claude Code, Codex, Cursor, OpenCode, OpenClaw, Hermes) | Yes | Yes | Yes | Yes | Yes |
| Team management | — | Yes | Yes | Yes | Yes |
@ -197,6 +199,8 @@ Usage meters are shared; **feature gates** differ by tier.
Coding plugins are available on every plan, including Free. Their API usage still consumes your plan's credits at the applicable rates.
PDF uploads use the extraction tier available to your organization automatically; no extra API parameter is required. On plans without advanced extraction, uploads still use standard extraction rather than returning a feature-gate error. Advanced extraction can also fall back to standard extraction if needed. This tier distinction applies to PDFs, not all supported file types.
Overrides can be applied per org in metadata (`featureOverrides`) for enterprise deals.
### HTTP behavior when gated