
Uploaded files organized in the document library
Documents (PDF, Word, RTF, TXT, Markdown, HTML)
Documents (PDF, Word, RTF, TXT, Markdown, HTML)
Spreadsheets (Excel, CSV)
Spreadsheets (Excel, CSV)
Presentations (PowerPoint)
Presentations (PowerPoint)
Images (OCR applied automatically)
Images (OCR applied automatically)
Email (EML, MSG, MBOX, PST)
Email (EML, MSG, MBOX, PST)
Attachments within email files are extracted and ingested as separate searchable documents, with a link back to the parent message.
Code and data (source files, JSON, XML)
Code and data (source files, JSON, XML)
Archives (ZIP, TAR, RAR, 7Z)
Archives (ZIP, TAR, RAR, 7Z)
When you upload an archive, the system extracts and ingests each file inside as a separate searchable document. Folder structure is preserved.
OCR for Scanned Documents
For image-based files and scanned PDFs:- OCR runs automatically on upload
- The resulting text is searchable like any other document
- The original image is preserved; click any citation to see the highlighted passage on the page
- Handwritten content is recognized where legible
- Common OCR artifacts (broken hyphens, mis-recognized characters) are cleaned up automatically
Size and Page Limits
Upload limits vary by plan - see your dashboard for current per-plan caps on file size, page count, and batch size. For very large files or batches, contact support to discuss enterprise ingestion options.What Happens After Upload
1
Safety scan
The file is scanned for safety. No untrusted code is executed.
2
Text extraction and OCR
Text is extracted; OCR runs on image-based content.
3
Indexing
The document is indexed for search across your matter.
4
Citation detection
Citations within the document are detected and linked to authorities.
5
Ready for analysis
The document is available for analysis in any tool.
Tips
Related
Document Search
Search across everything you have uploaded.
Email Ingestion
Forward emails to ingest documents and attachments.
Drive Import
Pull documents directly from Google Drive.

