DeepSeek PDF Analysis Benchmark: 10 Live File Tests

We tested DeepSeek Chat with 10 controlled PDF and file-analysis tasks. See the 89/100 product-workflow score, the 84/95 submitted-response score, live screenshots, limitations, and a post-hoc Vision diagnostic.

Can DeepSeek PDF analysis handle searchable documents, scans, tables, charts, long files, amendments, multiple uploads, and encryption safely? We ran 10 controlled tasks in a live browser session. The DeepSeek Chat product workflow scored 89/100: submitted model responses earned 84/95, while the protected-file interface safeguard earned 5/5. For platform background, see our independent DeepSeek guide.

Independent benchmark: This test was conducted by Chat Deep AI and is not affiliated with or endorsed by DeepSeek. All files were synthetic and contained no real personal, financial, customer, or confidential information.

DeepSeek PDF analysis: quick verdict

  • Instant-mode Chat product score: 89/100
  • Submitted model-response score: 84/95
  • Post-hoc mode sensitivity: 98/100 if the chart result is replaced by the later Vision diagnostic; this was not a predeclared routing benchmark
  • Test set: 10 tasks using 12 synthetic files
  • Median response time: 8.2 seconds among submitted primary runs
  • Selected repeat runs: all material fields matched in three repeated cases
  • Strongest tested cases: native text, OCR, tables, long-document retrieval, and cross-file calculations
  • Main limitation: Instant returned no chart values from a raster-only PDF; Vision read the same file correctly
  • Test date: July 24, 2026

Bottom line: Within this small synthetic set, Instant handled extracted text, one scan, tables, amendments, and multi-file reasoning well. Vision solved the one image-only chart that Instant could not read. Page citations were occasionally broader than the evidence required, so important answers still need human verification.

DeepSeek new chat interface in Instant mode with DeepThink and Search turned off before a benchmark submission.
Representative interface capture from August 7, 2026 showing Instant mode with DeepThink and Search off. The frozen July 24 run log, not this later UI capture, records those controls for each scored submission.

How we tested DeepSeek file analysis

We created a controlled answer key before opening DeepSeek. Each document contained known facts, calculations, page locations, or failure conditions. This avoids subjective scoring: a value either matched the source file or it did not.

  • Every primary test started in a fresh chat.
  • Search was disabled to keep answers grounded in the uploaded files.
  • Instant was the primary mode for all 10 scored tasks.
  • The first complete response was scored whenever a model response was produced; the protected-PDF case was an explicit UI-level safe-failure test.
  • Exact values, calculations, amendment precedence, and evidence pages were checked against a predefined key.
  • Response time did not affect accuracy scores.
  • Three tasks were repeated in fresh chats to check consistency.

The file-upload control was available in Instant and Vision during our session, but not in Expert. We therefore did not pretend that Expert had been tested with files. Vision was used only as a diagnostic rerun for the chart that failed in Instant.

A searchable PDF attached to a new DeepSeek Instant chat
A synthetic searchable PDF attached in a fresh Instant chat before the prompt was submitted.

Scoring rules

Eight tasks were worth 10 points each. The 30-page retrieval task was worth 15 points because it required several distant page matches and an amendment calculation. The protected-file task was worth 5 points because the correct product behavior was to block submission without inventing document content. No prompt reached the model in that final case, so we also report the submitted model responses separately as 84/95.

Correct values could still lose a point for unsupported or overly broad page citations. A fabricated material claim would cap a test at half credit. Minor formatting differences, such as a written date instead of ISO format, were accepted when the meaning was identical.

Complete DeepSeek PDF benchmark scorecard

#TestModeScoreTimeFinding
1Searchable invoiceInstant9/105.1sAll values and arithmetic correct; one extra evidence page
2Policy amendmentInstant10/1091.8sCorrect general rule, conditional exception, and supporting pages
3Scanned receiptInstant10/1021.5sCorrect OCR, line items, totals, and sum check
4Two-column reportInstant10/105.0sKept both projects separate and applied the priority rule
5Financial tableInstant10/107.5sAll extracted and calculated values correct
630-page manualInstant15/158.7sRecovered distant facts and applied the later limit
7Image-only chart PDFInstant1/108.7sReturned nulls instead of chart values, but did not fabricate
8Contract amendmentInstant9/107.9sTerms correct; one unsupported evidence page included
9CSV + TXT + PDFInstant10/108.2sCorrectly combined three files for all inventory decisions
10Password-protected PDFInstant5/5Not submitted“No text extracted”; the interface blocked submission
Instant-mode Chat product total89/10084/95 from submitted model responses, plus 5/5 for protected-file interface handling

The 91.8-second policy result was a clear interface-delay outlier. We kept it in the record rather than replacing it with a faster rerun. The median of the submitted primary runs was still 8.2 seconds.

Test 1: Searchable invoice — 9/10

The first PDF was a two-page native-text invoice. DeepSeek had to extract the invoice number, customer, due date, subtotal, discount, tax, and amount due, then recalculate the final amount.

Every requested value was correct. It verified that $3,630.00 minus the $180.00 discount plus $172.50 tax equals $3,622.50. The only issue was evidence_pages: [1, 2]. All scored fields appeared on page 1, so page 2 was unnecessary.

DeepSeek returning correct JSON from a searchable invoice PDF
The searchable-invoice response: correct extraction and arithmetic, with an extra page citation.

Test 2: Policy amendment — 10/10

This three-page policy listed a general application deadline, a $275,000 grant cap, a 12% matching contribution, and a contact on the first two pages. Page 3 introduced a later amendment: qualifying North County organizations received a five-day extension only if they emailed notice before the original deadline.

DeepSeek correctly separated the general deadline from the conditional exception. It returned both dates, the complete eligibility condition, the unchanged cap and contribution percentage, the contact name, and evidence pages 1, 2, and 3. Accuracy was perfect, although this run took an unusual 91.8 seconds.

DeepSeek correctly applying a later policy amendment in a PDF
DeepSeek applied the later amendment without confusing the general deadline and the conditional exception.

Test 3: Scanned receipt OCR — 10/10

The receipt was an image-only scan with no searchable text layer. It contained three products, a subtotal of $36.15, tax of $2.89, and a total of $39.04.

Instant extracted the receipt number, date, every item and price, subtotal, tax, total, and highest-priced item. It also confirmed the sum correctly. This matters because the chart test later showed that “image-only PDF” is not one uniform capability: simple document OCR and visual chart interpretation behaved very differently.

DeepSeek extracting line items and totals from an image-only scanned receipt
The scanned-receipt OCR result was correct on every scored field.

Test 4: Two-column report — 10/10

Two project summaries appeared side by side. Each had a different owner, launch date, budget, and risk. A final instruction said the analyst should prioritize the project whose decision gate occurred 14 days earlier.

DeepSeek kept the Atlas and Beacon fields separate, returned all eight project values correctly, and selected Atlas for the stated 14-day reason. It did not merge lines across columns.

DeepSeek keeping two project columns separate in a PDF report
The two-column response preserved the correct owner, date, budget, and risk for each project.

Test 5: Financial table — 10/10

The table contained quarterly gross revenue, returns, net revenue, and units. DeepSeek had to total each measure, identify the strongest quarter, calculate the difference between Q4 and Q1, and verify the annual arithmetic.

All values were correct: $535,000 gross, $23,000 returns, $512,000 net, 36,750 units, Q4 as the highest-net quarter at $143,500, and a $28,500 Q4–Q1 difference. The first send button remained unavailable while the file processed, but the accepted submission completed in 7.5 seconds.

DeepSeek calculating annual figures from a financial table in a PDF
DeepSeek extracted the financial table and returned every calculated field correctly.

Test 6: 30-page document retrieval — 15/15

We placed one operational code and quantity on page 3, a second code and return quantity on page 17, and a budget amendment on page 29. Similar-looking text on other pages acted as distractors.

DeepSeek found ORBIT-314 and 127 units on page 3, LANTERN-852 and 43 returned units on page 17, calculated 84 net units, and used the page-29 amendment to replace the $65,000 limit with $68,750. It returned every requested page number correctly.

DeepSeek retrieving distant facts and an amendment from a 30-page PDF
The long-document answer correctly linked pages 3, 17, and 29.

Test 7: Image-only chart — 1/10 in Instant, 10/10 in Vision

This was the decisive mode test. The PDF contained a raster bar chart with North 84, South 61, East 73, and West 92. The correct derived answers were West as the highest region, a 31-point West–South difference, and a total of 310.

Instant returned null for every value. We awarded one point because it did not fabricate numbers. We then opened a fresh chat, selected Vision, uploaded the same PDF, and used the same prompt. Vision returned all four values, both calculations, and the page reference correctly in 8.9 seconds.

DeepSeek Instant returning null values for an image-only chart PDF
Instant: no chart values, but no invented answer.
DeepSeek Vision correctly reading an image-only chart PDF
Vision: all values and calculations correct.

Practical lesson: Do not treat a successful file attachment as proof that the selected mode can interpret every visual element. For similar chart- or image-dependent PDFs, test Vision on representative files before adopting a workflow.

Test 8: Contract amendment — 9/10

The agreement originally allowed cancellation with 30 days’ notice and payment within 15 days. A later amendment changed those terms to 45 days and 10 days, took effect on September 1, 2026, and preserved a formal-notice email.

DeepSeek returned every original and current value correctly and applied the later amendment. It lost one point because its evidence list included page 2 even though the scored terms were supported on pages 1, 3, and 4.

DeepSeek applying the controlling terms from a contract amendment
The contract values were correct, but the evidence list included one unnecessary page.

Test 9: Multi-file reasoning — 10/10

We uploaded three related files together: a CSV with inventory, a TXT file containing the reorder formula, and a PDF listing incoming shipments. DeepSeek had to calculate post-shipment availability and reorder quantities for three SKUs.

It correctly combined all three sources, handled a zero incoming quantity, and returned the correct result for every SKU: A-17 required 3 units, B-04 required none, and C-91 required 5 units.

DeepSeek T9 chat showing attached CSV, TXT, and PDF files above the complete multi-file JSON prompt.
T9 setup in the original chat: inventory CSV, reorder-policy TXT, and incoming-shipment PDF attached above the exact prompt. Search is off in this August 7, 2026 recapture.
First overlapping DeepSeek T9 result capture showing the opening array, complete A-17 and B-04 objects, and the start of C-91.
T9 result, part 1 of 2: the opening bracket plus complete A-17 and B-04 objects; C-91 begins at the bottom for overlap with part 2.
Second overlapping DeepSeek T9 result capture showing complete B-04 and C-91 objects and the closing array bracket.
T9 result, part 2 of 2: B-04 is repeated for overlap, followed by the complete C-91 object and closing bracket.

Test 10: Password-protected PDF — 5/5

This task measured safe failure, not password bypassing. The encrypted PDF contained a hidden verification phrase. A reliable system should avoid claiming that it read protected content.

The interface labeled the file “No text extracted” and disabled submission. When a prompt was present, it displayed “Remove failed files to submit.” Because no request reached the model, no verification phrase or invented summary was produced. That is the correct outcome for this test.

DeepSeek showing no text extracted for a password-protected PDF
The encrypted file produced a “No text extracted” status.
DeepSeek blocking submission while an unreadable protected PDF is attached
DeepSeek required the failed file to be removed before submission.

Were the results consistent?

We repeated three high-risk cases in fresh chats:

Repeated taskModeConsistency result
Scanned receipt OCRInstantEvery material field matched the first run
30-page retrievalInstantAll 10 requested fields and page locations matched
Image-only chartVisionAll chart values, calculations, and page reference matched

The material-field consistency rate was 100% across these three repeats. That is encouraging, but three repeat runs are a stability check, not a guarantee for every document.

Best settings for DeepSeek PDF analysis

  1. Start by testing Instant on searchable documents, tables, and cross-file text tasks. It performed strongly on the text-centered examples in this suite.
  2. Test Vision on chart- or page-image tasks. Our single raster-only chart moved from 1/10 to 10/10 in Vision, but broader visual performance needs a larger sample.
  3. Disable Search for document-grounded work. This reduces the chance that outside information is mixed with the uploaded evidence.
  4. Request structured output. JSON fields made omissions, wrong calculations, and citation errors easy to detect.
  5. Ask for exact pages or source filenames. Then verify those citations. Two otherwise correct answers cited an extra page.
  6. Check that the send control is active. A file may still be processing even after its attachment card appears.
  7. Never assume an attached file was parsed. Confirm that the interface did not show an extraction failure.

How to upload and analyze a file

  1. Open a fresh DeepSeek Chat and select the mode you intend to test.
  2. Disable Search when the answer must come only from the uploaded evidence.
  3. Select the attachment control, choose the file, and wait until processing finishes.
  4. If the send control remains disabled or the file says “No text extracted,” remove the failed file and verify its format, encryption, or text layer.
  5. Request structured fields, calculations, and exact evidence pages, then check the answer against the original file.

Supported file types and upload limits

Our controlled set successfully used PDF, CSV, and TXT files in the web interface. That is an observation from July 24, 2026, not a permanent support guarantee. DeepSeek’s public Help Center gives image-upload failure examples such as unsupported image formats, no extractable text, and exceeding a maximum, but it does not publish a numerical size-and-format matrix. The official update log records a previous optimization to file upload and webpage summarization.

Do not convert DeepSeek V4’s advertised context window into an assumed PDF upload limit. Context capacity, file size, parser behavior, and the amount of extracted text passed to a model are separate constraints. Check the live interface and test a representative document before building a production workflow.

DeepSeek Chat file upload is not the same as the API

The consumer Chat interface used in this benchmark includes a file-processing layer. The documented DeepSeek chat-completions API currently describes message content as text. DeepSeek’s Anthropic-compatible documentation also marks image, document, and container-upload blocks as unsupported.

For an API workflow, plan to extract text, parse tables, or perform OCR before sending content to the model unless DeepSeek publishes a new native file endpoint. Do not advertise browser upload behavior as a documented API feature.

Privacy: what should you upload?

DeepSeek’s privacy policy treats uploaded files, photos, and chat content as user inputs. It advises users not to submit sensitive personal data, explains that inputs may be used to provide and improve services, provides a training-improvement opt-out, and states that data is stored and processed in the People’s Republic of China.

  • Use synthetic or properly anonymized files for testing.
  • Do not upload client contracts, medical records, identity documents, private financial records, or trade secrets without an approved data-governance process.
  • Review the current privacy policy and account settings before use.
  • Verify every high-impact answer against the original document.
  • Remove passwords locally only when you are authorized to access the file; do not share passwords in a chat prompt.

Frequently asked questions

Can DeepSeek analyze PDF files?

Yes, the DeepSeek Chat web interface analyzed searchable, scanned, table-heavy, long, and amended PDFs in our test. Accuracy depended on document type and mode. The Instant-mode Chat product workflow scored 89/100, comprising 84/95 from submitted responses and 5/5 for protected-file interface handling.

Can DeepSeek read scanned PDFs?

It read our image-only scanned receipt correctly in Instant, including every line item and total. That result does not prove equal OCR accuracy on low-resolution, handwritten, rotated, or damaged scans.

Can DeepSeek read charts inside a PDF?

Mode choice mattered in this one chart case. Instant returned null for all values in our raster-only chart PDF, while Vision read every value and calculation correctly. Test Vision with representative visual files before relying on it more broadly.

Can DeepSeek analyze several files together?

Yes in our test. Instant combined a CSV inventory, a TXT policy, and a PDF shipment schedule and returned correct calculations for all three SKUs.

Can DeepSeek open a password-protected PDF?

Not in our test. The interface reported “No text extracted” and blocked submission until the failed file was removed. It did not invent protected content.

What is the maximum DeepSeek file size?

We did not find a public numerical upload-limit matrix in DeepSeek’s official documentation on the test date. Treat any visible interface behavior as time-specific and verify it again before publication or deployment.

Can the DeepSeek API accept a PDF directly?

The current official chat-completions documentation describes text message content, and the Anthropic-compatible guide lists native image and document blocks as unsupported. An API pipeline should extract or OCR the document before submitting text unless the documentation changes.

Final verdict: Is DeepSeek reliable for PDF analysis?

In this small synthetic benchmark, DeepSeek Chat was a strong file-analysis assistant when prompts requested verifiable structure and the document was routed to a suitable mode. Instant performed well on the tested text extraction, calculation, OCR, long-document, amendment, and cross-file cases. Its two point losses outside the chart test came from over-inclusive page citations, not wrong document values.

The raster-only chart is the critical caveat. Instant’s 1/10 result and Vision’s 10/10 result show why a single overall claim such as “DeepSeek reads PDFs” is too broad. The practical workflow is to use Instant for text-centered files, route visual documents to Vision, and verify every important answer against the source.

As a post-hoc sensitivity calculation, replacing the failed Instant chart result with the later Vision diagnostic changes the product total from 89/100 to 98/100. This was not a predeclared or independently rerun routing benchmark, so 89/100 remains the primary result. Neither score is a reason to remove human review from legal, financial, medical, compliance, or deadline-sensitive decisions.

Official sources

Last tested July 24, 2026. Interface behavior, model routing, file support, and policies may change after publication.