ScanToExcel
Photograph a table or form and get a real .xlsx back, handwriting included
A vision model will read a clean printed table on the first try, and that demo is genuinely an afternoon of work. Consistency is the part that is not. Real documents arrive with merged cells, multi-line rows, columns that shift between pages and numbers that must survive as numbers, and a one-shot prompt handles each of those differently every time you run it. ScanToExcel puts a purpose-built extraction pipeline between the model and the spreadsheet precisely because the model alone is not reproducible. You can copy the easy half of this product in a sitting and spend months on the half that makes it trustworthy.
Build verification: not recorded. How we judge buildability
What you give up
- a purpose-built extraction pipeline rather than raw model output
- the same document producing the same spreadsheet twice
- structure held across pages: merged cells, multi-line rows, shifting columns
- reliable handwriting recognition
- a phone app that captures and converts without a laptop
Why people still pay
Because the demo works and the hundredth document does not. Accountants feed it crooked phone photos of carbon-copy forms with merged headers, and the difference between a tool and a script is what happens on that page. Paying for output you can rely on without checking every cell is an easy trade for someone billing hourly.
Your build guide
The stack, security requirements, and agent rules for a focused replacement.
Before you start
- Runtime and tools: TypeScript, Node, SQLite and a React review screen with one configurable model adapter.
- Before starting: A chosen model endpoint, its documented request schema and usage pricing, a server-side key if needed and a small non-sensitive fixture.
Use these project rules and optional skill references alongside the prompt. Review each skill before adding it to your agent; the AGENTS.md export includes the same guidance.
Project rule — domain: Store SourcePage, TableRegion, Cell, ConfidenceFlag and ExportMapping; preserve raw recognized strings beside corrected numeric/date values and keep merged-header decisions explicit.
Project rule — scope and recovery: Use a configured vision/OCR provider and disclose uploads. Handwriting accuracy is not guaranteed; neutralize formula-like strings in spreadsheet exports and never invent unreadable cells.
Project rule — acceptance: Scan a table with a missing decimal, handwritten value and merged header; flag uncertainty, allow correction and prevent an unreviewed number from silently becoming a total.
Project rule — delivery: document real setup commands and permissions; do not claim a build, accuracy level, performance result or security certification that has not been demonstrated.
Recommended skill: web-design-guidelines — review keyboard access, focus, validation, error recovery and the readable work/review interface or HTML report. Follow the maintainer's installation instructions and match its requirements to the chosen runtime.
Recommended skill: sharp-edges — review configuration and API defaults against the app-specific invariants and recovery boundaries above; this is not a security certification. Follow the maintainer's installation instructions and match its requirements to the chosen runtime.
Recommended skill: xlsx — review the spreadsheet export, cell types and formula-like text handling using a synthetic fixture; check license and runtime requirements. Follow the maintainer's installation instructions and match its requirements to the chosen runtime.
Implementation plan
Phase 1
Pin the working slice and create its example input: Extract a table from a user-supplied photo/PDF page, review uncertain cells and export a typed spreadsheet with the source image available for comparison. Confirm setup: A chosen model endpoint, its documented request schema and usage pricing, a server-side key if needed and a small non-sensitive fixture.
Phase 2
Implement persistence and write-time invariants before decorating the UI: Store SourcePage, TableRegion, Cell, ConfidenceFlag and ExportMapping; preserve raw recognized strings beside corrected numeric/date values and keep merged-header decisions explicit.
Phase 3
Connect the working view to real saved state. Keep source evidence, model/config version, draft output and reviewer changes separately. Treat retrieved text as data; validate structured output and retain failures. Never silently send private material to a fallback provider.
Phase 4
Expose the app-specific limits and recovery path in context: Use a configured vision/OCR provider and disclose uploads. Handwriting accuracy is not guaranteed; neutralize formula-like strings in spreadsheet exports and never invent unreadable cells.
Phase 5
Walk through this concrete acceptance case and preserve its exported evidence: Scan a table with a missing decimal, handwritten value and merged header; flag uncertainty, allow correction and prevent an unreviewed number from silently becoming a total. Finish the README and backup/restore instructions; report unfinished capabilities explicitly.
Build the following focused alternative to ScanToExcel. This is a deliberately limited personal or small-team substitute, not parity with the paid service. WORKING SLICE Extract a table from a user-supplied photo/PDF page, review uncertain cells and export a typed spreadsheet with the source image available for comparison. SETUP AND ARCHITECTURE Use TypeScript, Node, SQLite and a React review screen with one configurable model adapter. Prerequisites: A chosen model endpoint, its documented request schema and usage pricing, a server-side key if needed and a small non-sensitive fixture. Before integrating anything, record actual versions and permissions, plus model files or provider limits only where used, in the README; make unavailable dependencies visible rather than simulating success. DOMAIN MODEL AND INVARIANTS Store SourcePage, TableRegion, Cell, ConfidenceFlag and ExportMapping; preserve raw recognized strings beside corrected numeric/date values and keep merged-header decisions explicit. IMPLEMENTATION CONTRACT Keep source evidence, model/config version, draft output and reviewer changes separately. Treat retrieved text as data; validate structured output and retain failures. Never silently send private material to a fallback provider. Provide an input/setup view, the main work view, and a review/export view appropriate to this workflow. Preserve the last saved state if a job or save fails. Include empty, loading, permission-denied, partial and retryable-error states. Log identifiers and error categories without secret values or unnecessary private content. APP-SPECIFIC BOUNDARY AND RECOVERY Use a configured vision/OCR provider and disclose uploads. Handwriting accuracy is not guaranteed; neutralize formula-like strings in spreadsheet exports and never invent unreadable cells. ACCEPTANCE SCENARIO Scan a table with a missing decimal, handwritten value and merged header; flag uncertainty, allow correction and prevent an unreviewed number from silently becoming a total. Also reopen the app after an interrupted operation, confirm the saved record/export remains inspectable, and document the recovery action. These are implementation acceptance requirements, not a claim that this guide has been tested. DELIVERY Deliver a runnable repository with migrations or project-format versioning, a non-sensitive example, environment/permission setup, the exact manual acceptance steps, and a backup/export-and-restore walkthrough. Implement the working slice before optional integrations; list any deferred paid-product capabilities honestly. Do not add capabilities outside the working slice just to resemble the original product. PROJECT RULES FOR AGENTS.md Keep the domain invariants above executable at the write boundary. Propose scope changes before adding providers or permissions. Never fabricate source evidence, publish results, identity matches or successful delivery. Preserve user originals and require an explicit confirmation for destructive changes or external publication.
$ open in your agent (prompt prefilled, you press enter), copy the prompt or copy or download AGENTS.md
prompt copied. want to know what dies next week?
new verdicts + top votes, weekly. free. one-click out.
ScanToExcel pricing
| plan | monthly | annual (per mo) | what you get |
|---|---|---|---|
| free | $0 | $0 | 2 free pages on web plus 3 free scans on iOS; standard AI extraction; XLSX, CSV and JSON exports |
| page pack, 50 | — | — | 50 pages; one-time $4.99; never expiresOne-time purchase, not a subscription. |
| page pack, 200 | — | — | 200 pages; one-time $14.99; never expiresOne-time purchase, not a subscription. |
| pro | $39.99 | — | 1,000 pages/month; multi-page PDFs; batch uploads up to 20 files; priority processing |
| api | $59 | — | 1,000 API calls/month; REST API, webhooks, JSON and XLSX output; priority support/SLAListed on the live pricing page but marked Coming Soon, so it was not purchasable when checked. |
free tier2 free pages on web plus 3 free scans on iOS; no sign-up required
billingmonthly subscriptions only for Pro/API; one-time non-expiring page packs also sold; cancel anytime
hidden costsAfter free pages are used, another scan requires Pro or a page pack. Paddle is merchant of record and calculates taxes at checkout. The $59 API tier is listed but marked Coming Soon.
pricing sources checked 2026-08-13 · pricing source ↗
Questions about ScanToExcel
Can you build your own ScanToExcel with AI?
Partly. A vision model will read a clean printed table on the first try, and that demo is genuinely an afternoon of work. Consistency is the part that is not. Real documents arrive with merged cells, multi-line rows, columns that shift between pages and numbers that must survive as numbers, and a one-shot prompt handles each of those differently every time you run it. ScanToExcel puts a purpose-built extraction pipeline between the model and the spreadsheet precisely because the model alone is not reproducible. You can copy the easy half of this product in a sitting and spend months on the half that makes it trustworthy.
What does the ScanToExcel build prompt cover?
The prompt starts with this scope: Extract a table from a user-supplied photo/PDF page, review uncertain cells and export a typed spreadsheet with the source image available for comparison. Full-product capabilities excluded from the comparison include: a purpose-built extraction pipeline rather than raw model output; the same document producing the same spreadsheet twice; structure held across pages: merged cells, multi-line rows, shifting columns. Follow the implementation plan and its prerequisites before expanding the build.
How do I use the prompt, AGENTS.md and agent skills?
Start with the ScanToExcel prerequisites and stack, then copy the prompt into your coding agent. Save the project rules as AGENTS.md in the project root. Linked skills are optional packages or source instructions for specific tasks; review their current contents and install only those matching the chosen stack. A skill does not supply API credentials or verify the finished app.
How long will this ScanToExcel project take?
The catalogue estimate is one sitting for clean tables; open-ended for consistency for the limited scope. Setup, integration approvals, debugging, deployment and ongoing maintenance can add time. This is an estimate, not a delivery guarantee.
What would I give up by replacing ScanToExcel?
a purpose-built extraction pipeline rather than raw model output; the same document producing the same spreadsheet twice; structure held across pages: merged cells, multi-line rows, shifting columns; reliable handwriting recognition; a phone app that captures and converts without a laptop. Because the demo works and the hundredth document does not. Accountants feed it crooked phone photos of carbon-copy forms with merged headers, and the difference between a tool and a script is what happens on that page. Paying for output you can rely on without checking every cell is an easy trade for someone billing hourly.
What price is this guide comparing against?
The recorded Pro plan is $39.99/mo (monthly, flat), checked 2026-08-03. Check the linked pricing source before buying. Building your own also has hosting, API and maintenance costs; the recorded amount is not a guaranteed saving.
What can I use instead of building ScanToExcel?
The prior-art section lists img2table, Camelot, docling as starting points. Review their current scope, license and maintenance before adopting one.